Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA
An open agent-model system for post-deployment continual learning, combining a 744B GLM-5.2 base with a Mixture-of-LoRA architecture that composes specialist adapters for chat, agent, coding, and generative UI tasks; a smaller 50B Qwen3.6-based variant supports local deployment.
Why it matters
324 upvotes on HuggingFace Daily Papers; introduces a full co-designed stack (MinT post-training platform, LongStraw long-context RL, MindForge agentic RL) rather than a single isolated technique.
Importance: 3/5
HuggingFace Daily Papers with 324 upvotes (>=100 bump threshold applied).