In the news
GLM-4.7-Flash-Coder
White Circle · Published · 3 min read
In 30 seconds
- What happened
- GLM-4.7-Flash-Coder, a 30B mixture-of-experts model, improved agentic coding performance from 33.15% to 41.65% on SWE-rebench-V2 with 14% lower inference cost.
- Why it matters
- Engineers building code-generation agents or evaluating smaller models against larger competitors should assess this model's cost-performance tradeoff for their specific workloads.
- Watch out
- The model was trained on successful trajectories from a teacher model using rejection sampling; performance on novel or out-of-distribution tasks remains unclear.
The patterns behind this
Each one covers how the technique works, when it earns its cost, and where it breaks.
The Agent Architect
One pattern, one tradeoff, one production failure story. A short weekly briefing for people building agentic systems.
Weekly email, one-click unsubscribe. We only use your address to send the briefing.