agentic engineering
Octobench Finds a 5-Task Harness Swing
The same GLM model solved 24 coding tasks in one harness and 19 in another. Builders should benchmark the whole system.
agentic engineering
The same GLM model solved 24 coding tasks in one harness and 19 in another. Builders should benchmark the whole system.
policy geopolitics
Europe added 38 AI enforcers to an office reported above 140 people. Model vendors should rehearse the evidence request now.
enterprise ai work
Apollo finds a 6.7-point wage-growth gap in AI-exposed jobs without significant job loss. Employers should audit careers, not only cuts.
Quoted
“A coding model never arrives alone; retrieval, memory, stopping rules, and tests walk into production with it.”
“I’m Stephen, a software engineer with over two decades of industry experience. The Weighted Average is a daily reading of the AI economy — the models, the money, and what they do to us.”