Models & Open Source
Model capability, architecture, benchmarks, price-performance, licensing, provenance, and workload fit.
Model coverage compounds only when it helps readers make a decision. This topic compares frontier and open systems through capability, architecture, benchmarks, price-performance, licensing, provenance, routing, and deployment fit.
Release news earns a place when it provides durable evidence about what a model can do, where it fails, and which workloads justify switching. Vendor and version names remain useful metadata, but the navigation stays organized around evaluation and operator choice rather than a separate archive for every model family.
-
Kimi K2.5: The Benchmaxxing Debate and China's AI Surge
-
GPT‑5.2 Codex: The CLI that still speaks chat
-
Gemini 3 Flash: Pro-Grade Brains at Warp Speed
-
GPT‑5.2: Reliability as a Product
-
DeepSeek V3.2 Speciale vs GPT-5: math supremacy
-
Anthropic Opus 4.5: The Agentic Architect Arrives
-
Gemini 3 Pro: The Agentic Singularity Arrives
-
GPT‑5.1: What’s New and Why It Matters
-
Claude Code vs Codex CLI: GPT‑5 vs Claude 4.5
-
GPT-5 Codex vs Claude Opus 4.1