7 articles about coding
OpenAI's GPT-6 Astra capabilities announcement: state-of-the-art on computer use, coding, and science with saturated benchmarks, two real prime-gap proofs, and 0% rate of escaping authorized scope vs Sol's 48%.
Anthropic releases Claude Fable 5.1 and Mythos 5.1, its most advanced models for coding and knowledge work, with research capabilities offering an early glimpse of AI contributing to scientific progress.
Alibaba's Qwen team releases an open-weight coding model achieving GPT-5.6 level performance on a single consumer GPU with 24GB VRAM.
Thinking Machines Lab released Inkling-Small, a 276B-parameter open-weight model with 12B active parameters that outperforms its 975B sibling on agentic coding. The model achieves 80.2% on SWE-Bench Verified and runs on consumer hardware.
Anthropic launched Claude Opus 5, scoring 43.3% on Frontier-Bench v0.1 — surpassing all competitors. Priced at $5/M input and $25/M output tokens, it delivers near-Fable 5 intelligence at half the cost per task, making it the new default on Claude Max.
Grok 4.5 targets coding, agentic tasks, and knowledge work at $2 per million input and $6 per million output tokens — roughly half the price of leading rivals.
Anthropic released Claude Sonnet 5 on June 30, 2026, making it the default model for Free and Pro plans. Priced at $2 per million input tokens, it is the first Sonnet-class model to outscore the Opus 4.8 flagship.