Coding Agents
6 pieces on Coding Agents.
News & Analysis

Qwen 3.8 27B Is Capable but Defaults to Extreme Overthinking
Alibaba's Apache 2.0-licensed 27B vision model fits in 17 GB and handles agents, vision, and code — but its xhigh reasoning default is a trap.

Qwen 3.8 27B Is Strong but Overthinks by Default
Alibaba's 17 GB Qwen 3.8 27B excels at vision, tool use, and coding agents — but its xhigh reasoning default burns tokens on trivial prompts.

Z.ai GLM-5.3: Post-Training Gains on a Fixed 743B Base Model
GLM-5.3 reuses GLM-5.2's 743B base unchanged. Terminal-Bench 3.0 jumps from 4.6 to 28.3; CyberGym hits 84.5%, edging closed frontier models.
Codex Comes to the ChatGPT Mobile App in Preview
Codex is moving onto phones, letting developers start, steer, approve, and monitor AI work from the ChatGPT app.
DeepSWE Reshuffles the AI Coding Leaderboard and Puts GPT-5.5 on Top
A tougher coding benchmark shows wider gaps between frontier AI models, with GPT-5.5 leading and verifier quality becoming the real story.
OpenClaw Creator's $1.3M OpenAI Token Bill Shows the New Cost of Agentic Coding
A reported $1.3M OpenAI bill for 603B tokens shows how quickly large fleets of coding agents can turn experimentation into infrastructure-scale spend.