Categories Alternatives News Submit a Tool Advertise About
W33 Aug 3 – Aug 9, 2026
Issue W33 · Published every Monday

AI Coding Weekly

Signal, not noise. The models, tools, pricing changes and ecosystem shifts that actually matter for developers — curated once a week.

Read this week ↓ ~5 min read · 21 stories
3 Models Released Qwen Max · Muse · V4-Flash
80% Biggest API Price Cut GPT-5.6 Luna led
3 New Tools Launched Muse Code · Sec plugin · PilotDeck
86.7% Top Benchmark Claude Code · TerminalBench

Top Stories

What you need to know from this week

Benchmark Snapshot

Frontend Code Arena · human-preference voting

TerminalBench 2.1 — Terminal Agent Score

Claude Code
86.7%
Muse Code
82.9%
DeepSeek V4-Flash
82.7%
Codex
81.8%

Research & community-reported scores · TerminalBench 2.1

API Price — per 1M tokens (combined in + out)

DeepSeek V4-Flash cheapest open API
$0.42
GPT-5.6 Luna cheapest closed tier
$1.40
Muse Code Meta pay-as-you-go
$5.50
Gemini 3.6 Flash premium tier
$9.00
$0 $10

Research & community-reported prices

Models & Benchmarks

6 stories
Alibaba Aug 3

Qwen3.8-Max: First Open Max-Tier Model

Alibaba's 2.4T-param / 95B-active coder with ~1M context. Weights drop next week on Hugging Face and ModelScope — the strongest open model yet for repo-scale agent workflows.

DeepSeek Aug 3

DeepSeek V4-Flash-0731 Open-Sourced, Wired into Codex

MIT-licensed 284B/13B-active model scoring 82.7% on TerminalBench 2.1, now available through the Responses API and usable inside Codex.

Anthropic Jul 24

Claude Opus 5: Fable-Class Quality at Half Price

Anthropic's new GA Opus tier ($5/$25 per M) lands within 0.5% of Fable 5 on CursorBench at half the cost, with a 5-level effort dial and 1M context.

Meta Aug 5

Muse Spark 1.2: Co-Trained with Muse Code

Meta's coding model was trained in lockstep with the new agent — the key to its jump in terminal-task reliability over Muse Spark 1.1.

Moonshot AI Aug 5

Kimi K3 Stays Competitive as Managed API

The 2.8T open-weight model is quoted around $3/$15 per M on hosted platforms — frontier open weights now buy control and auditability, not a lower price.

OpenAI Jul 30

Sol Fast Mode: 2.5x Throughput at 2x Price

OpenAI's new speed tier for GPT-5.6 Sol, replacing Priority Processing. Effective cost per unit of throughput is ~1.25x better than standard.

Tools & Editors

4 stories
Meta Aug 5

Meta Muse Code: One-Line Install Terminal Agent

Handles planning, editing, and verification end-to-end with multiple parallel agents. Zero data retention — a top ask for enterprise buyers.

Anthropic Aug 4

Claude Security Plugin (Beta): Scanning in the Terminal

A multi-agent vulnerability scanner for injection, auth bypass, and more, using an adversarial voting panel to cut false positives.

Cursor Aug 5

Cursor 3: The Editor Becomes an Agent Platform

Composer keeps keystroke-level costs down by not routing everything to a frontier model; cloud agents run autonomous work in the background.

ShengShu AI Aug 4

PilotDeck: Open-Source Agent Operating System

From Tsinghua and ShengShu AI: workspace isolation, white-box memory, and native MCP support for running multiple coding agents safely.

Pricing & Business

4 stories
OpenAI Jul 30

GPT-5.6 Luna Cut 80%: $0.20 / $1.20 per M

Terra also dropped 20% ($2/$12); Sol unchanged. OpenAI cites kernel and harness optimizations — some done by Sol itself inside Codex.

Meta Aug 5

Muse Code: Pay-as-You-Go at $1.25 / $4.25 per M

Meta undercuts the $20/mo subscriptions of Claude Code and Codex. Contributor tier drops output to $0.20/M — over 10x cheaper.

Anthropic Aug 4

Heads-Up: Claude Sonnet 5 Intro Pricing Ends Aug 31

From Sep 1 it bills at $3/$15 per M — a 50% increase over the intro $2/$10. Batch and cache prices rise too. Budget for the list price.

OpenRouter Aug 6

Watch the Reseller Markup on Terra & Luna

OpenAI's 50% promo shows on OpenRouter, but the same models cost $2.20 via Bedrock and $2.50 via Azure vs $2.00 direct. Check your provider column.

Ecosystem & Community

4 stories
GCC Aug 4

GCC Policy: No Legally-Binding Code from LLMs

China's GCC issued guidance rejecting AI-generated code with legal significance, pushing provenance tooling and "Assisted-by" labeling.

European Commission Aug 2

EU AI Act Transparency Rules Take Effect

AI identity disclosure and deepfake labeling became mandatory Aug 2, with fines up to €15M. Affects AI code review and pair-programming vendors in the EU.

DeepInfra Aug 6

OpenWeights Keeps Beating the Price War

DeepSeek V4-Flash ($0.42/M), MiMo-V2.5 Flash ($0.40/M) and open models on DeepInfra set the commodity floor — 5–50x under closed APIs.

Community Aug 5

Claude Code vs Codex: Teams Keep Both

Community breakdowns note the two terminal agents trade wins by task — a growing pattern of developers running two or three AI tools at once.

Our Take

The editorial view
01

Meta entering with Muse Code at $0.20/M output marks the commodity moment for coding agents: the tooling is now so table-stakes that the battleground is price and data privacy. Claude Code keeps the quality crown (86.7% TerminalBench), but a 5-point gap against a 10x-cheaper entrant is a warning shot for the subscription model.

02

Qwen3.8-Max open-sourcing a 2.4T Max-tier model changes the ceiling for self-hosted coding. After Kimi K3, this is the second open-weight "frontier-class" release in a month — teams that care about data control now have real options, not just trade-offs.

03

Watch the promo trap: Luna's 80% cut and Terra's 50% promo are limited-time, Sonnet 5 jumps 50% on Sep 1, and reseller markups quietly add 10-25%. Budget on list prices, benchmark on total tokens per task — sticker rates no longer predict your bill.

Older issues are loaded on demand to keep this page fast.