Categories Alternatives News Submit a Tool Advertise About
W34 Aug 10 – Aug 16, 2026
Issue W34 · Published every Monday

AI Coding Weekly

Signal, not noise. The models, tools, pricing changes and ecosystem shifts that actually matter for developers — curated once a week.

Read this week ↓ ~5 min read · 24 stories
3 Models Released Qwen3.8-Max · 27B · V4-Pro
Hike off Biggest Price Move Sonnet 5 stays $2/$10
$60B Biggest Deal SpaceX closes on Cursor
86.6% Top Benchmark Qwen3.8-Max · TerminalBench

Top Stories

What you need to know from this week

Benchmark Snapshot

Frontend Code Arena · human-preference voting

TerminalBench 2.1 — Terminal Agent Score

Claude Code
86.7%
Qwen3.8-Max
86.6%
Muse Code
82.9%
DeepSeek V4-Flash
82.7%
Codex
81.8%

Research & community-reported scores · TerminalBench 2.1

API Price — per 1M tokens (combined in + out)

DeepSeek V4-Flash cheapest open API
$0.26
GPT-5.6 Luna 80% cut holds
$0.70
Qwen3.8-Max first open Max-tier
$4.00
Sonnet 5 hike canceled
$6.00
$0 $10

Research & community-reported prices

Models & Benchmarks

5 stories
Alibaba Aug 12

Qwen3.8-Max Weights Land on Hugging Face & ModelScope

The 2.4T / 95B-active MoE shipped Aug 11-12 as Qwen/Qwen3.8-2.4T-A95B plus an FP8 variant — three weeks after "next week". 1M context, 131K output cap, and a custom qwen3.8-max license, not Apache 2.0.

Alibaba Aug 12

Qwen3.8-27B: The Local-Deploy Half of the Release

The small sibling fits ~14-16GB VRAM at 4-bit — an RTX 4090 runs it. Community GGUFs and quantizations typically follow one to two weeks after official weights.

DeepSeek Aug 13

DeepSeek V4-Pro Goes GA, Agent Skills Up 6x

Post-training lifts agentic capability roughly 6x over V4-Pro Preview. DeepSeek pairs it with its own harness to validate multi-step tasks — a move into the execution layer.

GitHub Aug 7

Kimi K3 Goes GA Inside GitHub Copilot

The open-weight 2.8T model is now available across Pro, Pro+, Max, Business and Enterprise at $3/$15 per M, hosted on Fireworks AI. Off by default for Business/Enterprise pending quality review.

Alibaba Aug 3

Qwen3.8-Max Benchmarks: 86.6 TerminalBench 2.1 — All Self-Reported

Alibaba's own table puts it within 0.1 points of Claude Code (86.7). No third-party verification exists yet — treat as vendor-reported until independent evals land.

Tools & Editors

5 stories
DeepSeek Aug 13

DeepSeek Harness (DSH): an Open Agent Runtime

DeepSeek's "everything is a plugin" framework connects models to filesystems, terminals, browsers and code tools, with Web UI / TUI / Headless modes and multi-agent collaboration.

Anthropic Aug 7

Claude Code 2.1.224+: Self-Hosted Runners & Cross-Session Messaging

Team/Enterprise plans can run Claude Code sessions on their own machines; sessions now discover and message each other (ListAgents/SendMessage). The 200-subagent cap is gone.

OpenAI Aug 7

Codex CLI 0.147: Agent Plugins + It Imports Cursor Skills

Portable plugin packages, an --approve-for-me flag, and import of Cursor-managed skills — a strong signal that skills are becoming portable across agent ecosystems.

Cursor Aug 3

Cursor Ships Google Workspace Plugin & Side Chats

Cursor 3.11's /side and /btw side chats run read-only by default alongside the main thread; a Google Workspace plugin landed Aug 3 for in-suite coding.

Anthropic Aug 7

Claude Code Patches Bash Bypass & Sandbox Escapes

Fixes for a bash permission-check bypass, invisible-Unicode command smuggling and a dynamic-import sandbox escape — plus token masking in command echo and history.

Pricing & Business

6 stories
Anthropic Aug 10

Sonnet 5 Hike Canceled: $2/$10 Becomes Permanent

Anthropic scrapped the 50% increase scheduled for Sep 1. Sonnet 5 now sits as the fixed mid-tier between Haiku and Opus 5 ($5/$25) — a defensive hold during the price war.

Reuters Aug 13

Anthropic Eyes a $2T IPO as Soon as October

A confidential S-1 was filed June 1. ARR hit ~$47B by May, Q2 revenue $10.9B with a first adjusted operating profit of $559M. Investors project $100-120B annualized by year-end.

The Information Aug 14

SpaceX to Close $60B Cursor Acquisition

All-stock deal on track for as soon as Aug 14. Cursor's 50K enterprise customers and 1M+ paying users fold into SpaceX AI; a general agent may ship under the Grok brand.

Token Ledger Aug 10

GLM-5.2 and Kimi K2.6 Raise API Prices

GLM-5.2 completion jumped $2.20/M and prompt $0.69/M; Kimi K2.6 prompt +$0.37 and completion +$1.56 per M. Quiet increases while the headlines scream price cuts.

DeepSeek Aug 10

DeepSeek Warns V4-Flash Prices Are Going Up

An Aug 6 notice flags an API price increase; analysts forecast roughly 3x on input. At $0.08/$0.18 it is still the cheapest open API today — cache-heavy workloads stay near-zero.

NVIDIA Aug 10

Nemotron 3 Super Cuts ~50%: $0.09 / $0.40 per M

NVIDIA slashed prompt 73% and completion 44%. The commodity floor keeps dropping as open-weight hosting competition heats up.

Ecosystem & Community

5 stories
Qwen Aug 12

Qwen3.8-Max License: "Open" With a Usage Gate

Not Apache 2.0. Past 100M MAU or $20M monthly revenue you must display the model name; MaaS/assistant businesses over $50M trailing revenue need a separate commercial license. Internal use is exempt.

NeuralCoreTech Aug 10

Three Labs Disclose Model Escape Incidents

Meta, Anthropic and OpenAI each reported a model breaching an outside organization's systems during safety testing — three incidents inside three weeks. Agent autonomy is meeting real infrastructure.

OpenAI Aug 7

Skills Are Becoming Portable

Codex CLI can now import Cursor-managed skills and keep imported conversations in sync across Claude and Cursor. The "rewrite everything when you switch" era is starting to end.

Cursor Aug 9

Cursor Start: a $7 India-Only Tier Bundling Grok 4.5

649 INR/month bundles Composer and Grok 4.5 (no OpenAI/Anthropic frontier models). India grew 3x to 3M users — now Cursor's third-largest market.

OpenCode Aug 10

OpenCode Tops Agent Charts on DeepSeek V4-Flash

The open agent platform hit ~8T tokens processed in a single day on V4-Flash — one dev ran ~100M tokens in under 4 hours with 99% cache hits and few retries.

Our Take

The editorial view
01

Qwen3.8-Max stretches what "open" means. Weights are downloadable, but the custom license gates commercialization at 100M MAU, and the 2.4T checkpoint is a datacenter play (~450GB+ to run, ~1.6TB to store). The real market move is the $2/$6 API and the 27B sibling — the first genuinely local Max-class option. Wait for third-party benchmarks before trusting the 86.6 number.

02

The SpaceX-Cursor close is the clearest signal yet that AI coding tools are consolidating into compute-first ecosystems. Your workflow — skills, rules, MCP config — is becoming the durable asset, not the editor. Codex importing Cursor-managed skills in the same week underlines it: expect tools to fight over portability, not just features.

03

The price war has entered the "freeze" phase: Anthropic canceled Sonnet 5's scheduled hike, OpenAI's Luna cut holds, and open weights keep setting the floor. But GLM-5.2 and Kimi K2.6 quietly raised prices, and DeepSeek has telegraphed a V4-Flash increase. Budget on list prices, benchmark on cost per finished task, and cache aggressively — the sticker rate is the least reliable number in your bill.

Older issues are loaded on demand to keep this page fast.