Claude Code Browser, Codex 7M Users, Hunyuan 3 MTP
Dienstag, 14. Juli 2026 - AI News · (letzte 24h)
Claude Code ships an in-app sandboxed browser on desktop while Codex passes 7M users, overtaking Claude Code by public numbers.
Must read
- Claude Code desktop gains in-app sandboxed browser — Your overnight-agent-factory can now read docs, click through designs, and drive web UIs without external MCP wiring.
- Codex usage up 10x in 6 months to 7M users — Codex may have overtaken Claude Code on raw usage — worth pressure-testing your Claude-first bet before AIE World’s Fair.
- Claude Code v2.1.208 adds process wrapper and vim remaps — CLAUDE_CODE_PROCESS_WRAPPER lets you route every self-spawn through a corporate launcher — useful for RegTech sandboxing.
- Open-weight models hit 29% of Vercel AI Gateway volume — Concrete production-mix data supports your local-plus-cloud routing thesis; token prices flattening changes the LiteLLM calculus.
- Solo builder runs 24/7 local AI on five machines — Direct parallel to your overnight-agent-factory writing — Claude Code build-and-review loop on local hardware.
Tools & Frameworks
Claude Code v2.1.208
New CLAUDE_CODE_PROCESS_WRAPPER forces every self-spawn through a required wrapper executable; adds screen-reader mode and vim insert-mode remaps.
Why this matters: Wrapper hook is the sandboxing primitive RegTech shops need.
Claude Code in-app browser
Sandboxed, configurable browser inside Claude Code desktop can read, click, and interact with sites the way it does with local dev servers.
Why this matters: Removes a whole class of Playwright-MCP hand-rolling.
Cursor reportedly building a general-purpose agent
Cursor is building an agent that answers emails, organises spreadsheets, and handles engineering tasks beyond the IDE.
Why this matters: Cursor pushing outside the editor changes your model-selection story.
Vercel Agent Runs surfaces subagent activity
Eve projects now expose a Subagents tab with per-turn prompt, duration, failures, tool calls, cost and token usage on a shared timeline.
Why this matters: Observability template for your own subagent orchestration.
Vercel Flags targeting rules via CLI
New vercel flags rules command lets agents add, reorder and inspect flag targeting rules from the terminal using the same model as the dashboard.
Why this matters: Feature-flag CRUD your Claude Code agents can now drive directly.
Cline CLI v3.0.40
Adds manual API-key escape hatch for OAuth providers, fixes Bun global-install auto-update, and preserves session id across continuations.
Why this matters: Alternative headless agent worth watching alongside Claude Code.
Open Models & Local
llama.cpp adds Hunyuan 3 with MTP speculative decoding
b9993 lands Tencent Hunyuan 3 (hy_v3) MoE support with sigmoid router, shared expert, dense leading blocks and MTP speculative decoding.
Why this matters: Another frontier-class MoE runnable on your Apple Silicon rig.
llama.cpp b9994 adds Metal Q2_0
Metal backend gains Q2_0 quantisation support, expanding low-bit options for Apple Silicon inference.
Why this matters: More headroom for large coding models on M-series.
llama.cpp adds Minimax2 EAGLE3 speculative decoding
b9990 lands EAGLE3 speculative decoding for Minimax2, improving token throughput on the local runtime.
Why this matters: Speculative decoding is the practical latency lever for local coding agents.
Industry & Trends
Codex hits 7M users, up 10x in six months
Latent Space fact-checks OpenAI’s numbers: Codex added ~1M users in a day and appears to have overtaken Claude Code, which has gone quiet on public metrics.
Why this matters: Real distribution shift among coding agents — not just vibes.
Open-weight models reach 29% of gateway volume
Vercel’s July index shows open-weight models at 29% of tokens routed and price-per-token flattening across labs after a year of steep declines.
Why this matters: Hard data for your local-plus-cloud routing case.
Anthropic extends Claude Fable 5 and Code limits to July 19
Anthropic bumped Claude Fable 5 access and elevated Claude Code limits again through July 19; OpenAI removed GPT-5.6 Sol usage caps in parallel.
Why this matters: Capacity signals matter when you’re planning agent-factory throughput.
Tencent in talks for Manus at $2B after Meta deal unwound
Chinese regulators blocked Meta’s Manus acquisition; Tencent is negotiating to take the largest stake at the same $2B valuation to bolster its agent stack.
Why this matters: Watch, don’t act — but agentic M&A landscape is consolidating.
Apple sues OpenAI over alleged trade-secret theft
Apple accuses OpenAI and former Apple execs of directing employees to bypass security procedures and hand over confidential hardware information.
Why this matters: Background noise, but relevant if you’re betting on Apple-side on-device AI.
Simon Willison shows Opus 4.5-era code-frequency spike
GitHub code-frequency chart on Datasette illustrates a measurable output jump correlated with Opus 4.5-class coding agents in Willison’s own workflow.
Why this matters: Rare honest primary-source datapoint on agentic leverage.
Org & Leadership
OpenAI consolidates power under Brockman ahead of IPO
Greg Brockman takes over OpenAI’s key projects after Fidji Simo’s departure, tasked with revenue growth as ChatGPT share slips and IPO filings loom.
Why this matters: Leadership reshuffle at your primary non-Anthropic supplier.
Sources unavailable today: r/ChatGPTCoding top, r/ClaudeAI top, r/LocalLLaMA top, r/MachineLearning top
Auto-curated daily by Claude Opus 4.7 from Don’t Worry About the Vase (Zvi), Exponential View (Azeem Azhar), GitHub: anthropics/claude-code, GitHub: cline/cline, GitHub: ggml-org/llama.cpp, Google DeepMind blog, Latent Space, Lenny’s Newsletter, NVIDIA developer blog, SaaStr (Jason Lemkin), Simon Willison, TLDR AI, The Algorithmic Bridge (Alberto Romero), Tomasz Tunguz, Vercel blog. Source list and editorial profile maintained by Daniel.