Skip to content

← AI Tracker

AI Briefing

Cursor Router, AMD-Anthropic Deal, OpenAI Presence

Freitag, 24. Juli 2026 - AI News · (letzte 24h)

Cursor ships an intelligent model router claiming frontier quality at 60% lower cost, while AMD and Anthropic sign a multi-gigawatt MI450 deal.

Must read

Tools & Frameworks

Vercel MCP can now deploy code

The Vercel MCP server exposes a deploy_to_vercel tool that ships code to a new or existing project and returns a shareable URL.

Why this matters: Your team already deploys to Vercel — this closes the loop from Claude Code to production.

Vercel Sandboxes: dashboard connect and snapshots

Vercel Sandboxes now support in-dashboard shell, filesystem, port inspection, snapshots, and stop/resume for persistent instances.

Why this matters: Useful primitive for headless agent sessions with reproducible state.

Ling 3.0 Flash on Vercel AI Gateway

Ant Group’s Ling 3.0 Flash — 124B MoE, ~5.1B active, 256K context, thinking/non-thinking modes — is free through August 3rd on AI Gateway.

Why this matters: Free window makes it cheap to A/B against your LiteLLM defaults for agentic workloads.

LangChain Eval Engineering Skill

LangChain shipped an Eval Engineering Skill that maps agent repos and production traces into executable Harbor evaluations.

Why this matters: Templated eval-generation from real traces is exactly the discipline layer your agentic dev workflow needs.

Together AI production platform for open-weight inference

Together pitches SLO-driven deployment for open-weight models with cost/perf/quality controls and staged rollouts.

Why this matters: Watch as a hybrid-routing option alongside your LiteLLM gateway.

Anthropic Economic Index connector for Claude

New Claude connector lets you query Anthropic’s Economic Index data on occupations, tasks, and automation trends in natural language.

Why this matters: Useful signal source for your writing on AI-native team design.

Open Models & Local

Ollama v0.32.3

Adds chat, thinking, and tool-calling support for Laguna 2.1 with a Metal inference fix; restores Claude Code channels and fixes Anthropic thinking streams.

Why this matters: Direct relevance to your Apple Silicon local-LLM setup and Claude Code integration.

llama.cpp b10093: DeepSeek4 template fix

Fixes DeepSeek4 chat template, adds drop_reasoning flag, and hooks the DS3.2 parser for DS4 with tool-result reordering.

Why this matters: If you’re evaluating DeepSeek V4 locally, these template fixes matter for tool-calling correctness.

Inside Poolside’s model factory — Eiso Kant

Poolside’s co-CEO details how a small research team trained Laguna S, a 118B MoE that beats Thinky’s ~1T open-weights model.

Why this matters: Concrete data point on the small-team model-training curve now that Laguna 2.1 is landing in Ollama.

Genesis-Science-1: DOE + Arcee open-weight model

US DOE and Arcee announced GS1, an open-weight model for scientific computing with reproducible training records and open contribution windows.

Why this matters: Watch — open-weight scientific tuning recipes may transfer to domain-specific coding models.

OpenAI’s infrastructure plan hits $750B

OpenAI reportedly raised planned infrastructure spending through 2030 to $750B, anchored by a $20B, 3.2GW Georgia campus.

Why this matters: Sets the frontier-lab compute baseline your team will price against.

Treasury threatens sanctions over Moonshot distillation claims

US Treasury signals sanctions on Chinese labs after allegations Moonshot’s Kimi K3 was distilled from Anthropic’s Fable.

Why this matters: Direct impact on availability of Chinese open models in your local stack — worth tracking.

Will Kimi K3 change the economics of AI?

Azeem Azhar on how Kimi K3’s cost and capability profile pressures Western frontier pricing.

Why this matters: Pairs with the Treasury story to frame near-term model-routing decisions.

Anthropic building Claude ‘managed projects’

Anthropic is developing persistent, semi-autonomous ‘managed projects’ inside Claude for organised long-running task management.

Why this matters: Early signal on where Claude Code + Projects are heading — relevant to your overnight-agent-factory pattern.

Pragmatic Engineer: Chinese open models match closed frontier

Orosz covers Chinese open models now matching Anthropic/OpenAI on key evals, plus AWS billing incidents and Spotify reliability.

Why this matters: Useful cross-check on the Kimi/Laguna narrative from an engineering-leadership lens.


Sources unavailable today: r/ChatGPTCoding top, r/ClaudeAI top, r/LocalLLaMA top, r/MachineLearning top

Auto-curated daily by Claude Opus 4.7 from Ben’s Bites, Don’t Worry About the Vase (Zvi), Exponential View (Azeem Azhar), GitHub: cline/cline, GitHub: ggml-org/llama.cpp, GitHub: langchain-ai/langchain, GitHub: ollama/ollama, Hugging Face blog, LangChain blog, Latent Space, NVIDIA developer blog, One Useful Thing (Ethan Mollick), OpenAI blog, SaaStr (Jason Lemkin), Simon Willison, Sourcegraph blog, TLDR AI, The Pragmatic Engineer (Gergely Orosz), Together AI blog, Vercel blog, smol.ai news. Source list and editorial profile maintained by Daniel.