Kimi K3, Claude Code Bun/Rust, Cline Team Runs
Sonntag, 19. Juli 2026 - AI News · (letzte 24h)
Moonshot’s Kimi K3 lands with reviewers claiming frontier-level coding, while Claude Code quietly ships a Rust-rewritten Bun runtime.
Must read
- Kimi K3 surprise & AI economics — Moonshot’s K3 is the week’s model story; Cline already routes to it and it’s a candidate for your LiteLLM gateway.
- Claude Code uses Bun written in Rust now — Your overnight agent factory is now running on a Rust runtime — 10% faster startup, matters when you spawn dozens of headless sessions.
- Claude Code v2.1.215: /verify and /code-review now manual — Anthropic pulled auto-invocation of verify/code-review skills — update your CLAUDE.md and team workflows or you’ll silently lose that check.
- Cline SDK v0.0.65: team runs, Kimi K3, slimmer install — Team-runs semantics tightened (spawn hidden from teammates, errors report correctly) — relevant if you’re evaluating multi-agent orchestration beyond Claude Code.
Tools & Frameworks
Cline CLI v3.0.45 drops install from 640MB to 285MB
Claude Code and Codex providers are now optional peer deps, loaded on demand; OAuth refresh retries added; Kimi K3 in ClinePass fallback.
Why this matters: Lighter CLI matters for CI runners and ephemeral agent containers.
Controlling Reasoning Effort in LLMs
Raschka on how low/medium/high reasoning modes are actually trained into models, not just prompt-toggles.
Why this matters: Useful mental model when routing between reasoning tiers in your LiteLLM gateway.
Open Models & Local
llama.cpp b10068: DFlash K/V cache rotation with quantisation
Fixes injected K/V cache rotation when using quantised K/V for DFlash models — unblocks long-context runs at lower memory.
Why this matters: Direct impact if you run quantised local models on Apple Silicon for coding tasks.
llama.cpp b10067: DeepSeek-V4 quantisation fix
Excludes DeepSeek-V4’s i32 ffn_gate_tid2eid routing table from quantisation; llama-quantize no longer fails on the MoE index tensor.
Why this matters: Unblocks local DeepSeek-V4 quants — worth a retry if you shelved it after quantisation errors.
Moonshot’s Kimi K3: the frontier from another planet
Romero argues Kimi K3 matches or beats US frontier labs on coding and reasoning, at a fraction of cost.
Why this matters: Watch-and-test: if K3 weights land open, it becomes a real local/hybrid coding option.
Industry & Trends
Claude Fable 5 becomes permanent on Max and Team Premium
From July 20, Fable 5 is included in Max and Team Premium at 50% of standard limits; Pro/Team Standard get a one-time $100 credit.
Why this matters: Plan-mix decision for your team’s Claude seats — Fable access without burning credits changes the calculus.
Sources unavailable today: r/ChatGPTCoding top, r/ClaudeAI top, r/LocalLLaMA top, r/MachineLearning top
Auto-curated daily by Claude Opus 4.7 from Exponential View (Azeem Azhar), GitHub: anthropics/claude-code, GitHub: cline/cline, GitHub: ggml-org/llama.cpp, Lenny’s Newsletter, SaaStr (Jason Lemkin), Sebastian Raschka, Simon Willison, The Algorithmic Bridge (Alberto Romero). Source list and editorial profile maintained by Daniel.