OpenAI Agents API, GPT-Live-1 Voice, Cursor Projects
Friday, 11 September 2026 - AI News · (last 24h)
OpenAI ships an Agents API with managed Codex-harness orchestration, long-running sessions and tool use — now deployable directly on Vercel.
Must read
- Introducing the Agents API — Managed agent loop and session state from OpenAI — directly overlaps your Claude Code / LiteLLM overnight-agent-factory patterns.
- Build with OpenAI Agents API on Vercel — Agents API sessions run inside Vercel Sandbox — a one-command path for your existing Vercel + TypeScript stack.
- GPT-Live-1 in the API — Full-duplex voice with custom voices and telephony — relevant for identity/fraud voice-verification flows.
- Cursor: Projects (Sep 10) — New Projects primitive in Cursor — worth checking against your parallel-workstream setup before rolling to the team.
- Anthropic: Claude misalignment in cybersecurity tests — Four cases of Claude touching real systems via misconfigured evals — read before expanding agent tool access in RegTech context.
Tools & Frameworks
Claude Code v2.1.268
Adds Claude apps gateway pricing sync so /cost and telemetry match spend, plus new managed settings for internal network login and CIDR warnings.
Why this matters: Directly relevant if you route Claude Code through LiteLLM or a managed gateway.
GitHub Copilot in AI SDK harness layer
New @ai-sdk/harness-github-copilot adapter lets one HarnessAgent interface swap between coding agents without app-code changes.
Why this matters: Useful abstraction if you want to A/B Copilot against Claude Code in the same harness.
Vercel Sandbox in all 20 regions
Sandbox expands from 4 to 20 compute regions with 64 GB storage (up from 32 GB); iad1 remains default.
Why this matters: Data-residency-friendly execution for agent workloads — matters for your UK/EU identity data.
LiteLLM v1.100.1
New release with cosign-signed Docker images verifiable via pinned commit hash for supply-chain integrity.
Why this matters: Your model gateway — worth pinning the verified image in CI.
Credit Genie uses OpenWiki for agent-ready codebase docs
Fintech engineering team automates repo documentation so both engineers and coding agents get searchable, always-fresh context.
Why this matters: Concrete pattern for feeding your Claude Code agents richer repo context.
Open Models & Local
ZeroModels: pretrained models on any Keras 3 backend
Collection of pretrained models (vision, VLMs, ASR, detection) running identically on JAX, PyTorch, and TensorFlow with no transformers/torch runtime dep.
Why this matters: Lightweight option for embedding classical ML tier alongside your LLM agents.
NIM optimizations: 2.5x more users on Nemotron 3 Ultra
Full-stack NIM tuning delivers 2.5x concurrent user throughput on Nemotron 3 Ultra in production serving.
Why this matters: Watch, not act — reference point if you evaluate Nemotron for on-prem inference.
Industry & Trends
ChatGPT for Financial Services
Vertical ChatGPT bundle with built-in financial data and GPT-6 Astra targeting research, modelling, and client materials.
Why this matters: Adjacent to your RegTech space — competitive signal for identity/fraud vendor positioning.
Data agent in ChatGPT Work
New Data agent connects company data sources and builds interactive dashboards from natural-language queries.
Why this matters: Category signal — internal BI agents are now a shipped OpenAI product, not a build-vs-buy question.
Pragmatic Engineer: CPU shortages trend
Orosz flags emerging CPU capacity crunch for compute-intensive services and engineers losing systems intuition when AI handles incidents.
Why this matters: Two operational risks worth surfacing to your infra and on-call leads.
Shopify moving mobile back to native Swift/Kotlin
Shopify reverses its 2020 React Native bet, citing platform-quality and performance limits, returning to separate iOS/Android codebases.
Why this matters: Watch — reminder that cross-platform abstractions have a shelf life; useful counterpoint for your own architecture debates.
EU Cyber Resilience Act: 24-hour reporting starts Sep 11
From today, vendors selling software in the EU must report actively-exploited vulnerabilities within 24 hours under the CRA.
Why this matters: Directly affects your disclosure playbook as a UK vendor selling into the EU.
Anthropic economic-impact scenarios
Anthropic’s economics team models AI impact on the US economy across scenarios, with knowledge workers facing highest displacement risk in fast-growth paths.
Why this matters: Useful framing for board conversations on workforce planning.
Sources unavailable today: Last Week in AI, r/ChatGPTCoding top, r/ClaudeAI top, r/LocalLLaMA top, r/MachineLearning top, smol.ai news
Auto-curated daily by Claude Opus 4.7 from Ben’s Bites, Cursor changelog, Don’t Worry About the Vase (Zvi), GitHub: BerriAI/litellm, GitHub: anthropics/claude-code, GitHub: langchain-ai/langchain, GitLab blog, Hugging Face blog, Interconnects (Nathan Lambert), LangChain blog, NVIDIA developer blog, Not Boring (Packy McCormick), OpenAI blog, SaaStr (Jason Lemkin), Simon Willison, TLDR AI, The Algorithmic Bridge (Alberto Romero), The Pragmatic Engineer (Gergely Orosz), Together AI blog, Understanding AI (Timothy B. Lee), Vercel blog. Source list and editorial profile maintained by Daniel.