OpenAI Makes Codex-Powered Agents Public Beta #
OpenAI is releasing its Agents API as a public beta. The managed service lets developers build cloud agents that run autonomously for hours, execute code, use tools and hand off tasks to sub-agents. OpenAI says developers pay no additional platform fee beyond token usage.
OpenAI has also released GPT-Live-1 as a developer API for applications that listen and speak simultaneously. The full-duplex speech model scored 80.1% on interactivity tests, up from 45.4% for its predecessor, and costs $0.05 per minute.
DeepSeek's V4.1-Flash offers a different efficiency tradeoff: its 552 billion-parameter model uses one-quarter the KV-cache memory of its predecessor while activating 16 billion parameters per token. It narrowly beat Opus 5 and GPT-5.6 Sol on the DeepSWE coding benchmark. Separately, TechCrunch reports that AI agents are flooding public services with new requests; one researcher said the vast majority involved people claiming something they were entitled to receive.
Why it matters: For developers, OpenAI now manages orchestration, long-running sessions and tool use through a Codex-powered service, while DeepSeek's V4.1-Flash cuts KV-cache memory to one-quarter of its predecessor.
Key Takeaways
- GPT-Live-1 scored 80.1% on interactivity tests, compared with 45.4% for its predecessor, at $0.05 per minute
- DeepSeek V4.1-Flash contains 552 billion parameters but activates 16 billion per token and narrowly beat Opus 5 and GPT-5.6 Sol on DeepSWE
- TechCrunch reports AI agents are flooding public services with requests, and a researcher said most cases involved people claiming benefits or services they were entitled to receive