DeepSeek releases its efficient V4.1-Flash model for long-horizon agents, while OpenAI launches a public beta for its Agents API, expands GPT-6 Astra, introduces a finance chatbot, and releases GPT-Live-1.
DeepSeek launches highly efficient V4.1-Flash model for AI agents
DeepSeek's new V4.1-Flash model reduces memory strain by keeping only 16 billion of its 552 billion parameters active per token. The update cuts KV cache needs to a quarter of its predecessor.
OpenAI launches Agents API in public beta for long-running workflows
The managed service lets developers build cloud agents that execute code and run autonomously for hours. It carries no extra fees beyond standard token usage.
GPT-6 Astra expands to coding and math, straining OpenAI capacity
OpenAI rolled out GPT-6 Astra for coding and cybersecurity tasks, concurrently topping open math benchmarks. The resulting compute demand forced a temporary pause on new Pro subscriptions.
OpenAI introduces specialized ChatGPT for Wall Street and financial services
Powered by GPT-6 Astra, the tailored chatbot helps finance professionals develop research and execute modeling calculations. The platform integrates built-in financial data to generate client-ready materials.
OpenAI releases GPT-Live-1 API for full-duplex voice applications
The new model allows applications to talk and listen simultaneously without turn-taking delays. Priced at five cents per minute, it scored 80.1 percent in interactivity tests.