OpenAI released new AI-generated solutions to long‑standing math problems, demonstrating models can tackle geometry, cryptography and complexity.
Google's Gemini API now lets a single call coordinate reasoning, code execution and web retrieval inside an isolated cloud sandbox.
Anthropic clarified its stance amid US discussions on banning Chinese open‑weights models, countering claims it seeks to restrict them.
OpenAI's upcoming Astra model family will coordinate multiple agents for long‑running, complex tasks and will undergo U.S. government review.
AMD released Instella‑MoE‑16B‑A3B, a 16B‑parameter Mixture‑of‑Experts LLM with 2.8B active parameters and open weights.
OpenAI's Codex Security CLI and TypeScript SDK help locate and fix code vulnerabilities, requiring Node 22+ or Python 3.10+.
Moonshot Labs' Kimi K3 and Alibaba's Qwen 3.8 launch publicly, challenging top‑tier models by matching Anthropic's Fable 5 performance.
OpenAI field report shows AI coding agents can speed up legacy research software up to 60×, but they cannot verify scientific correctness.
Supabase open‑sourced Evals, a benchmark that runs Claude Code, Codex and OpenCode on real Supabase tasks, with a public leaderboard.
A new field report details how scientists use AI coding agents to modernize scientific software, accelerating development in genomics and data‑rich fields.
OpenAI offers 100,000 researchers free access to its most advanced ChatGPT models to boost scientific discovery.
OpenAI's ChatGPT for Small Businesses program equips entrepreneurs with AI tools to automate tasks and expand capabilities.