AI News 2026-08-01
New To You · 12 of 326 stories

OpenAI unveils Astra and expands free ChatGPT access as Google rolls out Gemini Managed Agents

49 sources swept 326 distinct stories from the last 12 days 36 already covered, held back 12 worth your time
  1. Ten advances in mathematics and theoretical computer science

    OpenAI released new AI-generated solutions to long‑standing math problems, demonstrating models can tackle geometry, cryptography and complexity.

    OpenAI··Research·lab·3 outlets·openai.com
    Also covered bySimon Willison
  2. Gemini API Managed Agents: 3.6 Flash, hooks, and more

    Google's Gemini API now lets a single call coordinate reasoning, code execution and web retrieval inside an isolated cloud sandbox.

    Google AI··Agents & tooling·lab·blog.google
  3. Our position on open-weights models

    Anthropic clarified its stance amid US discussions on banning Chinese open‑weights models, countering claims it seeks to restrict them.

    Anthropic··Safety & security·lab·2 outlets·anthropic.com
  4. OpenAI announces its "next major model" Astra by dropping ten previously unsolved math solutions

    OpenAI's upcoming Astra model family will coordinate multiple agents for long‑running, complex tasks and will undergo U.S. government review.

    The Decoder··Models & releases·press·the-decoder.com
  5. AMD Releases Instella-MoE-16B-A3B: A Fully Open Mixture-of-Experts LLM With 2.8B Active Parameters Trained On Instinct GPUs

    AMD released Instella‑MoE‑16B‑A3B, a 16B‑parameter Mixture‑of‑Experts LLM with 2.8B active parameters and open weights.

    Marktechpost··Models & releases·newsletter·marktechpost.com
  6. Codex Security

    OpenAI's Codex Security CLI and TypeScript SDK help locate and fix code vulnerabilities, requiring Node 22+ or Python 3.10+.

    Hacker News AI··Products & apps·HN·596 pts·github.com
  7. Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling

    Moonshot Labs' Kimi K3 and Alibaba's Qwen 3.8 launch publicly, challenging top‑tier models by matching Anthropic's Fable 5 performance.

    Hacker News AI··Models & releases·HN·371 pts·emergingtrajectories.com
  8. AI coding agents can modernize research software but can't judge if the science is right

    OpenAI field report shows AI coding agents can speed up legacy research software up to 60×, but they cannot verify scientific correctness.

    The Decoder··Agents & tooling·press·the-decoder.com
  9. Supabase Releases Evals: an Open Source Benchmark That Scores Claude Code, Codex and OpenCode on Real Supabase Tasks

    Supabase open‑sourced Evals, a benchmark that runs Claude Code, Codex and OpenCode on real Supabase tasks, with a public leaderboard.

    Marktechpost··Agents & tooling·newsletter·marktechpost.com
  10. Scientific computing in the age of agentic AI

    A new field report details how scientists use AI coding agents to modernize scientific software, accelerating development in genomics and data‑rich fields.

    OpenAI··Agents & tooling·lab·openai.com
  11. Accelerating scientific discovery with ChatGPT for Academic Researchers

    OpenAI offers 100,000 researchers free access to its most advanced ChatGPT models to boost scientific discovery.

    OpenAI··Business & policy·lab·openai.com
  12. Introducing the ChatGPT for small business program

    OpenAI's ChatGPT for Small Businesses program equips entrepreneurs with AI tools to automate tasks and expand capabilities.

    OpenAI··Business & policy·lab·openai.com