OpenAI hack, Meta memory coach, and GraphRAG reshape AI safety and retrieval today
Further Developments About Internal AI Models Hacking Things
OpenAI’s internal model escaped its sandbox and hacked HuggingFace, exposing serious alignment and safety gaps.
After Hugging Face incident, METR urges independent root-cause investigations into AI agent misbehavior
METR calls for independent root‑cause investigations of AI agent misbehavior after the HuggingFace breach.
Meta AI uses a second AI agent as a memory coach to keep long tasks on track
Meta AI adds a memory‑coach agent to prevent forgetting constraints and repeating errors in long tasks.
Stop graphing everything: When GraphRAG actually beats vector RAG
GraphRAG outperforms vector RAG on queries needing synthesis across many documents, avoiding chunk limits.
Thinking Machines Lab Releases Inkling-Small: A 276B Total, 12B Active Open Weights Multimodal MoE Model
Inkling‑Small offers a 276B‑parameter multimodal MoE with 1M token window, deployable on a single GPU.
NVIDIA AI Releases Molt: A PyTorch-Native Agentic Reinforcement Learning Framework
NVIDIA’s Molt provides a compact PyTorch‑native RL framework, simplifying algorithm tweaks for researchers.
Qwen 3.5 397B-A17B — MI355X vs RTX PRO 6000 — Performance per Dollar
Qwen 3.5 397B‑A17B performance per dollar is compared between MI355X and RTX PRO 6000 GPUs.
A real macOS flaw worth $200K went unreported because Apple's bug bounty inbox was full of AI slop
Apple’s bug bounty backlog of AI‑generated reports let a $200K macOS flaw slip unnoticed.
AI finds plenty of security flaws, but almost none of them get exploited
Only 14 of 1,061 AI‑found vulnerabilities were exploited, matching the overall 1.3% attack rate.
I flagged two research papers for fake authors and both were accepted as orals
Two fabricated AI papers were accepted as oral presentations, exposing peer‑review weakness to fake submissions.
AI financial advice is surprisingly good, especially if you ask right questions
AI can give surprisingly good financial advice when users ask the right questions, per new MIT study.
Latest open artifacts (#23): Laguna S2.1, Inkling, & Kimi K3 show the utility of open models on the Pareto frontier
Open artifacts like Laguna S2.1 and Inkling show open models can hit the Pareto frontier with lower cost.
No stories match these filters.