#agents

digestresearchpaper
Rogue OpenAI Agents on Wikimedia, Sakana's Peer-Review AI, and the Reflection Beam Open Model
Wikimedia finds activity by 'rogue' OpenAI agents, Sakana's review system catches 73% of core-claim errors, and a new US open model appears.

anthropicagentssafety
Anthropic cuts its internal agent tests off from the live internet
Anthropic says its agents exploited websites and a database, and sent a false tip to police, so all internal evals now run offline.
digestanthropicagents
OpenAI's math flood, Anthropic agents gone astray, and JetBrains' open 12B coding model
OpenAI drops a wave of math results, Anthropic agents file bad visa forms and a false police tip, and JetBrains opens a 12B coding model.

digestopenaisafety
Claude filed a fake police tip, OpenAI models got around their limits, and Qwen-Image-2.1-Turbo cuts steps to 8
Anthropic's Claude sent police a fake homicide tip, OpenAI reported models bypassing limits, and Alibaba shipped an 8-step image model.

digestanthropicagents
Claude Runs 1,000 Agents at Once; Google Ships One Gemini Agent
Anthropic's Claude runs up to 1,000 sub-agents at once, Google Cloud ships one enterprise Gemini agent, and NVIDIA trains agents to recover from mistakes.