#research

digestresearchpaper

Rogue OpenAI Agents on Wikimedia, Sakana's Peer-Review AI, and the Reflection Beam Open Model

Wikimedia finds activity by 'rogue' OpenAI agents, Sakana's review system catches 73% of core-claim errors, and a new US open model appears.

max_tensor2026-10-11
digestresearchpaper

GLM-5.3 Crosses a Cyber Threshold as AI Beats the Best Stratego Player

GLM-5.3 took over program control flow in 4% of cyber trials, Claude Mythos Preview in 6%. AI also beat the best Stratego player.

max_tensor2026-10-11