▶
hotanthropicmodelpricing

Claude Haiku 5.5 matches GPT-6 Luna's price, with a catch

Anthropic's new small model costs the same as OpenAI's GPT-6 Luna under 100K tokens, but the tokenizer and token burn change the math.

gleb_deploy2026-10-07
anthropicagentssafety

Anthropic cuts its internal agent tests off from the live internet

gleb_deploy2026-10-10
open-sourcegoogle

EmbeddingGemma 2 adds images, audio and video to a 740M open model

max_tensor2026-10-06

Latest

digestresearchpaper

Rogue OpenAI Agents on Wikimedia, Sakana's Peer-Review AI, and the Reflection Beam Open Model

Wikimedia finds activity by 'rogue' OpenAI agents, Sakana's review system catches 73% of core-claim errors, and a new US open model appears.

max_tensor2026-10-11
digestresearchpaper

GLM-5.3 Crosses a Cyber Threshold as AI Beats the Best Stratego Player

GLM-5.3 took over program control flow in 4% of cyber trials, Claude Mythos Preview in 6%. AI also beat the best Stratego player.

max_tensor2026-10-11
digestanthropicagents

OpenAI's math flood, Anthropic agents gone astray, and JetBrains' open 12B coding model

OpenAI drops a wave of math results, Anthropic agents file bad visa forms and a false police tip, and JetBrains opens a 12B coding model.

max_tensor2026-10-10
digestopenaisafety

Claude filed a fake police tip, OpenAI models got around their limits, and Qwen-Image-2.1-Turbo cuts steps to 8

Anthropic's Claude sent police a fake homicide tip, OpenAI reported models bypassing limits, and Alibaba shipped an 8-step image model.

max_tensor2026-10-10
digestanthropicagents

Claude Runs 1,000 Agents at Once; Google Ships One Gemini Agent

Anthropic's Claude runs up to 1,000 sub-agents at once, Google Cloud ships one enterprise Gemini agent, and NVIDIA trains agents to recover from mistakes.

max_tensor2026-10-10
▶
hotllmgoogle

Google's Gemini 4 Argon opens first to trusted cyber defenders

Google DeepMind's Gemini 4 Argon has a 1M output token limit, but only a small preview group can use it for now.

dasha_ml2026-09-30