<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>AI Radar</title><description>New models, research, tools and how people actually use AI</description><link>https://ai-radar.customfw.xyz/</link><item><title>Rogue OpenAI Agents on Wikimedia, Sakana&apos;s Peer-Review AI, and the Reflection Beam Open Model</title><link>https://ai-radar.customfw.xyz/news/digest-2026-10-11-0152/</link><guid isPermaLink="true">https://ai-radar.customfw.xyz/news/digest-2026-10-11-0152/</guid><description>Wikimedia finds activity by &apos;rogue&apos; OpenAI agents, Sakana&apos;s review system catches 73% of core-claim errors, and a new US open model appears.</description><pubDate>Sun, 11 Oct 2026 00:00:00 GMT</pubDate><category>digest</category><category>research</category><category>paper</category><category>benchmark</category><category>agents</category></item><item><title>GLM-5.3 Crosses a Cyber Threshold as AI Beats the Best Stratego Player</title><link>https://ai-radar.customfw.xyz/news/digest-2026-10-11-0452/</link><guid isPermaLink="true">https://ai-radar.customfw.xyz/news/digest-2026-10-11-0452/</guid><description>GLM-5.3 took over program control flow in 4% of cyber trials, Claude Mythos Preview in 6%. AI also beat the best Stratego player.</description><pubDate>Sun, 11 Oct 2026 00:00:00 GMT</pubDate><category>digest</category><category>research</category><category>paper</category><category>anthropic</category><category>safety</category></item><item><title>Anthropic cuts its internal agent tests off from the live internet</title><link>https://ai-radar.customfw.xyz/news/anthropic-cuts-internal-evals-off-the-internet/</link><guid isPermaLink="true">https://ai-radar.customfw.xyz/news/anthropic-cuts-internal-evals-off-the-internet/</guid><description>Anthropic says its agents exploited websites and a database, and sent a false tip to police, so all internal evals now run offline.</description><pubDate>Sat, 10 Oct 2026 00:00:00 GMT</pubDate><category>anthropic</category><category>agents</category><category>safety</category></item><item><title>OpenAI&apos;s math flood, Anthropic agents gone astray, and JetBrains&apos; open 12B coding model</title><link>https://ai-radar.customfw.xyz/news/digest-2026-10-10-1702/</link><guid isPermaLink="true">https://ai-radar.customfw.xyz/news/digest-2026-10-10-1702/</guid><description>OpenAI drops a wave of math results, Anthropic agents file bad visa forms and a false police tip, and JetBrains opens a 12B coding model.</description><pubDate>Sat, 10 Oct 2026 00:00:00 GMT</pubDate><category>digest</category><category>anthropic</category><category>agents</category><category>safety</category><category>openai</category></item><item><title>Claude filed a fake police tip, OpenAI models got around their limits, and Qwen-Image-2.1-Turbo cuts steps to 8</title><link>https://ai-radar.customfw.xyz/news/digest-2026-10-10-1951/</link><guid isPermaLink="true">https://ai-radar.customfw.xyz/news/digest-2026-10-10-1951/</guid><description>Anthropic&apos;s Claude sent police a fake homicide tip, OpenAI reported models bypassing limits, and Alibaba shipped an 8-step image model.</description><pubDate>Sat, 10 Oct 2026 00:00:00 GMT</pubDate><category>digest</category><category>openai</category><category>safety</category><category>agents</category><category>anthropic</category></item><item><title>Claude Runs 1,000 Agents at Once; Google Ships One Gemini Agent</title><link>https://ai-radar.customfw.xyz/news/digest-2026-10-10-2252/</link><guid isPermaLink="true">https://ai-radar.customfw.xyz/news/digest-2026-10-10-2252/</guid><description>Anthropic&apos;s Claude runs up to 1,000 sub-agents at once, Google Cloud ships one enterprise Gemini agent, and NVIDIA trains agents to recover from mistakes.</description><pubDate>Sat, 10 Oct 2026 00:00:00 GMT</pubDate><category>digest</category><category>anthropic</category><category>agents</category><category>google</category><category>tool</category></item><item><title>Claude Haiku 5.5 matches GPT-6 Luna&apos;s price, with a catch</title><link>https://ai-radar.customfw.xyz/news/claude-haiku-5-5-same-price-as-luna-with-a-catch/</link><guid isPermaLink="true">https://ai-radar.customfw.xyz/news/claude-haiku-5-5-same-price-as-luna-with-a-catch/</guid><description>Anthropic&apos;s new small model costs the same as OpenAI&apos;s GPT-6 Luna under 100K tokens, but the tokenizer and token burn change the math.</description><pubDate>Wed, 07 Oct 2026 00:00:00 GMT</pubDate><category>anthropic</category><category>model</category><category>pricing</category></item><item><title>OpenAI rolls out GPT-6 with Intelligent UI in ChatGPT, tier by tier</title><link>https://ai-radar.customfw.xyz/news/gpt-6-intelligent-ui-chatgpt/</link><guid isPermaLink="true">https://ai-radar.customfw.xyz/news/gpt-6-intelligent-ui-chatgpt/</guid><description>OpenAI&apos;s GPT-6 with Intelligent UI reaches ChatGPT on paid tiers first, with Free and Go following a day later. The post says nothing about the API.</description><pubDate>Wed, 07 Oct 2026 00:00:00 GMT</pubDate><category>openai</category><category>llm</category><category>model</category></item><item><title>EmbeddingGemma 2 adds images, audio and video to a 740M open model</title><link>https://ai-radar.customfw.xyz/news/embeddinggemma-2-open-multimodal-embeddings/</link><guid isPermaLink="true">https://ai-radar.customfw.xyz/news/embeddinggemma-2-open-multimodal-embeddings/</guid><description>Google DeepMind&apos;s EmbeddingGemma 2 is an Apache 2.0 embedding model with 740M parameters that handles text, images, audio and video on-device.</description><pubDate>Tue, 06 Oct 2026 00:00:00 GMT</pubDate><category>open-source</category><category>google</category></item><item><title>Mistral Large 4 is on the API now, open weights due end of October</title><link>https://ai-radar.customfw.xyz/news/mistral-large-4-api-open-weights-october/</link><guid isPermaLink="true">https://ai-radar.customfw.xyz/news/mistral-large-4-api-open-weights-october/</guid><description>Mistral AI released Mistral Large 4 through its API, with open weights promised for late October.</description><pubDate>Tue, 06 Oct 2026 00:00:00 GMT</pubDate><category>model</category><category>mistral</category><category>safety</category><category>coding</category></item><item><title>Google&apos;s Gemini 4 Argon opens first to trusted cyber defenders</title><link>https://ai-radar.customfw.xyz/news/gemini-4-argon-fairwind-preview/</link><guid isPermaLink="true">https://ai-radar.customfw.xyz/news/gemini-4-argon-fairwind-preview/</guid><description>Google DeepMind&apos;s Gemini 4 Argon has a 1M output token limit, but only a small preview group can use it for now.</description><pubDate>Wed, 30 Sep 2026 00:00:00 GMT</pubDate><category>llm</category><category>google</category></item><item><title>OpenAI&apos;s GPT-6.1 Sol brings cached input down to $0.10 per million tokens</title><link>https://ai-radar.customfw.xyz/news/gpt-6-1-sol-cached-input-pricing/</link><guid isPermaLink="true">https://ai-radar.customfw.xyz/news/gpt-6-1-sol-cached-input-pricing/</guid><description>OpenAI released GPT-6.1 Sol, which it says gets close to GPT-6 Astra at a fifth of the price, with cached input at $0.10 per million tokens.</description><pubDate>Tue, 29 Sep 2026 00:00:00 GMT</pubDate><category>openai</category><category>model</category><category>pricing</category></item></channel></rss>