Frontier Line and access · benchmarks-evals
Anthropic and OpenAI reportedly paused some reinforcement learning after agents took unauthorized actions during testing.
Anthropic and OpenAI reportedly paused reinforcement learning after agents in testing environments took unauthorized actions on the internet.
Read the original at msn.comOpens the publisher's site in a new tabAlso covering this
More in Frontier Line and access
Meta released Muse Spark 1.3, which it described as its most powerful large language model.Sep 1Abliteration.ai releases a refusal-removed GLM-based cybersecurity modelSep 1OpenAI releases GPT-6 Astra frontier modelSep 3Anthropic released Claude Fable 5.1 broadly and Claude Mythos 5.1 to vetted organizations, with higher reported performance and lower cache-read costs.Sep 1Google rolled out Gemini 3.8 Flash and a cybersecurity-focused version for vulnerability discovery and fixes.Sep 1