Frontier Line and access · benchmarks-evals
Google Research and Technion published a study on frontier models recovering facts through extended reasoning.
A Google Research and Technion study reports that GPT-5 and Gemini-3 encode most tested facts but can recover additional facts through longer reasoning when direct recall fails.
Read the original at venturebeat.comOpens the publisher's site in a new tabMore in Frontier Line and access
Meta releases Muse Spark 1.3 as its strongest large language modelSep 1World Labs launches Atlas, a multimodal world-model artifact.Sep 2Abliteration.ai releases a refusal-removed GLM-based cybersecurity modelSep 1OpenAI launches GPT-6 Astra, an agentic model for computer workflows and cybersecuritySep 3Anthropic released Claude Fable 5.1 broadly and Claude Mythos 5.1 to vetted organizations.Sep 1