Frontier Line and access · benchmarks-evals
Google DeepMind pilots cryptographic double-blind frontier-model evaluation
Google DeepMind announced a pilot using cryptographic methods to conduct a double-blind evaluation of a proprietary frontier AI model.
Read the original at glenrhodes.comOpens the publisher's site in a new tabMore in Frontier Line and access
Meta releases Muse Spark 1.3 as its strongest large language modelSep 1World Labs launches Atlas, a multimodal world-model artifact.Sep 2Abliteration.ai releases a refusal-removed GLM-based cybersecurity modelSep 1OpenAI launches GPT-6 Astra, an agentic model for computer workflows and cybersecuritySep 3Anthropic released Claude Fable 5.1 broadly and Claude Mythos 5.1 to vetted organizations.Sep 1