BLAST RADIUS

Agents and harnesses · Benchmarks evals

Specific Publishes Benchmark For Enterprise Coding Agents

Sep 12, 2026

Specific introduced Real-SWE, a benchmark evaluating AI models on private, real-world enterprise codebases.

Read the original at cognition.comOpens the publisher's site in a new tab

More in Agents and harnesses