
ResearchSep 11, 2026·4 min read
Specific Labs Introduces Real-SWE Benchmark Using Private Codebases
Specific Labs has launched Real-SWE, evaluating AI coding agents on licensed, private enterprise codebases across 640 scored rollouts.
Read more →
AnnouncementsJul 8, 2026·4 min read
xAI Releases Grok 4.5: Benchmarks, Integrations, and Pricing
xAI has launched Grok 4.5, a new large language model optimized for software engineering and knowledge tasks. It is trained on NVIDIA GB300 GPUs.
Read more →