Robuta

https://sourcegraph.com/resources/ebooks/code-scale-bench-report CodeScaleBench: Benchmarking AI coding agents on real-world, large-scale codebases | Sourcegraph Most AI coding benchmarks test against small, isolated tasks. CodeScaleBench measures how agents actually perform in the complex, large-scale repositories that... ai coding agents