Loading dataset details…
harbor run -d scale-ai/swe-atlas-twSWE-Atlas - Test Writing -- A benchmark of comprehensive test writing problems for coding agents. Checkout https://github.com/scaleapi/SWE-Atlas/ for instructions on running it.
harbor run -d scale-ai/swe-atlas-twA benchmark of comprehensive test-writing problems for coding agents.
See the repository for setup and usage instructions: