The Eval Index / Coding Eval / #189
openai/SWELancer-Benchmark
by openai · Coding Eval · updated 1y ago
This repo contains the dataset and code for the paper "SWE-Lancer: Can Frontier LLMs Earn $1 Million from Real-World Freelance Software Engineering?"
38
momentum
1,431
stars
136
forks
#189
rank