flâneur — a map of the web's best reading

openai/mle-bench: MLE-bench is a benchmark for measuring how well AI agents perform at machine learning engineering ·

github.com · 2,393 words · saved by 1 readers

MLE-bench is a benchmark for measuring how well AI agents perform at machine learning engineering

Explore this link on the map →

saved by

related reading