✳flâneur — a map of the web's best reading
evmbench.pdf
cdn.openai.com · saved by 1 readers
N/A
Explore this link on the map →related reading
- netcs.cornell.edu
- Making sure you're not a bot!inria.hal.science
- GitHub - openai/mle-bench: MLE-bench is a benchmark for measuring how well AI agents perform at machine learning engineering · GitHubgithub.com
- GitHub - openai/parameter-golf: Train the smallest LM you can that fits in 16MB. Best model wins! · GitHubgithub.com
- GitHub - getmetal/motorhead: 🧠 Motorhead is a memory and information retrieval server for LLMs. · GitHubgithub.com
- Engram/Engram_paper.pdf at main · deepseek-ai/Engram · GitHubgithub.com
- GitHub - chenglou/pretext: Fast, accurate & comprehensive text measurement & layout · GitHubgithub.com
- hamming.pdfhomepages.inf.ed.ac.uk
- 2401.10166-VMambaarxiv.org
- EconEvals: Benchmarks and Litmus Tests for LLM Agents in Unknown Environmentsarxiv.org
- Untitledmogami.neocities.org
- edit.pdfcs.ox.ac.uk