✳flâneur — a map of the web's best reading
nverse-scaling/prize: A prize for finding tasks that cause large language models to show inverse scaling
github.com · 6,547 words · saved by 1 readers
A prize for finding tasks that cause large language models to show inverse scaling
Explore this link on the map →saved by
related reading
- AI Model & API Providers Analysis | Artificial Analysisartificialanalysis.ai
- GitHub - brexhq/prompt-engineering: Tips and tricks for working with Large Language Models like OpenAI's GPT-4. · GitHubgithub.com
- There's An AI For That® — The front page of AItheresanaiforthat.com
- Parsed | Custom, interpretable AI systems that continuously learnparsed.com
- GitHub - jacobhilton/deep_learning_curriculum: Language model alignment-focused deep learning curriculum · GitHubgithub.com
- Branches · HazyResearch/intelligence-per-watt · GitHubgithub.com
- GitHub - google-research/tuning_playbook: A playbook for systematically maximizing the performance of deep learning models. · GitHubgithub.com
- ML Contestsmlcontests.com
- Neuronpedianeuronpedia.org
- GitHub - x1xhlol/system-prompts-and-models-of-ai-tools: FULL Augment Code, Claude Code, Cluely, CodeBuddy, Comet, Cursor, Devin AI, Junie, Kiro, Leap.new, Lovable, Manus, NotionAI, Orchids.app, Perplexity, Poke, Qoder, Replit, Same.dev, Tragithub.com
- GitHub - affaan-m/ECC: The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond. · GitHubgithub.com
- GitHub - open-compass/opencompass: OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, etc) over 100+ datasets. · GitHubgithub.com