BIG-bench/keywords_to_tasks.md at main · google/BIG-bench
Beyond the Imitation Game collaborative benchmark for measuring and extrapolating the capabilities of language models - BIG-bench/keywords_to_tasks.md at main · google/BIG-bench
This file maps descriptive keywords (and key phrases) to tasks described by that keyword. To add new keywords, directly edit keywords.md rather than this file. Summary table Keyword Number of tasks Description traditional NLP tasks contextual question-answering 22 identifying the meaning of a particular word/sentence in a passage context-free question answering 24 responses rely on model's knowledge base, but not on context provided during query time reading comprehension 36 a superset of contextual question-answering, measuring the degree to which a model understands the content of a…
saved by
related reading
- Perplexityperplexity.ai
- AI Model & API Providers Analysis | Artificial Analysisartificialanalysis.ai
- Datacurve | The data engine for frontier AIdatacurve.ai
- Parsed | Custom, interpretable AI systems that continuously learnparsed.com
- Vals AIvals.ai
- knowledgator/gliner-multitask-large-v0.5 · Hugging Facehuggingface.co
- GitHub - brexhq/prompt-engineering: Tips and tricks for working with Large Language Models like OpenAI's GPT-4.github.com
- FrontierSWEfrontierswe.com
- There's An AI For That® — The front page of AItheresanaiforthat.com
- Branches · HazyResearch/intelligence-per-watt · GitHubgithub.com
- OpenAI | Research & Deploymentopenai.com
- MAI-Thinking-1 | Microsoft AImicrosoft.ai