TableLlama: Towards Open Large Generalist Models for Tables
In this work, we introduce TableLlama and TableInstruct, the first large open-source generalist model and instruction tuning dataset for tables. Everything is open-source right now! TableInstruct is a large-scale instruction tuning dataset with diverse, realistic tasks based on real-world tables. TableInstruct boasts a collection of 14 datasets of 11 tasks in total, which is curated from 1.24M tables containing 2.6M instances: TableLlama is a large generalist model for tables based on Llama 2 (7B) and LongLoRA, which can: Figure 1: An overview of TableInstruct and TableLlama. TableInstruct includes a wide variety of realistic tables and tasks with instructions. We make the first step towards developing open-source generalist models for tables with TableInstruct and TableLlama. Semi-structured tables are ubiquitous. There has been a variety of tasks that aim to automatically interpret, augment, and query tables. Current methods often require pretraining on tables or special model archit
TableLlama: Towards Open Large Generalist Models for Tables TableLlama: Towards Open Large Generalist Models for Tables TableLlama: Towards Open Large Generalist Models for Tables --> Tianshu Zhang , Xiang Yue , Yifei Li , Huan Sun The Ohio State University Conferance name and year --> zhang.11535@osu.edu , yue.149@osu.edu , li.14042@osu.edu , sun.397@osu.edu * Indicates Equal Contribution --> 🤗 Dataset 🤗 Models Code arXiv 🌐 Twitter Introduction --> In this work, we introduce TableLlama and TableInstruct , the first large open-source generalist model and instruction tuning dataset for table
related reading
- Parsed | Custom, interpretable AI systems that continuously learnparsed.com
- Datacurve | The data engine for frontier AIdatacurve.ai
- GitHub - brexhq/prompt-engineering: Tips and tricks for working with Large Language Models like OpenAI's GPT-4.github.com
- Fine-Tuning Llama-2: Tailoring Models to Unique Applicationsanyscale.com
- GitHub - open-compass/opencompass: OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, etc) over 100+ datasets.github.com
- Hugging Face – The AI community building the future.huggingface.co
- Scaling Instruction-Finetuned Language Modelsarxiv.org
- Tinkerthinkingmachines.ai
- Expert Data for Frontier AI - AfterQueryafterquery.com
- Putting Task Expertise into RL Achieves State-of-the-Art Performance on Text-to-SQLthinkingmachines.ai
- GitHub - Hannibal046/Awesome-LLM: Awesome-LLM: a curated list of Large Language Modelgithub.com
- Language Models can Solve Computer Tasksarxiv.org