2410.12851
arxiv.org · 7,779 words · saved by 1 readers
N/A
Published as a conference paper at ICLR 2025 V IBE C HECK : D ISCOVER & Q UANTIFY Q UALITATIVE D IFFERENCES IN L ARGE L ANGUAGE M ODELS Lisa Dunlap Krishna Mandal Trevor Darrell Jacob Steinhardt Joseph Gonzalez UC Berkeley UC Berkeley UC Berkeley UC Berkeley UC Berkeley A BSTRACT…
related reading
- What We’ve Learned From A Year of Building with LLMs – Applied LLMsapplied-llms.org
- Discovering Language Model Behaviors with Model-Written Evaluations — LessWronglesswrong.com
- Emotion Concepts and their Function in a Large Language Modeltransformer-circuits.pub
- Probing Persona-Dependent Preferences in Language Modelsarxiv.org
- 2025: The year in LLMssimonwillison.net
- Things we learned about LLMs in 2024simonwillison.net
- LLM Evaluation doesn't need to be complicatedphilschmid.de
- [2303.17548] Whose Opinions Do Language Models Reflect?arxiv.org
- The bitter lesson of LLM evalsparsed.com
- [2605.21739] AttuneBench: A Conversation-Based Benchmark for LLM Emotional Intelligencearxiv.org
- [2412.00543] Evaluating the Consistency of LLM Evaluatorsarxiv.org
- GitHub - open-compass/opencompass: OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, etc) over 100+ datasets.github.com