LLM-Mod: Can Large Language Models Assist Content Moderation?
koustuv.com · 6,801 words · saved by 1 readers
N/A
LLM-Mod: Can Large Language Models Assist Content Moderation? Mahi Kolla∗ Siddharth Salunkhe∗ mkolla2@illinois.edu ss190@illinois.edu University of Illinois Urbana-Champaign University of Illinois Urbana-Champaign Urbana, Illinois, USA…
saved by
related reading
- 2025.naacl-long.441.pdfaclanthology.org
- Legilimens: Practical and Unified Content Moderation forLarge Language Model Servicesarxiv.org
- RigorLLM: Resilient Guardrails for Large Language Models against Undesired Contentarxiv.org
- 2312.06674arxiv.org
- What We’ve Learned From A Year of Building with LLMs – Applied LLMsapplied-llms.org
- Productizing Large Language Modelsblog.replit.com
- Debating the role of large language models in the kernel communitylwn.net
- Little Free Language Modelswritefutureslab.com
- Discovering Language Model Behaviors with Model-Written Evaluations — LessWronglesswrong.com
- Things we learned about LLMs in 2024simonwillison.net
- Language Models Learn to Mislead Humans via RLHFarxiv.org
- LLM-as-a-judge: a complete guide to using LLMs for evaluationsevidentlyai.com