[2305.15047] Ghostbuster: Detecting Text Ghostwritten by Large Language Models
arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website. Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them. Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs. arXiv Operational Status Get status notifications via email or slack
View PDF HTML (experimental) Abstract:We introduce Ghostbuster, a state-of-the-art system for detecting AI-generated text. Our method works by passing documents through a series of weaker language models, running a structured search over possible combinations of their features, and then training a classifier on the selected features to predict whether documents are AI-generated. Crucially, Ghostbuster does not require access to token probabilities from the target model, making it useful for detecting text generated by black-box models or unknown model versions. In conjunction with our…
saved by
related reading
- Spotting LLMs With Binoculars: Zero-Shot Detection of Machine-Generated Textarxiv.org
- Pangram 4 Technical Reportpangram-public.s3.us-east-1.amazonaws.com
- GPT-4openai.com
- Pangram 4 Technical Overviewpangram.com
- Why Perplexity and Burstiness Fail to Detect AI | Pangram Labspangram.com
- gpt-4.pdfcdn.openai.com
- I don't care (anymore) if your writing got a Pangram false positiveandrewwu.substack.com
- How does Pangram work?pangram.substack.com
- GitHub - salesforce/AuditNLG: AuditNLG: Auditing Generative AI Language Modeling for Trustworthinessgithub.com
- Seeing in Pangram Space | Pangrampangram.com
- How does Pangram work?substack.com
- The Dark Forest and Generative AImaggieappleton.com