flâneur

LMCA_dataset.pdf

andrew.cmu.edu · 7,838 words · saved by 1 readers

N/A

A dataset of rated conceptual arguments Emery Cooper* Caspar Oesterheld* Linh Chi Nguyen Alexander Kastner Ethan Perez 1 Introduction In recent years, the capabilities of large language models (LLMs) have progressed impressively across a wide range of domains. Over the past year or so, so-called reasoning models have made rapid progress on tasks with verifiable feedback, such as coding and math (Jaech et al. 2024; Guo et al. 2025; A. Yang et al. 2025). In light of concerns about the risks of AI development (such as misalignment…

saved by

related reading