flâneur — a map of the web's best reading

Implicit planning in LLMs Paper | Manifund

manifund.org · 635 words · saved by 1 readers

You're pledging to donate if the project hits its minimum goal and gets approved. If not, your funds will be returned. The recent Claude Poetry planning results in the Anthropic Biology paper suggest that Claude is doing implicit planning when writing poetry. But Anthropic only provides a single piece of evidence for this for a single prompt. Our goal is to provide detailed and quantitative evidence to show that LLMs are doing implicit planning in poetry and also provide case studies showing that LLMs are doing implicit planning in other contexts. In prior research we found that steering vectors work well to get a model to rhyme with a specific rhyme family (e.g. causing the model to end the line with a word rhyming with "rain" instead of "quick"). We look at various metrics to measure models' ability to plan / our ability to manipulate the planning behavior: Fraction where the model ends the line in a word from the correct rhyme family (for unsteered and steered) Fraction where the mo

Implicit planning in LLMs Paper | Manifund 1 Implicit planning in LLMs Paper Technical AI safety Jim Maar Active Grant $1,000 raised $1,000 funding goal Fully funded and not currently accepting donations. What are this project's goals? How will you achieve them? The recent Claude Poetry planning results in the Anthropic Biology paper suggest that Claude is doing implicit planning when writing poetry. But Anthropic only provides a single piece of evidence for this for a single prompt. Our goal is to provide detailed and quantitative evidence to show that LLMs are doing implicit planning in poet

Explore this link on the map →

related reading