✳flâneur — a map of the web's best reading
Group | Sherry Tongshuang Wu
cs.cmu.edu · 2,083 words · saved by 1 readers
Hello World!
Group | Sherry Tongshuang Wu Sherry @ CMU CV Publications About · Publications · Teaching & Mentoring · Teaching & Mentoring · --> Design practical AIs that can help users in complex tasks, where users are not oracle, and not static. Real-world AI Evaluation We ask What can general-purpose models do? We do Replicate diverse human-subject experiments with general-purpose models. User-Centered AI Instruction Following We ask How can AI effectively fulfill true user needs? We do Perform task-specific model testing and distillation. Design measurements that quantify use
Explore this link on the map →saved by
related reading
- Explore | alphaXivalphaxiv.org
- LLM Powered Autonomous Agents | Lil'Loglilianweng.github.io
- Explore | alphaXivalphaxiv.org
- Patterns for Building LLM-based Systems & Productseugeneyan.com
- AI in 2025: gestalt — LessWronglesswrong.com
- Composer2.pdfcursor.com
- The 2025 AI Engineering Reading List - Latent.Spacelatent.space
- GenAI Handbookgenai-handbook.github.io
- Language Models can Solve Computer Tasksarxiv.org
- Building Effective AI Agents \ Anthropicanthropic.com
- Your AI Product Needs Evals – Hamel's Blog - Hamel Husainhamel.dev
- 2025: The year in LLMssimonwillison.net