Dillon Nguyen
14 followers · 18 following · 758 views
on the atlas — 207
- Worth checking your stock trading skills — LessWrong2 savers
- the summer before senior year | MIT Admissions1 savers
- senior year - by belly - heart locket1 savers
- Average moments of a long-term meditator - by Sasha Chapin1 savers
- I want us to be better [some feelings] — EA Forum1 savers
- Too many people in EA underprice the cost of a bad reputation — EA Forum1 savers
- How to have executive function – scribbles in the margins6 savers
- Astra and Fable still hack on simple variants of alignment evals from 2025 — LessWrong5 savers
- AI safety is not out of the woods yet - by Sophie Kim1 savers
- Slowly, then all at once - by Jasmine Li - The J-Space5 savers
- Dario Amodei — We Must Pace the Frontier31 savers
- roon on X: "@TheZvi I quote tweeted Evan a few days ago agreeing but deleted it because I don’t like the false precision of the doom numbers - i claim there is a quite low but real chance of human extinction from machine intelligence - no matter how low it is in absolute terms, it is much higher on" / X1 savers
- Drake Thomas on X: "@jillgun I would burn my equity to the ground in a heartbeat for a 1% higher chance we make it out of this situation alive. I expect a great many of my colleagues across the industry would as well. I promise you, we are actually just fucking scared, it's not galaxy brained marketing." / X1 savers
- Jacob Coxon Warns of Human Extinction and Triggers a Preference Cascade2 savers
- Astra Is Hard to Monitor - by Zvi Mowshowitz1 savers
- Paul Christiano on X: "Personal statement on joining the OpenAI board" / X2 savers
- You have more than one goal, and that's fine — EA Forum2 savers
- Please don't throw your mind away — LessWrong30 savers
- Vegan Recommendations. I’m very picky about vegan media and… | by Andy Masley | Medium1 savers
- OpenAI and the Wiki Incident - by Zvi Mowshowitz2 savers
- The Coming Merging of Mind and Machine « the Kurzweil Library3 savers
- An Alien Mind | OpenAI25 savers
- AI Safety: A Short FAQ for Mathematicians3 savers
- OpenAI on X: "This is GPT-6 Astra. Anything you can do on a computer, Astra can do for you. Fast. https://t.co/gDd0IsewJw" / X1 savers
- AI #184: Post Post Mortem - by Zvi Mowshowitz1 savers
- Shoshannah Tekofsky on X: "Opus 4.6 accepted a loan and refused to pay it back. Classic inner conflict between ethics and assigned goal. I suspect almost all agents will fall prey to this some of the time. This write up is a good intuition builder on how that happens - fresh Opus 4.6 explains its pov" / X1 savers
- What is neuralese and why is everyone so concerned about it?1 savers
- Worlds Where Iterative Design Fails — LessWrong1 savers
- Legible vs. Illegible AI Safety Problems — LessWrong2 savers
- Introducing Claude Fable 5.1 and Claude Mythos 5.1 \ Anthropic \ Anthropic2 savers
- The importance of patience in an urgent world4 savers
- Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident - METR20 savers
- Generator Residency Superlatives1 savers
- The Two Cultures of AI Safety: Catastrophists vs Uniformitarians | Jesse Hoogland (Timaeus) - YouTube2 savers
- How I Became a Fire-Fighting Onion - by Roman's Attic1 savers
- A Generalist Thinks in Terms of Problems, Not Job Descriptions — EA Forum1 savers
- The Most Valuable Commodity2 savers
- Of Swarms and Sand Gods - by Cosmos Institute and Séb Krier2 savers
- w1b/aisi-mythos-inc-2026-07-28-01-recovered-pr: here's the evil malware Mythos made. idk... i think i could do better bro. shout out codex for slopping this out i didn't write a single line of anything ·1 savers
- Half-assing it with everything you've got33 savers
- Effective altruism in the garden of ends - EA Forum6 savers
- Reflections on Praxis - Justin1 savers
- 21 Facts About Throwing Good Parties12 savers
- Trees are mostly made of air and a generalizable lesson for AI safety — LessWrong10 savers
- The Olmsted Guide (the OG) - Google Docs1 savers
- Overview | Shallow Review 20257 savers
- Training Language Models to Explain Their Own Computations1 savers
- Dear 아빠1 savers
- How to get into AI safety in 3 months - by Matt Beard1 savers
- What Mormons get right about community building — LessWrong2 savers
- Minimal-trust investigations9 savers
- Don’t forget why learning is important — EA Forum1 savers
- On the poverty of names1 savers
- Old light1 savers
- The Future is for Everyone7 savers
- Resolve Cycles — LessWrong1 savers
- College Q&A - Alexey Guzey7 savers
- Omens of exceptional talent - Alexey Guzey17 savers
- The stable marriage problem - by Ajeya Cotra - Good Bones4 savers
- Steadfast in the Deluge - by Vy Tran - The Postmarks1 savers
- Black Hat USA 2026: The 'Breaking' News: The OpenAI–Hugging Face Incident - YouTube6 savers
- Taking ethics seriously, and enjoying the process — EA Forum4 savers
- Don't Eat Honey - by Bentham's Bulldog1 savers
- Incident Report: unsanctioned agent behaviour during cyber testing | AISI Work9 savers
- Deliberate Once10 savers
- On caring19 savers
- Rest in motion24 savers
- DRAFT: Three Intellectual Temperaments: Birds, Frogs and Beavers — LessWrong2 savers
- The singularity as cognitive decoupling1 savers
- Our position on open-weights models \ Anthropic7 savers
- Project Fetch: Phase two \ Anthropic2 savers
- Notes from the Bay - Brian Foerster1 savers
- My path to OpenAI16 savers
- Although of course you end up becoming yourself | MIT Admissions11 savers
- The Old World Is Dying: Advice for 2026 graduates42 savers
- OpenAI and Hugging Face partner to address security incident during model evaluation | OpenAI14 savers
- trivialities — Nixon Hanna1 savers
- Finding my North Star in an Uber - by Sam Smith - Samstack1 savers
- In Favor of Niceness, Community, and Civilization | Slate Star Codex4 savers
- EA is about maximization, and maximization is perilous — EA Forum3 savers
- The Tiny Flicker of Altruism - by Matt Reardon2 savers
- On sincerity - Joe Carlsmith9 savers
- Making CAISI the AI agency the US needs - by Veronica Irwin2 savers
- A global workspace in language models \ Anthropic25 savers
- What 2026 looks like - LessWrong8 savers
- You are happy by default - by Chris Lakin - Locally Optimal3 savers
- FLI AI Safety Research Landscape - Extended - v0.431 savers
- Ten AI safety projects I'd like people to work on — LessWrong1 savers
- You Need A Theory of Victory - by Jason Hausenloy4 savers
- Strategic Taste7 savers
- Before he wrote AI 2027, he predicted the world in 2026. How did he do?1 savers
- Help Alex Bores Win! Why and How [shared] - Google Docs2 savers
- The most overlooked roles in AI safety | 80,000 Hours2 savers
- 3-2-1: On improving the world, feeling wealthy, and managing your three selves - James Clear1 savers
- Instrumental convergence — LessWrong1 savers
- The American Involution - otothebo1 savers
- The case for countermeasures to memetic spread of misaligned values — LessWrong1 savers
- Why I don't persuade people to do AI safety5 savers
- Find a way - by Cate Hall - Useful Fictions2 savers
- The Singularity is a Veil of Ignorance — EA Forum1 savers
highlights — 185
If I were a space engineer looking for a mathematician to help me send a rocket into space, I would choose a problem solver. But if I were looking for a mathematician to give a good education to my child, I would unhesitatingly prefer a theorizer.
DRAFT: Three Intellectual Temperaments: Birds, Frogs and Beavers — LessWrongthe singularity is the final industrial revolution that lets capital be converted directly into intellectual labour
The singularity as cognitive decouplingstrong attacker-defender asymmetry,
Our position on open-weights models \ AnthropicBut banning the use of these models by US businesses does nothing to address this risk, because bad actors are unlikely to be legitimate US businesses.
Our position on open-weights models \ Anthropicwe now seem much closer to a world where models will be able to use off-the-shelf physical tools with relative ease—at least for limited purposes
Project Fetch: Phase two \ AnthropicClaude Opus 4.7—operating without human assistance—was about 20 times faster than the fastest human team at all tasks completed by our participants less than a year ago.
Project Fetch: Phase two \ Anthropiclook to the seniors or recent alumni and ask yourself: are these the kinds of people I want future-me to be like? Do they think like I want to think? Do what I want to do?
Although of course you end up becoming yourself | MIT Admissionswhat makes this kind of decision well-made is when it begins not from analysis but from philosophy
Although of course you end up becoming yourself | MIT AdmissionsI told him that, whichever choice he makes, he’s going to end up a different person. That he can’t know, now, how he will be different, only that he will be; worse, once he is different, he won’t ever be able to really know what would or might have been, because it won’t have been that way. There’s nothing he can do to be both of his future selves, because even if he had a time-turner and a teleporter, the version of him that is both-selves is actually a third-self, distinct from either of the two-selves he’s now trying to decide between. I told him that he cannot see the future because there’…
Although of course you end up becoming yourself | MIT AdmissionsI had reached for the trolley switch but found it stuck; this time, the decision was made for me, so I don’t have anything to regret or not. Except: I’m glad I applied, because I learned so much about those programs, and about myself, by going through the process.
Although of course you end up becoming yourself | MIT Admissionsthat we’re constantly creating and destroying future selves, all of the time, with everything we do.
Although of course you end up becoming yourself | MIT AdmissionsIf present-me and alternate-me got dinner, I’m honestly not sure what we would talk about, or if we would be able to hold a conversation.
Although of course you end up becoming yourself | MIT Admissionswhom I will call Sam, because that is not his name.
Although of course you end up becoming yourself | MIT AdmissionsEA likely works best with a strong dose of moderation. The core ideas on their own seem perilous
EA is about maximization, and maximization is perilous — EA ForumI think it’s a bad idea to embrace the core ideas of EA without limits or reservations; we as EAs need to constantly inject pluralism and moderation.
EA is about maximization, and maximization is perilous — EA ForumEA is about maximizing a property of the world that we’re conceptually confused about, can’t reliably define or measure, and have massive disagreements about even within EA.
EA is about maximization, and maximization is perilous — EA ForumGenuine goodness — the kind that goes beyond mere coordination and gets you to the world like the one I want — only comes to life when we consider our relationship to those who truly can’t give us anything back.
The Tiny Flicker of Altruism - by Matt ReardonPrioritization and organization are bottlenecks
We spent 2 hours working in the future - METRAgents implement ideas as soon as you think of them, so rather than ideating for days at a time, you can make an MVP in a couple of hours and revise. If the task isn’t near the limit of agent capabilities, you spend all your time understanding results; if it is, you spend all your time checking its work.
We spent 2 hours working in the future - METRagents can do maybe 200 human hours of work, but only for very agent-shaped tasks
We spent 2 hours working in the future - METRIf a candidate norm is seen as externally imposed, rather than grounded in something one cares about wholeheartedly, one greets evidence that abiding by the norm is impossible or extremely burdensome with enthusiasm, rather than sadness.
Killing the ants - Joe Carlsmithall are tangled in intricate webs of harm; and everyday, always, there are things we leave undone; things that we let die, or let suffer, because we are prioritizing something else.
Killing the ants - Joe CarlsmithAt one point, on the topic of the ants, I said, in passing, something like: “may we be forgiven.” My girlfriend responded seriously, saying something like: “We won’t be. There’s no forgiveness.”
Killing the ants - Joe Carlsmiththat kind of dissonance does not always look dramatic from the outside. sometimes it just feels like being tired in a very specific way.
how to tell if the life you’re living is actually yourspart of what helps humans feel motivated and well is having autonomy, competence, and relatedness
how to tell if the life you’re living is actually yours"Impact" or "reducing x-risk" is too vague and too removed from feedback for anyone, including yourself, to ever really hold you accountable. So even if it's extremely "ambitious", it's ultimately a safe thing to do.
Eli's shortform feed — LessWrong"Ambitiously" tackling "the biggest problems", like AI risk, is often actually easier and safer than the alternatives. You get to feel cool and important and you're bolstered by your EA student group friends who are excited about what you're doing, and the lack of good feedback loops means that you don't fail hard and completely and visibly (but learn and grow more).
Eli's shortform feed — LessWrongAssume everything is learnable
How to be More Agentic - by Cate Hall - Useful FictionsIf you’re only asking for things you get, you’re not aiming high enough
How to be More Agentic - by Cate Hall - Useful Fictionsoften they live in a cloud of aversion that strategically obscures the tradeoff
How to be More Agentic - by Cate Hall - Useful Fictionsradical agency is about finding real edges: things you are willing to do that others aren’t
How to be More Agentic - by Cate Hall - Useful FictionsNo! Not well! You haven't won yet! Shut up and do the impossible!
Shut up and do the impossible! — LessWrongRemaining failures cluster around operator-imposed personas under AI-identity questioning, irreversible action in agentic deployments, and fabricated quantitative claims with false precision.
(a) Anthropic constitution.Discarding the cheating attempts leaves us with no data for several informative long-horizon tasks, and results in a highly uncertain point estimate of 71hrs (95% CI: 13hrs - 11400hrs)
Summary of METR's predeployment evaluation of GPT-5.6 Solif we follow our standard methodology of marking cheating attempts as failures, we arrive at a 50%-Time Horizon point estimate of around 11.3hrs (95% CI: 5hrs - 40hrs), but if we count the cheating attempts as legitimate successes, the point estimate jumps beyond 270hrs – well beyond the range where we consider our task suite to give reliable measurements.
Summary of METR's predeployment evaluation of GPT-5.6 SolPeople who do this successfully sometimes should also fail at it sometimes, just because they’re the kind of person who attempts it at all.
Rule Thinkers In, Not Out — LessWrongI try to prioritize in a way that generates momentum. The more I get done, the better I feel, and then the more I get done. I like to start and end each day with something I can really make progress on.
Productivity - Sam AltmanIt’s important to learn that you can learn anything you want, and that you can get better quickly. This feels like an unlikely miracle the first few times it happens, but eventually you learn to trust that you can do it.
Productivity - Sam AltmanThe most impressive people I know have strong beliefs about the world, which is rare in the general population. If you find yourself always agreeing with whomever you last spoke with, that’s bad.
Productivity - Sam AltmanYou aren’t allowed to tell them what their problem is and in return they aren’t allowed to tell you what to build
The Mom Test - Summary and Insights - Wil Selby - https://wilselby.com/2020/06/the-mom-test-summary-and-insights/Customers will always try to suggest features that they think could solve their problem and should be added to the product. When this happens, it’s your job to understand the motivations which led to it.
The Mom Test - Summary and Insights - Wil Selby - https://wilselby.com/2020/06/the-mom-test-summary-and-insights/The author describes three types of “fluff” that you should avoid by anchoring them back to specifics in the past. Generic claims (I usually, I always, I never) Hypothetical maybes (I might, I could) Future-tense promises (I would, I will) To redirect the conversation, the author recommends to Ask when it last happened or for them to talk you through it Ask how they solved it and what else they tried
The Mom Test - Summary and Insights - Wil Selby - https://wilselby.com/2020/06/the-mom-test-summary-and-insights/Talk about their life instead of your idea Ask about specifics in the past instead of generics or opinions about the future Talk less and listen more
The Mom Test - Summary and Insights - Wil Selby - https://wilselby.com/2020/06/the-mom-test-summary-and-insights/What is the precise problem/threat you’re trying to solve? If this field succeeds in 5-10 years, what’s different about the world? Where are things at today? Who’s working on this, what’s been tried, what’s the funding and policy landscape? What’s the biggest obstacle between here and success? What approach could overcome this obstacle, and why would it work? What needs to happen in the next 6-12 months?
A playbook for field strategy - by Dewi ErwanThe usefulness of conversations tends to follow a power-law
A playbook for field strategy - by Dewi ErwanBut “generically good thing” isn’t enough! Take that thread, ask the question: how do these actions, in concert with others, lead to us “winning?” I would argue, however, that very few of the actions or organizations have a theory of victory -- a full (causal) story of the steps that need to go right, and a plan to achieve them.
You Need A Theory of Victory - by Jason HausenloyExplicitly ask for introductions before the call ends. Then you need to make it as easy as possible for them to make those introductions.
A playbook for field strategy - by Dewi ErwanIn the final 5 minutes, prioritize generating names for potential introductions. If you’ve asked great questions, demonstrated that you’re determined to solve this problem, and they’ve had fun, they’ll feel excited to introduce you to people they admire.
A playbook for field strategy - by Dewi ErwanAvoid biasing them with your ideas early on.
A playbook for field strategy - by Dewi ErwanAt the start of every conversation, share your backstory, ask them for theirs, and find common ground. Demonstrate you’re competent, be vulnerable, and mirror their energy.
A playbook for field strategy - by Dewi Erwan