Aviral Kumar on X: "🚨🚨 New paper on flow-matching value functions Last year, we showed training RL value functions with a flow-matching loss achieved SOTA results. But why does it work? And what could it possibly tell us about other things that have nothing to do with VFs or even RL? Short https://t.co/x5trTotUO4" / X
To view keyboard shortcuts, press question mark View keyboard shortcuts Home Explore 1 Notifications Chat Grok Premium Bookmarks Creator Studio Articles Profile More Post sean lee @infinitefun_ Post See new posts Conversation Aviral Kumar @aviral_kumar2 New paper on flow-matching value functions Last year, we showed training RL value functions with a flow-matching loss achieved SOTA results. But why does it work? And what could it possibly tell us about other things that have nothing to do with VFs or even RL? Short answer: iterative compute used correctly can address feature plasticity in continual learning! 1:38 PM · Mar 5, 2026 · 59.3K Views 3 79 481 397 Relevant View quotes Post your reply Reply Aviral Kumar @aviral_kumar2 · Mar 5 Background: last year we developed floq https:// x.com/aviral_kumar2/ status/1965396309377966490 …. floq uses flow matching to make scalar predictions. Key idea: Transform state/action scalar value with an integration process that integrate
@aviral_kumar2: New paper on flow-matching value functions Last year, we showed training RL value functions with a flow-matching loss achieved SOTA results. But why does it work? And what could it possibly tell us about other things that have nothing to do with VFs or even RL? Short answer: iterative compute used correctly can address feature plasticity in continual learning! @aviral_kumar2: Background: last year we developed floq https:// x.com/aviral_kumar2/ status/1965396309377966490 …. floq uses flow matching to make scalar predictions. Key idea: Transform state/action scalar v
Explore this link on the map →saved by
related reading
- RL without TD learningseohong.me
- Flow Matching Policy Gradientsflowreinforce.github.io
- Visualizing Flow Matching in Roboticsabhinavjha.xyz
- LoRA Without Regret - Thinking Machines Labthinkingmachines.ai
- [2506.03719] On the Closed-Form of Flow Matching: Generalization Does Not Arise from Target Stochasticityarxiv.org
- pistar06.pdfpi.website
- A (Long) Peek into Reinforcement Learning | Lil'Loglilianweng.github.io
- Learning the integral of a diffusion model – Sander Dielemansander.ai
- Diffusion Meets Flow Matchingdiffusionflow.github.io
- π*0.6: a VLA That Learns From Experiencephysicalintelligence.company
- Finite Difference Flow Optimizationmcallisterdavid.com
- Composer2.pdfcursor.com