flâneur

Jubayer Ibn Hamid

jubayer-ibn-hamid.github.io · 169 words · saved by 1 readers

I work in artificial intelligence and reinforcement learning. I am currently a PhD student at Stanford University and a Student Researcher at Google DeepMind. I am affiliated with Stanford Artificial Intelligence Laboratory (SAIL), where my research is advised by Dorsa Sadigh and Chelsea Finn. Previously, I studied mathematical physics as an undergraduate at Stanford University. Outside of AI, I am interested in pure mathematics, especially abstract algebra and neighbouring fields. I work on reinforcement learning and sequential decision-making with the goal of developing learning and inference algorithms aiming towards solving extremely complex problems, especially in scientific discovery. These days, I am interested in exploration-exploitation strategies, and training-time and inference-time search. Selected Papers: SPIRAL: Learning to Search and Aggregate. Jubayer Ibn Hamid*, Ifdita Hasan Orney*, Michael Y. Li, Omar Shaikh, Yoonho Lee, Dorsa Sadigh, Chelsea Finn, Noah Goodman. Prepr

SPIRAL: Learning to Search and Aggregate. Jubayer Ibn Hamid*, Ifdita Hasan Orney*, Michael Y. Li, Omar Shaikh, Yoonho Lee, Dorsa Sadigh, Chelsea Finn, Noah Goodman. Poly-EPO: Training Exploratory Reasoning Models. Ifdita Hasan Orney*, Jubayer Ibn Hamid*, Shreya Ramanujam, Shirley Wu, Hengyuan Hu, Noah Goodman, Dorsa Sadigh, Chelsea Finn. Neural Garbage Collection: Learning to Forget while Learning to Reason. Michael Y. Li, Jubayer Ibn Hamid, Emily B. Fox, Noah D. Goodman. Polychromic Objectives for Reinforcement Learning. Jubayer Ibn Hamid*, Ifdita Hasan Orney*, Ellen Xu, Chelsea Finn,…

saved by

related reading