dark

Jubayer Ibn Hamid

Profile Photo

I work in artificial intelligence and reinforcement learning.

I am currently a PhD student at Stanford University and a Student Researcher at Google. I am affiliated with Stanford Artificial Intelligence Laboratory (SAIL), where my research is advised by Dorsa Sadigh and Chelsea Finn.

Previously, I studied mathematical physics as an undergraduate at Stanford University. Outside of AI, I am interested in pure mathematics, especially abstract algebra and neighbouring fields.

Selected Papers
SPIRAL: Learning to Search and Aggregate. Jubayer Ibn Hamid*, Ifdita Hasan Orney*, Michael Y. Li, Omar Shaikh, Yoonho Lee, Dorsa Sadigh, Chelsea Finn, Noah Goodman. Preprint (ongoing work), 2026.
Poly-EPO: Training Exploratory Reasoning Models. Ifdita Hasan Orney*, Jubayer Ibn Hamid*, Shreya Ramanujam, Shirley Wu, Hengyuan Hu, Noah Goodman, Dorsa Sadigh, Chelsea Finn. Preprint (under submission), 2026.
Polychromic Objectives for Reinforcement Learning. Jubayer Ibn Hamid*, Ifdita Hasan Orney*, Ellen Xu, Chelsea Finn, Dorsa Sadigh. International Conference on Learning Representations (ICLR), 2026.

(* denotes co-lead)

All Papers
SPIRAL: Learning to Search and Aggregate. Jubayer Ibn Hamid*, Ifdita Hasan Orney*, Michael Y. Li, Omar Shaikh, Yoonho Lee, Dorsa Sadigh, Chelsea Finn, Noah Goodman. Preprint (ongoing work), 2026.
Poly-EPO: Training Exploratory Reasoning Models. Ifdita Hasan Orney*, Jubayer Ibn Hamid*, Shreya Ramanujam, Shirley Wu, Hengyuan Hu, Noah Goodman, Dorsa Sadigh, Chelsea Finn. Preprint (under submission), 2026.
Neural Garbage Collection: Learning to Forget while Learning to Reason. Michael Y. Li, Jubayer Ibn Hamid, Emily B. Fox, Noah D. Goodman. Preprint (under submission), 2026.
Polychromic Objectives for Reinforcement Learning. Jubayer Ibn Hamid*, Ifdita Hasan Orney*, Ellen Xu, Chelsea Finn, Dorsa Sadigh. International Conference on Learning Representations (ICLR), 2026.
RoboCade: Gamifying Robot Data Collection. Suvir Mirchandani*, Mia Tang*, Jiafei Duan, Jubayer Ibn Hamid, Michael Cho, Dorsa Sadigh. International Conference on Robotics and Automation (ICRA), 2026.
Bidirectional Decoding: Improving Action Chunking via Guided Test-Time Sampling. Yuejiang Liu*, Jubayer Ibn Hamid*, Annie Xie, Yoonho Lee, Max Du, Chelsea Finn. International Conference on Learning Representations (ICLR), 2025. (Website)
Tripod: Three Complementary Inductive Biases for Disentangled Representation Learning. Kyle Hsu*, Jubayer Ibn Hamid*, Kaylee Burns, Chelsea Finn, Jiajun Wu. International Conference on Machine Learning (ICML), 2024.
What Makes Pre-trained Visual Representations Successful For Robust Manipulation? Kaylee Burns, Zach Witzel, Jubayer Ibn Hamid, Tianhe Yu, Chelsea Finn, Karol Hausman. Conference on Robot Learning (CoRL), 2024.

(*) denotes co-first authorship

Notes
Set Reinforcement Learning: Principles and Approaches.
Category Theory and Algebraic Geometry.
In progress.
Rings, Modules and Categories.
In progress.
Group Theory.
Whitney's Embedding Theorem and Immersion Theorem.
Deep Reinforcement Learning.
Trust Region Optimization Methods.
Teaching
CS 224R - Deep Reinforcement Learning.
Head CA. Spring, 2025.
CS 229 - Machine Learning.
CA. Winter, 2025.