Calibration Is Not Control
Knowing that an agent might fail does not tell us whether intervening will help. We study intervention advantage through action-conditioned decisions and counterfactual branching from the same trajectory prefix.
Learning through interaction.
I work on learning and decision-making, with recent projects in online reinforcement learning and generative policies.
01 / About
I am Chubin Zhang (张楚彬), currently a Research Assistant in Prof. Bo An's group at Nanyang Technological University.
My recent work has explored online reinforcement learning and expressive generative policies. More broadly, I am interested in how learning systems make decisions and adapt through interaction.
For research conversations, reach me at z1175085991@gmail.com.
02 / News
03 / Research
Knowing that an agent might fail does not tell us whether intervening will help. We study intervention advantage through action-conditioned decisions and counterfactual branching from the same trajectory prefix.
04 / Publications
NeurIPS 2026
NeurIPS 2026
ICML 2026
ICML 2026
AAAI 2026
ICML 2026 Workshop on SPIGM
ICLR 2026 Workshop on ReALM-GEN
arXiv preprint, 2026
05 / Background
My research background includes online reinforcement learning, generative modeling, and language-model research.
Browse publicationsResearch Assistant in Prof. Bo An's group / Singapore
Research Intern with Prof. Ivor Tsang / Singapore
Research Intern with Prof. Zhiyuan Liu / Beijing
06 / Contact
For research conversations, collaborations, or opportunities, reach me by email.