Abhijeet Sinha
PhD researcher at the National University of Singapore working on reinforcement learning and generative AI.
Abhijeet Sinha
National University of Singapore
Singapore
I am a PhD researcher in Machine Learning at the National University of Singapore, where I work in the Artificial Scientific Intelligence Lab under the supervision of Dr. Dianbo Liu. My research focuses on reinforcement learning, generative modeling, LLM post-training and alignment, diversity, creativity, and interpretable AI.
My current work studies outcome-level mode collapse in reinforcement learning and reward-proportional sampling. I develop methods that encourage diverse, high-reward generation while preserving alignment and task performance. I also work on interpretable generative systems, including reinforcement-learning agents that edit discrete latent representations.
Previously, I completed a dual Bachelor’s and Master’s degree in Biotechnology at IIT Madras and worked on brain-inspired visual attention, neurological healthcare applications, document understanding, and enterprise AI systems.
selected publications
- AAAIDensity-Aware Reward Scaling: A Reinforcement Learning Framework for Reward-Proportional SamplingSubmitted to AAAI 2027