About

I am currently a Master’s student in Machine Learning at MBZUAI, advised by Dr. Junpei Komiyama. My research lies at the intersection of foundation model training and mechanistic interpretability, focusing on pretraining and post-training dynamics, machine reasoning, and multi-agent systems. I am particularly interested in how reasoning and internal mechanisms emerge during training, and how understanding them can lead to more reliable, robust, and aligned models.

Previously, I worked on latent reasoning in vision-language models with Dr. Biwei Huang at UC San Diego and on inference-time alignment of language models with Dr. Berkay Celik at Purdue University. I also had the opportunity to collaborate with Dr. Ferdinando Fioretto at the University of Virginia on understanding misalignment in self-evolving coding agents during iterative self-improvement.

News

Sep 2026 Our work on Prefix-Consistency was accepted to NeurIPS 2026!
Sep 2026 Attended the ELLIS Summer School at TU Munich, Germany!
Aug 2026 Our new work on Prefix-Denoising Consistency is now up on arXiv!
Aug 2026 Started MSc in Machine Learning at MBZUAI!
Aug 2026 Our new work on On-Policy Self-Distillation is now up on arXiv!
May 2026 STARS was accepted at the SPIGM Workshop at ICML 2026!
May 2026 Our new work on Prefix-Consistency is now up on arXiv!
Feb 2026 Joined MBZUAI as a Visiting Researcher to work on reasoning in language models!
Oct 2025 Served on the Program Committee for NeurIPS 2025 Workshops!
Oct 2025 Two papers accepted at NeurIPS workshops — Efficient Reasoning and FM4LS!
Aug 2025 Served on the Program Committee for AAAI 2026!

Publications * indicates equal contribution.

  1. NeurIPS 2026
    Reliable Chain-of-Thought via Prefix Consistency
    Naoto Iwase , Yuki Ichihara , Mohammad Atif Quamar , and Junpei Komiyama
    Advances in Neural Information Processing Systems (NeurIPS) 2026
  2. Preprint
    Adaptive Blockwise Search: Inference-Time Alignment for Large Language Models
    Mohammad Atif Quamar* , Mohammad Areeb* , Nishant Sharma , Ananth Shreekumar , Jonathan Rosenthal , Muslum Ozgur Ozmen , Mikhail Kuznetsov , and Z. Berkay Celik
    arXiv preprint, 2025
  3. NeurIPS-W 2025
    Logit–Entropy Adaptive Stopping Heuristic for Efficient Chain-of-Thought Reasoning
    Mohammad Atif Quamar and Mohammad Areeb
    NeurIPS 2025 Workshop on Efficient Reasoning
  4. ICML-W 2026
    STARS: Synchronous Token Alignment for Robust Supervision in Large Language Models
    Mohammad Atif Quamar* , Mohammad Areeb* , Mikhail Kuznetsov , Muslum Ozgur Ozmen , and Z. Berkay Celik
    ICML 2026 Workshop on Structured Probabilistic Inference & Generative Modeling
  5. Preprint
    Privileged Solutions or Context-Induced Teacher Behavior? Dissecting On-Policy Self-Distillation
    Yuki Ichihara , Naoto Iwase , Mohammad Atif Quamar , and Junpei Komiyama
    arXiv preprint, 2026
  6. Preprint
    Prefix-Denoising Consistency: Test-Time Verification for Diffusion Language Models
    Yuki Ichihara , Naoto Iwase , Mohammad Atif Quamar , and Junpei Komiyama
    arXiv preprint, 2026
  7. Preprint
    Learning Modal-Mixed Chain-of-Thought Reasoning with Latent Embeddings
    Yifei Shao , Kun Zhou , Ziming Xu , Mohammad Atif Quamar , Shibo Hao , Zhen Wang , Zhiting Hu , and Biwei Huang
    arXiv preprint, 2026
  8. NeurIPS-W 2025
    Decoding Histone Modification Signatures of Non-Coding RNAs via Foundation Models
    Nishant Sharma , Mohammad Atif Quamar , and Pengtao Xie
    NeurIPS 2025 2nd Workshop on Multi-modal Foundation Models and Large Language Models for Life Sciences