About

I am currently a Master’s student in Machine Learning at MBZUAI, advised by Dr. Junpei Komiyama. My research spans the training and reasoning of foundation models, including pretraining and post-training dynamics, reasoning, and multi-agent systems. I am particularly interested in understanding how reasoning abilities emerge through training and designing methods that make them more reliable, robust, and aligned.

Previously, I worked on latent reasoning in vision-language models with Dr. Biwei Huang at UC San Diego and on inference-time alignment of language models with Dr. Berkay Celik at Purdue University. I also had the opportunity to collaborate with Dr. Ferdinando Fioretto at the University of Virginia on understanding misalignment in self-evolving coding agents during iterative self-improvement.

News

Aug 2026 Our new work on Prefix-Denoising Consistency is now up on arXiv!
Aug 2026 Started MSc in Machine Learning at MBZUAI!
Aug 2026 Our new work on On-Policy Self-Distillation is now up on arXiv!
May 2026 STARS was accepted at the SPIGM Workshop at ICML 2026!
May 2026 Our new work on Prefix-Consistency is now up on arXiv!
Feb 2026 Joined MBZUAI as a Visiting Researcher to work on reasoning in language models!
Oct 2025 Served on the Program Committee for NeurIPS 2025 Workshops!
Oct 2025 Two papers accepted at NeurIPS workshops — Efficient Reasoning and FM4LS!
Aug 2025 Served on the Program Committee for AAAI 2026 conference!

Publications * indicates equal contribution.

  1. Adaptive Blockwise Search: Inference-Time Alignment for Large Language Models
    M. Atif Quamar* , M. Areeb* , N. Sharma , A. Shreekumar , J. Rosenthal , M. Kuznetsov , M. Ozgur Ozmen , and Z. Berkay Celik
    Under Review
  2. ICML-W 2026
    stars.png
    STARS: Synchronous Token Alignment for Robust Supervision in Large Language Models
    M. Atif Quamar* , M. Areeb* , M. Kuznetsov , M. Ozgur Ozmen , and Z. Berkay Celik
    ICML 2026 - Structured Probabilistic Inference & Generative Modeling Workshop
  3. NeurIPS-W 2025
    entropy.png
    Logit–Entropy Adaptive Stopping Heuristic for Efficient Chain-of-Thought Reasoning
    M. Atif Quamar and M. Areeb
    NeurIPS 2025 - Efficient Reasoning Workshop
  4. Reliable Chain-of-Thought via Prefix Consistency
    N. Iwase , Y. Ichihara , M. Atif Quamar , and J. Komiyama
    Under Review
  5. Privileged Solutions or Context-Induced Teacher Behavior? Dissecting On-Policy Self-Distillation
    Y. Ichihara , N. Iwase , M. Atif Quamar , and J. Komiyama
    Under Review
  6. Prefix-Denoising Consistency: Test-Time Verification for Diffusion Language Models
    Y. Ichihara , N. Iwase , M. Atif Quamar , and J. Komiyama
    Under Review
  7. Learning Modal-Mixed Chain-of-Thought Reasoning with Latent Embeddings
    Y. Shao , K. Zhou , Z. Xu , M. Atif Quamar , S. Hao , Z. Wang , Z. Hu , and B. Huang
    Under Review
  8. NeurIPS-W 2025
    histone.png
    Decoding Histone Modification Signatures of Non-Coding RNAs via Foundation Models
    N. Sharma , M. Atif Quamar , and P. Xie
    NeurIPS 2025 - FM4LS Workshop