Shubhashis Roy Dipta, PhD Researcher in NLP and Multimodal AI

Shubhashis Roy Dipta

PhD Researcher · UMBC

sroydip1@umbc.edu


Amazon Science (Alexa)
Seattle, WA
Applied Scientist Intern
Summer 2026
Manager: Dr. Lichao Wang
Amazon Science (Alexa)
Seattle, WA
Applied Scientist Intern
Summer 2025
Manager: Dr. Lichao Wang
Scale AI
San Francisco, CA
Machine Learning Research Intern
Summer 2024
Manager: Dr. Adrian Lam
Mentor: Vijay Kalmath
See more
University of Maryland, Baltimore County
Ph.D. in Computer Science
Fall 2023 - Present
Grade: 4.00/4.00
Publications: See Here (From 2022)
University of Maryland, Baltimore County
M.Sc. in Computer Science
Spring 2021 - Spring 2023
Awards: Phi Kappa Phi
Grade: 4.00/4.00
Morgan State University
Research Assistant
2017 - 2019
Publications: 4 Journal
UniShopr.com
Bangladesh
Founder
2017 - 2021

Upcoming Travel

  • NeurIPS 2026 in Atlanta, GA (Dec 9-13)
  • ICLR 2027 in San Francisco, CA (Apr 26-30)
  • NAACL 2027 in San Francisco, CA (Jun 1-5)

Previous

  • ✅ ACL 2026 in San Diego, CA (Jul 3-7)
  • ✅ NeurIPS 2025 in San Diego, CA (Dec 2-7)
  • ❌ AACL 2025 in Mumbai, India (Dec 20-24) (canceled)
👋 I'm open to meet! Email me to schedule a chat!

Peer Review

Reviewed 37+ papers across top venues (2023-2027).

Conferences
ACLNeurIPSNAACLCoNLLEACLCOLING*SEM
Workshops
SemEvalTrustNLPSRWW-NUTELVM
Journals
Scientific ReportsBMC BioinformaticsPlant MethodsComputational and Structural Biotechnology

I’m a final-year CS Ph.D. researcher at the University of Maryland, Baltimore County (UMBC), advised by Dr. Frank Ferraro, with research internships at Amazon Science (Alexa AI; Summer 2025 + 2026) and Scale AI (Summer 2024). I make LLMs more reliable largely through decomposition and reinforcement learning, spanning reasoning, agentic, and multimodal settings. My work on agentic LLMs earned a $20K Google Cloud Gemini Academic Program Award (2026).

  • Reasoning & Decomposition
    • Semi-supervised RL for traceable decomposition-based claim verification [DecomposeRL]
    • Atomic, presupposition-free decomposition for robust claim verification [De-Presuppose]
    • Token-efficient math reasoning via distractor-aware computational graphs [DAGGER]
    • Curriculum-driven GRPO for math reasoning in under-resourced languages [GanitLLM]
    • Hierarchical event abstraction for compositional sequence modeling [SHEM]
  • Agentic LLMs & Reinforcement Learning
    • Tool-calling alignment via policy-grounded deliberation [PA3]
    • Multi-agent benchmarks for diagnosing collaboration failures [AgentCollabBench]
    • Mechanistic analysis of token saliency in on-policy distillation [Rock Tokens]
    • Metacognitive control in LLMs under resource constraints [TRIAGE]
  • Multimodal Learning & Evaluation
    • Reference-free factuality metric for video captions [VC-Inspector]
    • Calibrated abstention under modality conflict in omni-modal models [OMD]
    • Zero-shot multilingual text-to-video retrieval via temporal event decomposition [Q2E]

Graduating Spring 2027 · No visa sponsorship needed · actively seeking Research Scientist roles in NLP / Multimodal AI. Please reach out if you have an opening.

Recent News (See All)

Sep 6, 2026 🚀 New preprint - OracleZoom: recursive image super-resolution that keeps zooming to 256× with ground truth only at 4×. A VLM judge prefers it over Chain-of-Zoom in 78% of decided comparisons at 256×. Try the demo.
Aug 22, 2026 🎉 Four papers accepted at EMNLP 2026 - PA3: Policy-Aware Agent Alignment through Chain-of-Thought to the main conference, plus †DAGGER, Many Dialects, Many Languages, One Cultural Lens, and Register Shifts Break LLM Safety to Findings.
Jun 16, 2026 🎉 Awarded a $20,000 Google Cloud research grant through the Gemini Academic Program to support my LLM agentic research at UMBC. Featured by UMBC and on LinkedIn.
Jun 1, 2026 🎉 Joined Amazon Science (Alexa AI) for my second summer - researching self-distillation with RL to push LLM reasoning on agentic tasks.
May 27, 2026 🚀 New preprint - DecomposeRL: a 7B claim-verifier that matches GPT-4.1-mini across 11 benchmarks - with fully inspectable reasoning traces.
May 26, 2026 🥳 AgentCollabBench accepted at the FAGEN workshop @ ICML 2026 - 900 tasks that catch when a multi-agent LLM team’s final answer is right but the reasoning quietly broke.

Beyond Research

I’ve competed internationally in algorithms and robotics, ranking 8th out of 300+ teams at the 2018 ACM ICPC Asia Dhaka Regional with multiple regional and national placements, reaching the top 70 on Kaggle 🥉 in the Birdcall Identification competition, and placing 9th at the University Rover Challenge 2015 (Utah, USA) and 22nd at the European Rover Challenge 2016 (Poland). Full list of awards →

Before the PhD, I also founded UniShopr (2017-2021), a cross-border e-commerce platform serving consumers in Bangladesh.

Featured Publications

A few papers that best represent my work. See all publications

  1. OracleZoom: On-Policy Self-Distillation Inspired Reference-Constrained Recursive Image Super Resolution
    Under review
    OracleZoom: On-Policy Self-Distillation Inspired Reference-Constrained Recursive Image Super Resolution
    Shubhashis Roy Dipta*, Sourajit Saha*, Shaswati Saha, and 1 more author
    Under review
    * Equal contribution
  2. DecomposeRL: Learning to Ask Useful, Informative, and Diverse Questions for Semi-Supervised, Traceable Claim Verification
    Under review
    DecomposeRL: Learning to Ask Useful, Informative, and Diverse Questions for Semi-Supervised, Traceable Claim Verification
    Shubhashis Roy Dipta, Ankur Padia, and Francis Ferraro
    Under review
  3. PA3: Policy-Aware Agent Alignment through Chain-of-Thought
    EMNLP
    PA3: Policy-Aware Agent Alignment through Chain-of-Thought
    Shubhashis Roy Dipta, Daniel Bis, Kun Zhou, and 4 more authors
    EMNLP 2026
    Work done during internship at Amazon Alexa AI
  4. †DAGGER: Distractor-Aware Graph Generation for Executable Reasoning in Math Problems
    EMNLP
    †DAGGER: Distractor-Aware Graph Generation for Executable Reasoning in Math Problems
    Zabir Al Nazi, Shubhashis Roy Dipta, and Sudipta Kar
    EMNLP 2026
  5. GanitLLM: Difficulty-Aware Bengali Mathematical Reasoning through Curriculum-GRPO
    ACL
    GanitLLM: Difficulty-Aware Bengali Mathematical Reasoning through Curriculum-GRPO
    Shubhashis Roy Dipta, Khairul Mahbub, and Nadia Najjar
    ACL 2026
  6. VC-Inspector: Advancing Reference-free Evaluation of Video Captions with Factual Analysis
    ACL
    VC-Inspector: Advancing Reference-free Evaluation of Video Captions with Factual Analysis
    Shubhashis Roy Dipta, Tz-Ying Wu, and Subarna Tripathi
    ACL 2026
  7. Multimodal Unlearning Across Vision, Language, Video, and Audio: Survey of Methods, Datasets, and Benchmarks
    ACL
    Multimodal Unlearning Across Vision, Language, Video, and Audio: Survey of Methods, Datasets, and Benchmarks
    Nobin Sarwar, Shubhashis Roy Dipta, Zheyuan Liu, and 1 more author
    ACL 2026
  8. Q2E: Query-to-Event Decomposition for Zero-Shot Multilingual Text-to-Video Retrieval
    AACL
    Q2E: Query-to-Event Decomposition for Zero-Shot Multilingual Text-to-Video Retrieval
    Shubhashis Roy Dipta, and Francis Ferraro
    AACL 2025
  9. Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning
    MathAI @NeurIPS
    Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning
    Ningning Xu, Yuxuan Jiang, and Shubhashis Roy Dipta
    MathAI @NeurIPS 2025

Patents

  • AI-Based System and Method for Code Localization Using Dependency Graph Retrieval for Bug Identification in Software Repositories
    US provisional application filed, March 2026
  • System and Method for Anchor-Based Memory Indexing and Policy-Driven Retrieval Across Heterogeneous Sources
    US provisional application filed, March 2026