Agentic & Multimodal Reasoning
Reasoning, grounding, and execution scaffolds for multimodal and agentic systems.

“Talk is cheap. Show me the code.” — Linus Torvalds
I am an AI Researcher
I am a third-year Ph.D. student in Computer Science at the University of Alabama at Birmingham, co-advised by Prof. Tianyang Wang and Prof. Min Xu. I'm also a student researcher at Oak Ridge National Laboratory working with Dr. Xiao Wang since May 2025.
My research focuses on post-training techniques (e.g., reinforcement learning) for LLMs and MLLMs. I am currently working on two directions: (i) alignment and reasoning for LLMs and MLLMs, and (ii) long-horizon LLM agents capable of reliable planning, interaction, and adaptation.
Reasoning, grounding, and execution scaffolds for multimodal and agentic systems.
Parameter-efficient adaptation, prompt discovery, and personalized vision-language models.
Representation geometry, alignment objectives, and post-training signals for reliable intelligence.
Efficient generation, dataset distillation, weather modeling, and scalable vision systems.
Talks & Media
Academic Service