Agentic & Multimodal Reasoning
Reasoning, grounding, and execution scaffolds for multimodal and agentic systems.
“Talk is cheap. Show me the code.” — Linus Torvalds
I am an AI Researcher
I am a third-year Ph.D. student in Computer Science at the University of Alabama at Birmingham, co-advised by Prof. Tianyang Wang and Prof. Min Xu. I'm also a student researcher at Oak Ridge National Laboratory working with Dr. Xiao Wang since May 2025.
My earlier research focused on understanding and shaping the intelligence and behavior of foundation models through post-training techniques. I'm currently extending this focus to agents, exploring how reinforcement learning (RL) and other post-training techniques can improve their capabilities and behavior, with a particular focus on agent memory, harness engineering, recursive self-improvement (RSI), and reasoning.
Reasoning, grounding, and execution scaffolds for multimodal and agentic systems.
Parameter-efficient adaptation, prompt discovery, and personalized vision-language models.
Representation geometry, alignment objectives, and post-training signals for reliable intelligence.
Efficient generation, dataset distillation, weather modeling, and scalable vision systems.
Talks & Media
Academic Service