TikTok · Machine Learning Engineer Intern · May – Aug 2026
LLM / VLM AI agent RL post-training Reward design Distributed training E-commerce Query recommendation
Schematic of the serving and training pipeline. Internal work; no public paper.
Internal project — no public release
Xpeng Motors · Machine Learning Intern · May 2025 – Jan 2026
LLM / VLM Autonomous driving Vision-language-action BEV representation End-to-end driving Distributed training
Two-stage training: pretrain on rendered nuPlan planning data, then fine-tune on nuScenes with BEV maps from a frozen perception model.
Manuscript under review
Amazon · Applied Scientist Intern · May – Aug 2024
LLM / VLM AI agent Web agents Context compression
A shared history compressor distills each verbose past web state into a fixed-length representation before the action-prediction transformer.
UIUC · Dissertation chapter · Jan 2025 – Jun 2026
Video generation / understanding AI agent LLM / VLM Diffusion models / visual generative models RL post-training Reward design Chain-of-thought planning
The generator and its metrics are treated as an environment; language agents plan, inspect, regenerate, and roll back around a frozen video diffusion model.
UIUC · Jun 2025 – Jan 2026
AI for healthcare LLM / VLM AI agent RL post-training Reward design Report generation LLM-as-judge
Figure 2 of the paper: the MedQPA evaluation loop (question proposing and answering), reflective prompting, and the RL update against the MedQPA reward model.
UIUC · Nov 2023 – Feb 2025
Diffusion models / visual generative models AI for healthcare Video generation / understanding MRI super-resolution Inverse problems 3D representation
At each denoising step, perpendicularly trained 2D diffusion models give initial estimates and a lightweight 3D network learns to fuse them in score space; alignment modules inject hierarchical 2D features.
UIUC · Jul 2022 – Aug 2023
Autonomous driving Diffusion models / visual generative models BEV map segmentation Generative priors
An off-the-shelf perception model gives a noisy estimate; a generative encoder, transformer sampler, and decoder turn it into realistic, diverse map layouts.
UIUC · Jan – Oct 2022
Autonomous driving Diffusion models / visual generative models LiDAR generation Score-based models
Point clouds are generated by score-based denoising in the equirectangular range/intensity view.