Ph.D. Student in Computer Science, Purdue University
I'm a PhD student specializing in applied reinforcement learning in planning applications from path planning, multi-agent navigation, and large language models with spatial reasoning capabilities. Also, I enjoy building tools and interactive systems that translate technology into meaningful products, spanning HCI research and independent side projects.
Proposed a two-stage framework that decomposes spatial reasoning into atomic building blocks and their composition. Applied supervised fine-tuning on elementary spatial transformations (rotation, translation, scaling), then trained lightweight LoRA adapters with GRPO to enable closed-loop, multi-step planning in puzzle-based environments.
Developing an interactive, story-based language-learning game using GPT that helps second-language learners practice slang and conversational English through spoken dialogue and reflection tasks. Designed as part of a study evaluating vocabulary retention among international students at Purdue University.
Collaborating with Lightspeed Studios on scalable traffic-flow simulation. Integrating Graph Neural Networks (GNN) with Multi-Agent PPO to derive environment-level coordination rules for complex multi-agent systems.
Path planning with Value Iteration and Random-based methods. Improved performance in complex maps by combining reinforcement learning with rapidly-exploring random trees.
Manage course logistics for 369 students, including lab instruction, exam coordination, and TA scheduling. Lead teaching assistants in developing instructional materials and ensuring smooth course operations.
Product Life-cycle, PRD, RoadMaps Tools (Roadmunk), Work-flows (Waterfall, Scrum, Kanban).
Created content on the X platform, increasing impressions from ~100 to 10k within six months.
Teaching assistants in developing instructional materials and ensuring smooth course operations.
Led logistics and communication efforts for academic events and visits, coordinating between professors and students to organize departmental gatherings.
Reinforcement Learning (PPO), LLMs (SFT, GRPO), Data Analysis, GNNs
User Research, Rapid Prototyping, A/B Testing, Market Analysis, Content Creation
Python (PyTorch, Pandas), JavaScript, C++, MATLAB, SolidWorks