Visual Servoing with Residual Reinforcement Learning
Combining classical IBVS with a PPO residual policy for robust eye-in-hand manipulation on a Franka Panda robot.
Combining classical IBVS with a PPO residual policy for robust eye-in-hand manipulation on a Franka Panda robot.
A production-grade pipeline for generating photorealistic lip-synced face videos from arbitrary speech audio, combining SyncNet with a U-Net generator.
Fine-tuning LLaMA with LoRA to convert free-text medical abstracts into structured JSON — covering synthetic data generation, multi-layer quality filtering, targeted loss design, and field-level F1 evaluation.
Comparison of Bi-directional RRT and Weighted A* for collision-free path planning across seven 3D obstacle environments.
An EKF-based visual-inertial SLAM pipeline that fuses high-rate IMU prediction with stereo observations for 3D trajectory and landmark mapping.
Fine-tuning SmolVLA across 5 data regimes on LIBERO-Spatial to investigate how many robot demonstrations a Vision-Language-Action model actually needs.