Visual Servoing with Residual Reinforcement Learning
Combining classical IBVS with a PPO residual policy for robust eye-in-hand manipulation on a Franka Panda robot.
Combining classical IBVS with a PPO residual policy for robust eye-in-hand manipulation on a Franka Panda robot.
A production-grade pipeline for generating photorealistic lip-synced face videos from arbitrary speech audio, combining SyncNet with a U-Net generator.
Fine-tuning LLaMA with LoRA to convert free-text medical abstracts into structured JSON — covering synthetic data generation, multi-layer quality filtering, targeted loss design, and field-level F1 evaluation.
Comparison of Bi-directional RRT and Weighted A* for collision-free path planning across seven 3D obstacle environments.
An EKF-based visual-inertial SLAM pipeline that fuses high-rate IMU prediction with stereo observations for 3D trajectory and landmark mapping.
Fine-tuning SmolVLA across 5 data regimes on LIBERO-Spatial to investigate how many robot demonstrations a Vision-Language-Action model actually needs.
Qihao Qian
arXiv
Enhances Yolact instance segmentation with edge detection to improve rail mask precision for autonomous train obstacle avoidance.
Zhirui Dai*, Qihao Qian*, Tianxing Fan, Nikolay Atanasov
Submitted to IROS 2026
A hybrid method combining explicit gradient-augmented octree interpolation with implicit neural residuals for efficient and accurate online SDF reconstruction.
Published:
This is a description of your talk, which is a markdown file that can be all markdown-ified like any other post. Yay markdown!
Published:
This is a description of your conference proceedings talk, note the different field in type. You can put anything in this field.
Undergraduate course, University 1, Department, 2014
This is a description of a teaching experience. You can use markdown like any other post.
Workshop, University 1, Department, 2015
This is a description of a teaching experience. You can use markdown like any other post.