I have been a researcher at the Robotics Institute, Carnegie Mellon University (CMU), since 2019. My research spans efficient AI, embodied AI, computer vision, wireless perception, and healthcare.
Before research, I worked in industry on GPU systems and AI algorithms at NVIDIA, and on telecom embedded software, Linux systems, and distributed network programming at Motorola, Agilent Technologies, etc.
My work has appeared at NeurIPS, ICLR, CVPR, ECCV, AAAI, IJCAI, EMNLP, AAMAS, and IROS, and in IEEE TCAD and ACM TECS. I serve as a reviewer for NeurIPS, ICLR, CVPR, AAAI, AAMAS, and EMNLP.
Education
- Ph.D.Northeastern University, Boston, USA
- M.S. and B.S.Beihang University, Beijing, China
Honors
- 2026IEEE Senior Member
Selected publications
- AAMAS 2026Structured Agent Distillation for Large Language Models
- Findings of EMNLP 2026RCR-Router: Efficient Role-Aware Context Routing for Multi-Agent LLM Systems with Structured Memory
- IROS 2026IndoorR2X: Indoor Robot-to-Everything Coordination with LLM-Driven Planning
- Findings of EMNLP 2026End-to-end on-device quantization-aware training for LLMs at inference cost
- CVPR 2026Roots Beneath the Cut: Uncovering the Risk of Concept Revival in Pruning-Based Unlearning for Diffusion Models
- NeurIPS 2025Harmony in Divergence: Towards Fast, Accurate, and Memory-efficient Zeroth-order LLM Fine-tuning
- ICLR 2025Mutual Effort for Efficiency: A Similarity-based Token Pruning for Vision Transformers in Self-Supervised Learning
- AAAI 2025Toward adaptive large language models structured pruning via hybrid-grained weight importance assessment
- ICASSP 2025RoRA: Efficient Fine-Tuning of LLM with Reliability Optimization for Rank Adaptation
- IJCAI 2025fairgnn-wod: Fair graph learning without complete demographics
- IJCAI 2025FairSMOE: Mitigating Multi-Attribute Fairness Problem with Sparse Mixture-of-Experts
- ACM TACO 2025Mobile-3DCNN: An Acceleration Framework for Ultra-Real-Time Execution of Large 3D CNNs on Mobile Devices
- PLOS Digital Health 2025AI-driven healthcare: Fairness in AI healthcare: A survey
- IEEE TCAD 2024TSLA: A Task-Specific Learning Adaptation for Semantic Segmentation on Autonomous Vehicles Platform
- ECCV 2024Instructgie: Towards generalizable image editing
- ACM SIGKDD Explorations 2025Graph Fairness via Authentic Counterfactuals: Tackling Structural and Causal Challenges
- GLSVLSI 2025Towards Memory-Efficient and Sustainable Machine Unlearning on Edge using Zeroth-Order Optimizer
- ICCAD 2025Perturbation-efficient zeroth-order optimization for hardware-friendly on-device training
- ASP-DAC 2025A computation and energy efficient hardware architecture for SSL acceleration
- ICASSP 2024Df-vton: Dense flow guided virtual try-on network
- WACV 2024Physical-space multi-body mesh detection achieved by local alignment and global dense learning
- ACMM 2023 WorkshopA scalable real-time semantic segmentation network for autonomous driving
- ICRA 2024Mono Wi-Fi Scene
- ICRA 2024Universal Monocular 3D Human Recovery Engine
- ACM TECS 2022Mobile or FPGA? A comprehensive evaluation on energy efficiency and a unified optimization framework
- IEEE RTAS 2021Work in progress: Mobile or FPGA? A comprehensive evaluation on energy efficiency and a unified optimization framework