Pengwei Xie

I am a researcher in embodied intelligence at AgiBot. Previously, I worked at Huawei Noah's Ark Lab (2025-2026). I received my Ph.D. from Tsinghua University in 2024, conducting research at the Visual Computing Laboratory (VCLab), Department of Electronic Engineering. From 2021 to 2023, I was a visiting scholar at UCSD working with Prof. Hao Su. I received my B.S. from Tsinghua University's Department of Electronic Engineering in 2018.

Pengwei Xie

Research

I'm interested in computer vision, robotic manipulation, reinforcement learning and sim2real. I focus on scalable algorithms to achieve AGI in complex real-world environments.

τ0-VLA architecture

τ0-VLA: a Hierarchical Robot Foundation Model with World-Model-Guided Test-Time Computation

Preprint, 2026

project page / paper / code / model

Learning While Deploying fleet-scale reinforcement learning framework

Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies

arXiv, 2026

project page / arxiv / paper

OmniEVA

OmniEVA: Embodied Versatile Planner via Task-Adaptive 3D-Grounded and Embodiment-aware Reasoning

International Conference on Learning Representations (ICLR), 2026

project page / arxiv

More than A Point

More than A Point: Capturing Uncertainty with Adaptive Affordance Heatmaps for Spatial Grounding in Robotic Tasks

arXiv, 2025

project page / arxiv

APeG

Active-Perceptive Language-Oriented Grasp Policy for Heavily Cluttered Scenes

IEEE Robotics and Automation Letters (RA-L), 2025

paper / code

FlexLoG

Rethinking 6-DoF Grasp Detection: A Flexible Framework for High-Quality Grasping

Pattern Recognition (PR), 2025

paper / arxiv / code

GenH2R: Learning Generalizable Human-to-Robot Handover via Scalable Simulation Demonstration and Imitation

IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2024

project page / arxiv / code

RegionNormalizedGrasp

Region-Aware Grasp Framework with Normalized Grasp Space for Efficient 6-Dof Grasping

Conference on Robot Learning (CoRL), 2024

arxiv / code

GAP-RL

GAP-RL: Grasps As Points for RL Towards Dynamic Object Grasping

IEEE Robotics and Automation Letters (RA-L), 2024

arxiv / code

Target-Oriented

Target-Oriented Object Grasping via Multimodal Human Guidance

European Conference on Computer Vision Workshops (ECCVW), 2024

arxiv

Part-Guided 3D RL

Part-Guided 3D RL for Sim2Real Articulated Object Manipulation

IEEE Robotics and Automation Letters (RA-L), 2023

arxiv / code

HGGD

Efficient Heatmap-Guided 6-Dof Grasp Detection in Cluttered Scenes

IEEE Robotics and Automation Letters (RA-L), 2023

arxiv / code

ManiSkill2

ManiSkill2: A Unified Benchmark for Generalizable Manipulation Skills

International Conference on Learning Representations (ICLR), 2023

project page / arxiv / code