Papers
23
Total Citations
1,344
H-Index
13
About
Karl Pertsch is a robotics and machine learning researcher whose work sits at the intersection of large-scale robot learning, foundation models, and generalist robot control. He has made significant contributions to some of the most influential robotics projects of the past few years, including the landmark RT-1 and RT-2 papers (512 and 267 citations respectively), which demonstrated how transformer-based architectures and vision-language models trained on internet-scale data can be leveraged to dramatically improve robotic generalization and real-world performance. Pertsch has been a key contributor to the development of open-source generalist robot policies, including Octo and OpenVLA, democratizing access to powerful robotic foundation models. His work on the DROID large-scale manipulation dataset and the π₀ vision-language-action flow model reflects a sustained commitment to building robust data infrastructure and scalable control frameworks. Earlier work on skill priors for reinforcement learning highlights his foundational interest in knowledge transfer and sample-efficient learning. More recent contributions, such as FAST action tokenization and language-correction methods for robots, demonstrate his range across both practical deployment and core algorithmic innovation. Across his career, Pertsch has helped define how modern AI techniques are reshaping robot learning at scale.
Research Focus
Key Achievements
Top Papers
- 1RT-1: Robotics Transformer for Real-World Control at Scale512 citations · 2023
- 2RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control267 citations · 2023
- 3π₀: A Vision-Language-Action Flow Model for General Robot Control127 citations · 2025
- 4DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset108 citations · 2024
- 5Octo: An Open-Source Generalist Robot Policy66 citations · 2024
- 6OpenVLA: An Open-Source Vision-Language-Action Model39 citations · 2024
- 7RT-1: Robotics Transformer for Real-World Control at Scale38 citations · 2022
- 8Yell At Your Robot: Improving On-the-Fly from Language Corrections34 citations · 2024
- 9FAST: Efficient Action Tokenization for Vision-Language-Action Models28 citations · 2025
- 10Accelerating Reinforcement Learning with Learned Skill Priors19 citations · 2020