惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

S
SegmentFault 最新的问题
博客园 - 三生石上(FineUI控件)
WordPress大学
WordPress大学
博客园 - 【当耐特】
月光博客
月光博客
Vercel News
Vercel News
D
Docker
I
InfoQ
Apple Machine Learning Research
Apple Machine Learning Research
博客园 - 叶小钗
MongoDB | Blog
MongoDB | Blog
GbyAI
GbyAI
有赞技术团队
有赞技术团队
雷峰网
雷峰网
博客园 - 聂微东
小众软件
小众软件
Y
Y Combinator Blog
腾讯CDC
L
LangChain Blog
The GitHub Blog
The GitHub Blog
宝玉的分享
宝玉的分享
Stack Overflow Blog
Stack Overflow Blog
大猫的无限游戏
大猫的无限游戏
T
The Blog of Author Tim Ferriss

School of Computer Science News

Robotics Innovation Center Earns LEED Platinum Certification for Sustainable Construction AI4MiddleSchools Expands Nationwide Effort To Prepare Students for an AI-Powered Future Carvalho Earns NSF CAREER Award To Study Motivation and Learning Season Three of 'Does Compute' Now Available Rare Ventures Partners Rings NYSE Opening Bell Bringing Images to Life Through Touch - Robotics Institute Carnegie Mellon University Fried Receives NSF CAREER Award - Language Technologies Institute - School of Computer Science - Carnegie Mellon University PAIR Helps Students Find Their Place in AI Research Navigating the AI Era with a CMU Focus on Critical Thinking Kaess Named to Inaugural Chief of Naval Research Fellows Program Navigating the Moon Koedinger Wins Lifetime Achievement Award Carnegie Mellon Names Damion Shelton Associate VP and Executive Director of the Swartz Center for Entrepreneurship Hong Shen Discusses AI Safety at WEF Annual Meeting Erickson Earns NSF CAREER Award - Robotics Institute Carnegie Mellon University Satya Honored With Test of Time Award From Proof to Program: CMU and the Rise of AI-Driven Mathematics SCS Researchers Named to Inaugural ACM SIGSOFT Software Engineering Academy Healthcare Blind Spots: AI Models Prone To Fabricating Diagnoses - Robotics Institute Carnegie Mellon University You Can't Remove Humans From Software Engineering Designing the Future of Tech Governance AI, Single-Cell Technology Reveal How 3D Genome Differs in People With Alzheimer's Disease Tepper School of Business and School of Computer Science Partner to Launch AI for Business Executive Education Program Carnegie Mellon Researchers Lead Three DOE Genesis Mission Awards to Advance the Future of AI-Enabled Scientific Discovery Snake Robots Support Earthquake Search and Rescue in Venezuela Lindlbauer Receives NSF CAREER Award for Adaptive Extended Reality Interfaces Fredrikson Earns Test of Time Award for AI Security CMU Advances Defense Manufacturing and Military Education at Pennsylvania Defense and Innovation Summit Looking Ahead: AI Needs UI Liu Receives NSF CAREER Award
CMU Researchers Train Robots With Internet Videos - Robot...
Mallory Lindahl · 2026-06-17 · via School of Computer Science News
Warning: You are viewing this site with an outdated/unsupported browser. Please update your browser or consider using a different one in order to view this site without issue.
For a list of browsers that this site supports, see our Supported Browsers page.
Skip to content

CMU Researchers Train Robots With Internet Videos

VideoManip Converts Videos of People and Objects Interacting Into Training Data

06/17/2026    Mallory Lindahl

The Breakdown: 

  • VideoManip teaches robots manipulation skills using videos of people interacting with objects.
  • It reconstructs movements and estimates how people make contact with objects.
  • The system helps robots learn new skills without time-consuming, human-operated demonstrations.

* * * 

Researchers in Carnegie Mellon University’s School of Computer Science are developing a new way for robots to learn everyday tasks like grasping a coffee mug or opening a drawer by watching videos of people interacting with objects. Their system, VideoManip, converts these videos into training data that can be used to teach robots dexterous manipulation skills.

VideoManip could eliminate the need for robot demonstrations or specialized motion-capture systems, making it easier for robots to learn how to use everyday objects.

“Humans have designed the world for human hands,” said Jeffrey Ichnowski, an assistant professor in the Robotics Institute (RI). “The tools, devices and objects around us are meant to be operated with human hands. A common challenge in robotics is determining how robots can interact with those same objects the way people do.”

VideoManip addresses one of the field’s biggest bottlenecks: data collection.

Teaching robots dexterous manipulation tasks requires large amounts of training data. Traditionally, collecting that data has involved specialized equipment, wearable sensors and hours of teleoperated demonstrations.

“While AI systems like ChatGPT can learn from massive amounts of internet data, robot learning — teaching robots how to physically interact with the world — has struggled to scale in the same way,” said Hongyi Chen, an RI Ph.D. student and lead researcher on the VideoManip team. “Collecting examples of people grasping, moving and manipulating objects is much more difficult than gathering text or images from the internet.”

VideoManip doesn’t rely on specialized equipment or human-operated robot demonstrations for training. Instead, it can use any video that focuses on a person using or manipulating an object. The system analyzes a video, reconstructs 3D movements of a person’s hands and the object, and estimates how the person makes contact with that object. It then translates those actions into movements a robotic hand can perform. The researchers found that robotic platforms trained using VideoManip could perform a variety of real-world manipulation tasks using a multifinger robotic hand.

The system can also create many additional training examples from a single video using computer simulation. By generating different versions of the same task, VideoManip gives robots more opportunities to practice and learn.

“We’re trying to enable robots to learn from videos that are readily available on the internet,” said Zackory Erickson, an assistant professor in the RI. “When we can start leveraging that, we have a clear way forward to solving the lack-of-data problem for robot learning.”

The project reflects a broader shift occurring across robotics, where researchers are increasingly exploring internet-scale datasets and foundation-model approaches to robot learning. By transforming ordinary human videos into robot training data, VideoManip offers a path toward teaching robots new skills simply by watching people perform them.

Along with Ichnowski, Chen and Erickson, the VideoManip team included Tony Dong, an SCS undergraduate student; RI master’s students Tiancheng Wu and Yash Jangir; RI Ph.D. student Yaru Niu; and RI Ph.D. graduate Homanga Bharadwaj. The RI researchers also collaborated with Liquan Wang, a Ph.D. student at the Georgia Institute of Technology, and Yufei Ye, a researcher at Stanford University.    

To learn more about VideoManip, visit the project website.

For More Information: Aaron Aupperlee | 412-268-9068 | aaupperlee@cmu.edu

2026-06-17T13:29:47-04:00

Share This Story!