惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

N
Netflix TechBlog - Medium
G
Google Developers Blog
H
Hackread – Cybersecurity News, Data Breaches, AI and More
T
The Blog of Author Tim Ferriss
Microsoft Azure Blog
Microsoft Azure Blog
GbyAI
GbyAI
L
LangChain Blog
云风的 BLOG
云风的 BLOG
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
aimingoo的专栏
aimingoo的专栏
P
Proofpoint News Feed
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
小众软件
小众软件
WordPress大学
WordPress大学
A
About on SuperTechFans
大猫的无限游戏
大猫的无限游戏
C
Check Point Blog
月光博客
月光博客
Stack Overflow Blog
Stack Overflow Blog
美团技术团队
Jina AI
Jina AI
T
Tailwind CSS Blog
Google DeepMind News
Google DeepMind News
D
Docker

Interesting Engineering

US firm to scale laser-based nuclear fusion ‘breakthrough’ with new partnership Military Archives - Interesting Engineering World’s first non-nuclear lead-cooled reactor to generate electricity begins installation US scientists devise new process to turn sewage sludge into 99% pure natural gas US firm unveils submarine-hunting drone with 9,200-mile-range, 35 mph top speed Military Archives - Interesting Engineering Supercomputer finds lithium-titanium tweak to boost sodium-ion batteries for grids Lockheed Martin demonstrates vertical launch missile system for mobile drone defense China’s 1116 MWe Taipingling Unit 1 reactor goes online, set to generate 9bn kWh yearly ChatGPT Images 2.0 update combines reasoning, research, and design with 2K output US Navy tests plug-and-play laser system on USS Bush carrier, downs drones at sea China’s CATL reveals 621-mile EV battery, under-7-minute charging to challenge BYD US uses world’s first exascale supercomputer to model supernovae, fusion reactors AI and Robotics Archives - Interesting Engineering First-in-human study confirms safety of graphene-based brain interface Tesla’s Optimus humanoid robot greets runners, poses for photos at Boston Marathon Interlocking materials offer high strength and flexibility for robotics, infrastructure US redeploys 100,000-ton nuclear-powered aircraft carrier in Red Sea after repairs US scientists unveil concept for ‘world’s first neutrino laser’ to unlock breakthroughs New military tech can maintain communication in contested electronic warfare environments Got a dark personality? Psychologists can help you choose your career wisely Humidity boosts performance of 3D-printed nanogenerator instead of degrading it China demonstrates microwave beam that recharges drones in flight, continues power delivery Scientists run compact free-electron laser for eight hours, cracks FEL stability problem China’s PLA considers to use minelaying underwater drones to enforce Taiwan blockade: Report 1-ton sharks may struggle for survival in waters exceeding 62.6°F, study suggests US firm’s thorium nuclear fuel bundles move to manufacturing for commercial reactors Tesla hits 0% charge in remote Chilean desert as YouTuber uses hood-mounted solar Humanoid robot surpasses human world record in Beijing half-marathon, clocking 50:26 mins New method extracts maximum work from unknown quantum states using symmetry tricks
One brain for all: China builds unified AI model to handl...
Neetika Walt · 2026-04-30 · via Interesting Engineering

Motubrain combines video, language, and action to help robots complete complex tasks in real-world settings.

Robots trained with Motubrain can adapt mid-task, correcting errors and retrying actions without prior instruction.

Robots trained with Motubrain can adapt mid-task, correcting errors and retrying actions without prior instructionShengShu Technology

ShengShu Technology has unveiled Motubrain, a unified AI model designed to act as a general-purpose brain for robots, combining perception, reasoning, prediction, and action into a single system.

The company says the model replaces fragmented, task-specific architectures typically used in robotics with a single framework capable of handling multiple tasks and environments. The approach aims to reduce dependence on separate modules for sensing, planning, and execution.

Motubrain has already shown strong benchmark performance, achieving a 63.77 score on WorldArena and averaging 96.0 across 50 tasks on RoboTwin 2.0. It is also reported to be the only model to exceed 95.0 in randomized environments.

The system builds on ShengShu’s earlier work in generative video through its Vidu platform, using large-scale video data to train robots to understand and interact with real-world environments.

One brain, many tasks

Motubrain is designed as a unified multimodal model that learns from video, language, and action simultaneously. This allows robots to process their surroundings, predict outcomes, and act in real time without switching between separate systems.

“A true world model must be able to build a unified representation of the real world and predict how it evolves,” said Jun Zhu, Founder of ShengShu Technology.

The model uses a three-stream Mixture-of-Transformers architecture to integrate inputs from different modalities. This setup enables robots to understand instructions, anticipate environmental changes, and generate appropriate actions in one continuous loop.

Unlike conventional systems that rely heavily on labeled datasets, Motubrain is trained using a broader mix of unlabelled video, simulation data, and multi-robot task recordings. A latent action framework extracts motion patterns directly from these inputs, reducing the need for manual annotation.

This training approach allows the model to scale more efficiently. In internal evaluations, Motubrain maintained higher success rates than competing systems as both task complexity and training data increased.

From data to action

Motubrain can execute multi-step tasks involving up to 10 atomic actions, significantly more than the typical 2–3 handled by many current robotic systems. This enables robots to complete more complex, real-world activities in a single sequence.

“We believe general world models should not be built as stitched-together modules, but as a unified architecture that brings together perception, reasoning, prediction, generation, and action in a single system.”

In real-world tests, robots trained with Motubrain demonstrated the ability to adapt during execution. For example, when a task failed mid-action, such as picking up an object unsuccessfully, the system could recognize the failure and retry without prior training on that specific scenario.

The company says the model is already being used by robotics firms in active training programs across industrial, commercial, and home environments. Partnerships with companies including Astribot, SimpleAI, and Anyverse Dynamics aim to further expand deployment.

Backed by a $293 million Series B led by Alibaba Cloud, ShengShu is positioning Motubrain as a key step toward general-purpose embodied AI systems capable of operating across diverse real-world settings.

The Blueprint

Get the latest in engineering, tech, space & science - delivered daily to your inbox.

With over a decade-long career in journalism, Neetika Walter has worked with The Economic Times, ANI, and Hindustan Times, covering politics, business, technology, and the clean energy sector. Passionate about contemporary culture, books, poetry, and storytelling, she brings depth and insight to her writing. When she isn’t chasing stories, she’s likely lost in a book or enjoying the company of her dogs.