AIToday
RoboticsAI Safety & AlignmentTechCrunch AIPublished: Jul 27, 2026, 10:00 JST

Brain waves join robot training data arsenal

Brain waves join robot training data arsenal

3 Key Points

  1. What happened

    Encord, a company that builds data tools for AI model training, is running a trial with Zander Labs—a German neuroscience startup—to tag robot training data with brain-wave measurements. Pilots wearing headsets that track both eye movement and brain activity perform physical tasks like pulling blocks from a Jenga tower, creating annotated training data that captures mental states such as error, intent, and surprise.

  2. Why it matters

    The robotics industry faces a severe shortage of real-world physical training data—the constraint that may now matter more than model architecture itself. Encord's internal analysis suggests it will take a data set roughly five times the size of YouTube's video corpus to teach robots manipulation at the scale LLMs achieved with internet text. Since physical data must be manufactured rather than scraped, the cost and scarcity of high-quality labeled training data has become a business bottleneck; brain-wave signals could reveal which moments matter most for model training, potentially improving efficiency.

  3. What to watch

    Encord is conducting this as a trial to evaluate whether brain-wave-tagged data actually improves robotic model performance before deciding to scale. The company's San Leandro warehouse is also experimenting with other new data modalities, including forearm sensors that detect electrical signals in muscles to create 3D hand-position data. Encord draws egocentric video from factories globally and uses remote-operated robotic arms (leader-follower rigs) to generate task-specific training data.

Not sure about something? Ask the AI

Summaries like this, in your inbox every morning.

Context & Analysis

The constraint limiting physical AI has shifted from model architecture to the sheer scarcity of real-world training data. Unlike large language models, which were built on text scraped from the internet at near-zero cost, robots require hands-on, physically grounded training data that must be deliberately manufactured. Encord's pivot from simply managing customer data to producing it reflects this new reality: the company realized that as robotics firms attempted end-to-end learning for manipulation tasks, "the data simply does not exist." This economic shift—from collection to manufacturing—changes the entire equation for physical AI development.

Encord's experiment with brain-wave measurement represents an attempt to increase the informational density of each training sample. By capturing not just video but also the pilot's mental states (error detection, intent, surprise), the company hopes to create a richer signal for model training. Velmurugan's claim that such dense annotation is worth 100 times as much as raw video, while costing only 20 times more, suggests a genuine efficiency gain—but "20 times more" remains a substantial cost barrier. The company's multi-modal approach—combining egocentric video, remote-operated robotic arms, muscle electrical signals, and now brain waves—reflects a broader bet that the next competitive advantage in robotics will belong to whoever can manufacture the highest-fidelity training data most efficiently.

FAQ
What does the brain-wave headset measure?
The headset, built by Zander Labs, measures brain activity to deduce mental states like error, intent, and surprise. Zander's neuroscientist Lucas Gehrke explained that the amount of brain activity at any point during a task offers clues for model builders trying to figure out when they need to deploy their highest-effort models.
How much more expensive is annotated physical training data compared to raw video?
Velmurugan estimates that densely annotated data (with physical descriptions like "right hand tightens bolt") is worth 100 times as much as unannotated egocentric video for training specific tasks, and costs only 20 times more to produce.
How much training data would robots need to match what LLMs achieved?
Velmurugan says it will take a data set roughly five times the size of YouTube's video corpus to break through—a scale that helps explain why data generation itself has become a business and not just a research problem.

Get the latest Robotics news every morning

For example, today's edition would include:

  • Stellantis taps Inbolt's sub-$600 vision to stop robot freezesExponential Industry · 50m ago
  • Wiwynn adds cobots at El Paso, Mexico plants for rack testingDIGITIMES Asia · 12h ago
  • NTTドコモ builds robot ML pipeline in 3 days with AI-DLCTop Companies AI · 19h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleArista Networks scores low on value metrics despite 5-year 7x return