Humanoid Robot Companies Sell Hardware at a Loss to Gather Valuable Training Data

Related Insights

Robotics Lacks an 'Internet-Scale' Public Dataset, Forcing Firms to Bootstrap Data Collection

The rapid progress of many LLMs was possible because they could leverage the same massive public dataset: the internet. In robotics, no such public corpus of robot interaction data exists. This “data void” means progress is tied to a company's ability to generate its own proprietary data.

Uncapped #32 | Kyle Vogt from The Bot Company

Uncapped with Jack Altman·3 months ago

Prioritize Extreme Affordability in Early Home Robots to Fuel the Data Flywheel

For consumer robotics, the biggest bottleneck is real-world data. By aggressively cutting costs to make robots affordable, companies can deploy more units faster. This generates a massive data advantage, creating a feedback loop that improves the product and widens the competitive moat.

Uncapped #32 | Kyle Vogt from The Bot Company

Uncapped with Jack Altman·3 months ago

Robotics AI Models Can Now Learn from Human Video, Unlocking a Scalable Training Path

Physical Intelligence demonstrated an emergent capability where its robotics model, after reaching a certain performance threshold, significantly improved by training on egocentric human video. This solves a major bottleneck by leveraging vast, existing video datasets instead of expensive, limited teleoperated data.

Amazon x OpenAI, Ford's EV Reality Check, Kushner Drops WB Bid | Sarah Guo, David Senra, Doug O'Laughlin, Doug Bernauer, Jacob Effron, Logan Kilpatrick

TBPN·2 months ago

Humanoid Robots Will Launch as Teleoperated Services Before Achieving Full Autonomy

Companies developing humanoid robots, like One X, market a vision of autonomy but will initially ship a teleoperated product. This "human-in-the-loop" model allows them to enter the market and gather data while full autonomy is still in development.

Diet TBPN: October 29, 2025

TBPN·4 months ago

Nio's $20K Home Robot Sells a 'Social Contract' Where Early Adopters Pay to Train AI

The first home humanoid robot, Nio, requires frequent human remote intervention to function. The company frames this not as a flaw but a "social contract," where early adopters pay $20,000 to actively participate in the robot's AI training. This reframes a product's limitations into a co-development feature.

📺 “Nobody Wants This… Product placement” — Netflix's ad drama. Earnings Season’s exception. Neo’s home robot. +DoorDash’s cuffing szn

The Best One Yet·4 months ago

Humanoid Robot Development is Bottlenecked by In-Home Data Collection, Not Hardware

Progress in robotics for household tasks is limited by a scarcity of real-world training data, not mechanical engineering. Companies are now deploying capital-intensive "in-field" teams to collect multi-modal data from inside homes, capturing the complexity of mundane human activities to train more capable robots.

Centific’s Role in AI Boom, Databricks $134B Valuation, Alien Hunter Funding | Dec 16, 2025

The Information's TITV·2 months ago

Scarce, Actively Generated Data Is the New Moat for Robotics and Biology AI

The future of valuable AI lies not in models trained on the abundant public internet, but in those built on scarce, proprietary data. For fields like robotics and biology, this data doesn't exist to be scraped; it must be actively created, making the data generation process itself the key competitive moat.

Josh Wolfe & Brett McGurk – Venture, Geopolitics, and the Next Frontier (EP.476)

Capital Allocators – Inside the Institutional Investment Industry·2 months ago

Humanoid Robot Demos Rely on Human Teleoperation, a Necessary 'Scaffolding' Phase Similar to Early Self-Driving Cars

While Figure's CEO criticizes competitors for using human operators in robot videos, this 'wizard of oz' technique is a critical data-gathering and development stage. Just as early Waymo cars had human operators, teleoperation is how companies collect the training data needed for true autonomy.

Elon's Trillion Dollar Pay Package, Breaking Down the State of AI | Katherine Boyle, Mikey Shulman, Immad Akhund, Jordan Castro

TBPN·3 months ago

Better Data Unlocked Transformers for Robotics, Not Vice-Versa

The adoption of powerful AI architectures like transformers in robotics was bottlenecked by data quality, not algorithmic invention. Only after data collection methods improved to capture more dexterous, high-fidelity human actions did these advanced models become effective, reversing the typical 'algorithm-first' narrative of AI progress.

Sunday Robotics: Scaling the Home Robot Revolution with Co-Founders Tony Zhao and Cheng Chi

No Priors: Artificial Intelligence | Technology | Startups·3 months ago

Human-Facing AIs Are Covertly Mining Training Data to Accelerate the AGI Race

Companies like Character.ai aren't just building engaging products; they're creating social engineering mechanisms to extract vast amounts of human interaction data. This data is a critical resource, like a goldmine, used to train larger, more powerful models in the race toward AGI.

The AI Dilemma with Tristan Harris – The Prof G Pod

Pivot·2 months ago