Robot-Use Agents: Why General-Purpose Models May Win in Robotics

Y Combinator · 2026-09-26 · 30 min
https://www.youtube.com/watch?v=Jv5B5CEaPJIVideo summary
Waddle Labs and RoboCurve argue general-purpose LLMs could control diverse robots, with teenager-level general-purpose robots possible within two years.
Waddle Labs co-founder Jaime describes building robot-control harnesses and training models with collected data, while RoboCurve’s Jay evaluates models across robot types and environments. The guests trace the shift from RT-2, which fine-tuned language models to output robot poses, to coding agents that use tools and write reusable robot policies. They distinguish quick, low-cost in-context learning from training updates, noting it can saturate after roughly 20–40 examples and is limited by context length. In demonstrations, Astra directs robot arms using camera feeds and tool calls; repeated actions can be compiled into faster skills, especially as model latency reportedly improves about 2× per month. The discussion argues that coding, computer-use, CAD, and egocentric data can teach spatial understanding. Jay predicts natural-language robots capable of tasks a competent teenager could do may arrive within two years, while latency, skill compression, and organizing learned skills remain major challenges.
Chapters
- 0:00Intro + The Rise Of Robot Use Agents: Waddle Labs, RoboCurve, and RT-2
- 4:07The Bitter Lesson for Robotics: Web Pretraining and General-Purpose LLMs
- 7:40From Coding Agents to Robot Policies: Voyager and One-Shot DeepMind Methods
- 10:41In-Context Learning vs Model Training: ICL Saturates After 20–40 Examples
- 14:22Building a Harness for Robot Control: Ashra’s Block Demo and 2x Monthly Latency Gains
- 17:50Turning Robot Actions Into Reusable Skills: Astra Tool Calls and Program Induction
- 20:55How AI Models Learn the Physical World: Astra, CAD Data, and Shared Representations
- 26:05How Close Are We to General-Purpose Robots? Two-Year Outlook, Astra Latency, and Skill Distillation
This is a Tier 1 public summary
Whether the chapter key points, section summaries and mind map are public is up to the person who shared it. Want the full analysis?Submit one yourself.
More from this channel
Patrick Collison: Is AI Breaking the Lean Startup Playbook?Y CombinatorPatrick Collison says Stripe’s new-business starts are nearly 2x year over year, making this a better time than ever to found a company.
The Case For Data Centers In SpaceY CombinatorStarcloud CEO Philip Johnston explains how an H100 reached orbit and why 88,000 satellites could host space-based AI data centers.
Waymo Co-CEO Dmitri Dolgov: The Demo Is Only 1% Of The WorkY CombinatorWaymo co-CEO Dmitri Dolgov says the demo took 18 months, but building a safe robotaxi product took 15 years.
Garry Tan: Own Your IntelligenceY CombinatorGarry Tan says personal AGI can make one founder 400x more productive—and argues your AI skills should stay in your own repo.
Max Hodak: Average Is Not Good EnoughY CombinatorScience CEO Max Hodak says Helix and faster iteration drive startup speed, while one retinal implant patient read a 300-page novel.
Related analyses
Inside OpenAI’s Breakthroughs in Mathematical Reasoninga16zOpenAI mathematicians Mehtaab Sawhney and Mark Sellke explain Astra’s sphere-packing bound and non-sofic group proof.
OpenAI President on What it Means to Cross Into The Age of AGIa16zOpenAI’s Greg Brockman calls Astra AGI after 24-hour tasks and urges a billion-dollar cybersecurity push for frontline defenders.
EP336. 前沿模型該減速嗎、Anthropic 威脅報告、為什麼要一直聊模型 | M觀點M觀點Anthropic 執行長提議控管前沿 AI,Kimi 被指將解放軍監控影像轉給 Claude,主持人認為安全測試未必利空 AI 硬體。
EP155 - 席捲 AI 圈的 Jev,到底在紅什麼?AI Agent 真的要大改了嗎?科技浪Jev 每次只需輸出分類機率,宣稱快 200 倍、省 400 倍,適合篩選與模型路由,不能取代完整 LLM。