TGViewer
All about AI, Web 3.0, BCI All about AI, Web 3.0, BCI @alwebbci · 3.9K subscribers
Post #3796 780
Physical intelligence introduced a new model π*0.6

π*0.6 can more than double throughput over a base model trained without RL, and can perform real-world tasks: making espresso drinks, folding diverse laundry, and assembling boxes.

Team trained a general-purpose value function on all of own data, which tells the π*0.6 VLA which actions are good or bad. By asking π*0.6 to produce only good actions, researchers get better performance. Team call this method Recap.

π*0.6 can then collect more autonomous data, which can be used to further train the value function and further improve π*0.6.

During autonomous data collection, a teleoperator can also intervene and provide corrections for significant mistakes, coaching π*0.6 further.

Quantitatively, training π*0.6 with RL can more than double throughput (number of successful task executions per hour) on the hardest tasks and cut the number of failures by as much as a factor of two.
  • 🔥 5
  • 🥰 3
  • 👏 3
More from @alwebbci
  1. Oct 7, 2026Sui partners with Alibaba Cloud to enable stablecoin-based per-call payments for AI agent…
  2. Oct 7, 2026Sierra announced Personal Agent Protocol - an open standard It will help define how person…
  3. Oct 6, 2026Vitalik Buterin: AI’ll become the new UI and traditional on-chain applications may no long…
  4. Oct 6, 2026Mistral launched a preview of new model, Mistral Large 4 (ML4), aka le Chonk ML4 is a 1T-p…
  5. Oct 6, 2026Reflection introduced Beam: an agentic open model with 501B total parameters and 23B activ…
  6. Oct 5, 2026Anthropic removes Cowork's local option for Pro/Max tomorrow. New tasks run in the cloud.…
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →