Teahose.
SIGN IN
NEW HERE — WHAT TEAHOSE DOES
We read the entire AI & tech firehose — so you don't have to.
PODPodcastsAll-In, No Priors, Acquired…
NEWNewslettersStratechery, Newcomer…
PAPPapersPhysical AI research
PHProduct Huntdaily launches
VCInvestor ScoutSequoia, a16z, Benchmark…
CLAUDE DISTILLS →
7 reads, 30 sec each — free, 6 AM ET.
+ a live graph of the companies, people & themes underneath.
HOME/THEMES/REINFORCEMENT LEARNING FOR ROBOTICS
// THEME

Reinforcement Learning for Robotics

Research labs and platforms applying deep reinforcement learning directly to robot skill acquisition and control, enabling robots to learn dexterous and locomotion tasks from reward signals rather than demonstrations.

COMPANIES 48VELOCITY — STABLECAPITAL 28D $617.3M · 12 DEALS
TOP INVESTORS: amazon (10) · andreessen horowitz (7) · index ventures (5) · nvidia (5) · bain capital (4)

CAPITAL FIGURES ARE MEDIA-EXTRACTED ESTIMATES, NOT VERIFIED FILINGS.

Capital burst driven by mega-rounds, now normalizing
◀ $5.0B · wk of 07-132026-06-08 ── 2026-08-31 · WEEKLY
Series B dominates deals; seed checks balloon
unknown
$671M · 9 DEALS
series b
$1.6B · 6 DEALS
seed
$283M · 4 DEALS
debt
$0M · 2 DEALS
series a
$5.5B · 2 DEALS
Mention momentum
MENTIONS / WEEK · PEAK 70

EXTRACTED FROM 25+ PODCASTS & VC NEWSLETTERS · MEDIA-REPORTED FIGURES, NOT VERIFIED FILINGS

// THE LEAD
▲ STRENGTHENING

RL fine-tuning of generalist VLAs is the new training paradigm

Physical Intelligence's π0.5 — built atop their flow-matching π0 backbone — now outperforms competing approaches on RoboTwin 2.0 (75.8% vs 49.2% for Temporal GRPO vs π0 baseline), validating the thesis that RL fine-tuning of pretrained VLAs is the dominant path to deployable robot policies. CMU researchers continue to anchor the field's talent pipeline, with multiple first-author and co-senior-author hires traced to the university. The RoboBRIDGE orchestration framework challenges the assumption that raw VLA scaling is sufficient, demonstrating that structured reasoning layers — boosting RoboCasa success from 3.7% to 7.5% — are needed alongside RL. Physical Intelligence's inclusion in Elad Gil's Conviction embed cohort further signals elite investor consensus around this architectural bet.

// TRENDS
▲ STRENGTHENINGSimulation speed unlocks practical real-world robot RL at scale

ManiSkill3 from UC San Diego slashes RL training time from nearly 8 hours (RLBench baseline) to 27 minutes using 32 parallel GPU environments — an 18x speedup that makes iterative reward-signal learning economically viable for commercial teams. RLBench remains the established benchmark comparator, and the Unitree G1 humanoid is emerging as a preferred low-cost physical test platform for validating sim-trained policies. Together, faster simulation and affordable hardware remove two of the three main barriers to real-world RL deployment.

Why it matters · Startups that integrate GPU-parallelized simulation into their training stack can compress robot skill acquisition timelines by an order of magnitude, creating durable cost advantages over demo-based competitors.

▲ NEWAdaptive locomotion RL closes the sim-to-real hardware gap

A new Rapid Embodiment Adaptation module demonstrates that inferring robot hardware parameters in 0.4 seconds — and feeding them explicitly into the controller — significantly outperforms implicit, end-to-end sensor-history approaches for quadrupedal locomotion. This explicit parameter-estimation strategy represents a structural shift in how locomotion RL policies handle real-world hardware variance without retraining.

Why it matters · Hardware-agnostic locomotion policies that adapt in under a second dramatically reduce deployment costs for robot operators managing heterogeneous fleets.

▲ NEWFrequency-adaptive diffusion policies resolve contact-rich RL tradeoffs

Shanghai Jiao Tong University researchers (Lifeng Zhuo, Chuan Wen, Wendi Chen, advised by Cewu Lu of Noematrix) identified a fundamental bottleneck in standard diffusion policies — fixed inference frequency forces a tradeoff between pre-contact multimodality and post-contact reactivity. Their FA-RDP (Frequency-Adaptive Reactive Diffusion Policy) resolves this by dynamically adjusting sampling steps and frequency during an episode, pointing toward a new generation of RL policies capable of handling contact-rich manipulation without architectural compromise.

Why it matters · Solving the reactivity-vs-multimodality tradeoff is a prerequisite for deploying robot RL policies in unstructured industrial and household environments where contact dynamics are unpredictable.

▲ NEWSeed-stage capital surging into human-robot interaction platforms

Enigma raised a $71M seed round backed by Index Ventures and Ribbit Capital — an unusually large seed for a company focused on new paradigms of human-intelligent-machine interaction, signaling that top-tier generalist VCs are entering robot RL adjacencies at formation stage. With 4 seed deals totaling $283M in the 90-day stage mix alongside 7 Series B deals at $1.77B, capital is bifurcating: large checks consolidating platform bets at Series B while oversized seeds fund frontier interaction-layer experiments.

Why it matters · The $71M Enigma seed sets a new price floor for human-robot interface startups, raising competitive pressure on incumbents and signaling that index-style VCs see interaction-layer differentiation as defensible.

CORROBORATED · 2 SOURCE TYPESParsers VC · Aug 4T1 Scout · Aug 3The VC Corner · Aug 2
// COMPANIES
48 COMPANIES
01
Mind Robotics
mindrobotics.ai
$500M · SERIES A · ANDREESSEN HOROWITZ + ACCEL · AUG 31
9 SIGNALS · LAST SEEN AUG 31, 2026
02
Amazon
amazon.com
$75M · GROWTH · T. ROWE PRICE + MICROSOFT · AUG 29
229 SIGNALS · LAST SEEN SEP 2, 2026
03
Mercor
mercor.com
UNKNOWN · AUG 29
33 SIGNALS · LAST SEEN AUG 29, 2026
04
Hugging Face
huggingface.co
149 SIGNALS · LAST SEEN SEP 5, 2026
05
Unitree Robotics
unitree.com
$904M · IPO · AUG 7
35 SIGNALS · LAST SEEN SEP 2, 2026
06
Enigma
$71M · AUG 4
6 SIGNALS · LAST SEEN AUG 4, 2026
07
SkildAI
UNKNOWN · SEQUOIA + ECLIPSE · JUL 17
1 SIGNAL · LAST SEEN JUL 17, 2026
08
Physical Intelligence
physicalintelligence.company
$5.0B · SERIES A · JUL 16
143 SIGNALS · LAST SEEN AUG 27, 2026
09
LeRobot
2 SIGNALS · LAST SEEN SEP 5, 2026
10
Galaxea
6 SIGNALS · LAST SEEN SEP 3, 2026
11
Boston Dynamics
bostondynamics.com
13 SIGNALS · LAST SEEN SEP 3, 2026
12
Stanford
16 SIGNALS · LAST SEEN AUG 28, 2026
13
UC San Diego
ucsd.edu
17 SIGNALS · LAST SEEN AUG 27, 2026
14
Carnegie Mellon University
cmu.edu
42 SIGNALS · LAST SEEN AUG 26, 2026
15
arXiv Physical AI
13 SIGNALS · LAST SEEN AUG 25, 2026
16
Harbin Institute of Technology
hit.edu.cn
6 SIGNALS · LAST SEEN AUG 21, 2026
17
RLBench
3 SIGNALS · LAST SEEN AUG 4, 2026
18
Shanghai Jiao Tong University
sjtu.edu.cn
36 SIGNALS · LAST SEEN JUL 30, 2026
19
Robotics Institute Germany
robotics.de
2 SIGNALS · LAST SEEN JUL 28, 2026
20
ALOHA
2 SIGNALS · LAST SEEN JUL 28, 2026
21
University of Maryland
6 SIGNALS · LAST SEEN JUL 27, 2026
22
Fudan TEAI Team
fudan.edu.cn
6 SIGNALS · LAST SEEN JUL 26, 2026
23
Honda Research Institute Europe GmbH
honda-ri.de
4 SIGNALS · LAST SEEN JUL 22, 2026
24
UT Austin
utexas.edu
1 SIGNAL · LAST SEEN JUL 16, 2026
25
KAIST
kaist.ac.kr
13 SIGNALS · LAST SEEN JUL 15, 2026
26
ThinkingVLA
2 SIGNALS · LAST SEEN JUL 5, 2026
27
TARS Robotics
6 SIGNALS · LAST SEEN JUL 2, 2026
28
Dream Labs
1 SIGNAL · LAST SEEN JUL 1, 2026
29
UC Berkeley
berkeley.edu
37 SIGNALS · LAST SEEN JUN 30, 2026
30
University of Washington
uw.edu
3 SIGNALS · LAST SEEN JUN 11, 2026
31
AIRe Lab
1 SIGNAL · LAST SEEN JUN 10, 2026
32
Harvard University
harvard.edu
8 SIGNALS · LAST SEEN JUN 9, 2026
33
National University of Singapore
nus.edu.sg
4 SIGNALS · LAST SEEN JUN 9, 2026
34
Cornell University
cornell.edu
4 SIGNALS · LAST SEEN JUN 8, 2026
35
Indian Institute of Science (IISc)
iisc.ac.in
2 SIGNALS · LAST SEEN JUN 5, 2026
36
Technical University of Munich
tum.de
5 SIGNALS · LAST SEEN JUN 5, 2026
37
University of Michigan
umich.edu
2 SIGNALS · LAST SEEN JUN 1, 2026
38
RoboTwin
1 SIGNAL · LAST SEEN JUN 1, 2026
39
TU Delft
0 SIGNALS · LAST SEEN MAY 29, 2026
40
SERL
1 SIGNAL · LAST SEEN MAY 28, 2026
41
Chinese Academy of Sciences Institute of Automation
ia.ac.cn
1 SIGNAL · LAST SEEN MAY 28, 2026
42
Vicarious
vicarious.com
1 SIGNAL · LAST SEEN MAY 25, 2026
43
Great Bay University
gbu.edu.cn
4 SIGNALS · LAST SEEN MAY 18, 2026
44
RLWRLD
rlwrld.ai
16 SIGNALS · LAST SEEN MAY 15, 2026
45
ETH Zurich
ethz.ch
3 SIGNALS · LAST SEEN MAY 14, 2026
46
HKUST (Guangzhou)
hkust-gz.edu.cn
1 SIGNAL · LAST SEEN MAY 4, 2026
47
Technical University of Darmstadt
tu-darmstadt.de
4 SIGNALS · LAST SEEN APR 30, 2026
48
Finite Robotics
0 SIGNALS · LAST SEEN JAN 1, 2024