Teahose.
SIGN IN
NEW HERE — WHAT TEAHOSE DOES
We read the entire AI & tech firehose — so you don't have to.
PODPodcastsAll-In, No Priors, Acquired…
NEWNewslettersStratechery, Newcomer…
PAPPapersPhysical AI research
PHProduct Huntdaily launches
VCInvestor ScoutSequoia, a16z, Benchmark…
CLAUDE DISTILLS →
7 reads, 30 sec each — free, 6 AM ET.
+ a live graph of the companies, people & themes underneath.
HOME/PEOPLE/CHANGWEN ZHENG
// PERSON

Changwen Zheng

AT CHINESE ACADEMY OF SCIENCESMENTIONS 1LAST SEEN AUGUST 13, 2026
// RECENT MENTIONS
// SIGNALS
1 SIGNAL
01
mention·arXiv Physical AI·AUGUST 13, 2026

Temporal GRPO keeps changes on the preceding stages close to zero and produces the largest positive improvement at md, indicating that it preserves acquired preceding behaviors and concentrates the update on the stage responsible for the rollout difference.

Source

AI-extracted from podcast / newsletter / paper summaries. May contain errors.

Changwen Zheng · Chinese Academy of Sciences — 1 mention on Teahose