// PERSON
Hang Gao
AT CHINESE ACADEMY OF SCIENCESMENTIONS 1LAST SEEN AUGUST 13, 2026
// RECENT MENTIONS
// SIGNALS
1 SIGNAL
01
mention·arXiv Physical AI·AUGUST 13, 2026
“We characterize and formulate trajectory-level credit aliasing in outcome-driven VLA reinforcement learning, where rollouts with different stage progress can receive the same final-outcome advantage, causing successful preceding actions and later failed actions to be updated with the same advantage.”
Source→AI-extracted from podcast / newsletter / paper summaries. May contain errors.