arxiv:2607.02642
Hao Li
lh152
ยท
AI & ML interests
None yet
Recent Activity
upvoted a paper about 1 month ago
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement authored a paper about 1 month ago
GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation authored a paper about 1 month ago
ViVa: A Video-Generative Value Model for Robot Reinforcement Learning