arxiv:2606.14885
ZhuofengLi
ZhuofengLi
AI & ML interests
Agents, Reasoning LLMs/VLLMs, RL
Recent Activity
upvoted a paper about 10 hours ago
World Editing: Intervening on Executable Worlds at Increasing Depth upvoted a paper 7 days ago
EasyPPO: Stabilizing the Critic Is Key updated a dataset 23 days ago
ZhuofengLi/Harbor-SWE-Trajectory