Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Xiangxin Zhou's picture

Xiangxin Zhou

zhouxiangxin
5 26 4
Datawitch-Programmer's profile picture 21world's profile picture Gargaz's profile picture
·
https://zhouxiangxin1998.github.io/

AI & ML interests

None yet

Recent Activity

upvoted a paper 30 days ago
OS Agents: A Survey on MLLM-based Agents for General Computing Devices Use
upvoted a paper 30 days ago
Exploring the Design Space of Reward Backpropagation for Flow Matching
upvoted a paper 30 days ago
TempAct: Advancing Temporal Plausibility in Autoregressive Video Generation via Planner-Executor RL
View all activity

Organizations

ezetimibe's profile picture benzeneRing's profile picture ProteinBench's profile picture Axon RL's profile picture AIFirstScience's profile picture substance0723's profile picture cruise0724's profile picture substance0724's profile picture bioterminal's profile picture Tencent-Hunyuan-Multimodal-RL's profile picture harnessRL's profile picture harnessRL2's profile picture

commented 2 papers 2 months ago

Rethinking the Divergence Regularization in LLM RL

Paper • 2606.09821 • Published Jun 8 • 34 •
4

Rethinking the Divergence Regularization in LLM RL

Paper • 2606.09821 • Published Jun 8 • 34 •
4
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs