Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Zhicheng Cai's picture

Zhicheng Cai

Aiolus-X
6

AI & ML interests

LLM&Agentic RL

Recent Activity

authored a paper 8 days ago
Beyond Euclidean Clipping: Overcoming Exploration Collapse in LLM RL via Riemannian Isometric Policy Optimization
authored a paper 8 days ago
FLEX: Continuous Agent Evolution via Forward Learning from Experience
authored a paper 8 days ago
Enigmata: Scaling Logical Reasoning in Large Language Models with Synthetic Verifiable Puzzles
View all activity

Organizations

None yet

upvoted a paper 8 days ago

Beyond Euclidean Clipping: Overcoming Exploration Collapse in LLM RL via Riemannian Isometric Policy Optimization

Paper • 2607.10169 • Published 20 days ago • 14
upvoted a paper 14 days ago

Spectral Rewiring for Exploration, Purification, and Model Merging

Paper • 2607.03065 • Published 28 days ago • 25
upvoted a paper 17 days ago

Weak-to-Strong Generalization via Direct On-Policy Distillation

Paper • 2607.05394 • Published 23 days ago • 139
upvoted 3 papers 6 months ago

Learning to Discover at Test Time

Paper • 2601.16175 • Published Jan 22 • 45

LLM-in-Sandbox Elicits General Agentic Intelligence

Paper • 2601.16206 • Published Jan 22 • 87

FLEX: Continuous Agent Evolution via Forward Learning from Experience

Paper • 2511.06449 • Published Nov 9, 2025 • 14
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs