Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Hanyang Wang's picture

Hanyang Wang

ssyhw7
5

AI & ML interests

None yet

Recent Activity

upvoted a paper 9 days ago
DivOPD: Spread Wide, Look Close for Asynchronous On-Policy Distillation of Multi-turn Agents
upvoted a paper 16 days ago
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses
upvoted a paper 3 months ago
BiPACE: Bisimulation-Guided Policy Optimization with Action Counterfactual Estimation for LLM Agents
View all activity

Organizations

None yet

upvoted a paper 9 days ago

DivOPD: Spread Wide, Look Close for Asynchronous On-Policy Distillation of Multi-turn Agents

Paper • 2609.34838 • Published 11 days ago • 1
upvoted a paper 16 days ago

RRSI: Regularized Recursive Self-Improvement of Agent Harnesses

Paper • 2609.24972 • Published 18 days ago • 223
upvoted a paper 3 months ago

BiPACE: Bisimulation-Guided Policy Optimization with Action Counterfactual Estimation for LLM Agents

Paper • 2606.25556 • Published Jun 24 • 1
upvoted a paper 6 months ago

The Detection--Extraction Gap: Models Know the Answer Before They Can Say It

Paper • 2604.06613 • Published Apr 8 • 2
upvoted a paper 8 months ago

SkillRL: Evolving Agents via Recursive Skill-Augmented Reinforcement Learning

Paper • 2602.08234 • Published Feb 9 • 76
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs