Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
HHY's picture

HHY

Jaderoof
11 9
chriszhouwei's profile picture
·

AI & ML interests

None yet

Organizations

meituan's profile picture

upvoted a paper 3 months ago

V_{0.5}: Generalist Value Model as a Prior for Sparse RL Rollouts

Paper • 2603.10848 • Published Mar 11 • 17
upvoted a paper 5 months ago

Self-Distilled Agentic Reinforcement Learning

Paper • 2605.15155 • Published May 14 • 118
upvoted a paper 7 months ago

QuantaAlpha: An Evolutionary Framework for LLM-Driven Alpha Mining

Paper • 2602.07085 • Published Feb 6 • 101
upvoted 5 papers 8 months ago

ScaleEnv: Scaling Environment Synthesis from Scratch for Generalist Interactive Tool-Use Agent Training

Paper • 2602.06820 • Published Feb 6 • 15

CoBA-RL: Capability-Oriented Budget Allocation for Reinforcement Learning in LLMs

Paper • 2602.03048 • Published Feb 3 • 32

Look Back to Reason Forward: Revisitable Memory for Long-Context LLM Agents

Paper • 2509.23040 • Published Sep 27, 2025 • 12

V_0: A Generalist Value Model for Any Policy at State Zero

Paper • 2602.03584 • Published Feb 3 • 22

LongCat-Flash-Thinking-2601 Technical Report

Paper • 2601.16725 • Published Jan 23 • 186
upvoted an article over 1 year ago
view article
Article

Visualize and understand GPU memory in PyTorch

qgallouedec
•
Dec 24, 2024
• 275
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs