Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
LuYanFCP's picture

LuYanFCP

LuYanFCP123
9

AI & ML interests

None yet

Organizations

None yet

upvoted a paper 2 months ago

DVAO: Dynamic Variance-adaptive Advantage Optimization for Multi-reward Reinforcement Learning

Paper • 2605.25604 • Published May 25 • 139
upvoted 8 papers 10 months ago

When Judgment Becomes Noise: How Design Failures in LLM Judge Benchmarks Silently Undermine Validity

Paper • 2509.20293 • Published Sep 24, 2025 • 8

Thinking While Listening: Simple Test Time Scaling For Audio Classification

Paper • 2509.19676 • Published Sep 24, 2025 • 5

MOSS-ChatV: Reinforcement Learning with Process Reasoning Reward for Video Temporal Reasoning

Paper • 2509.21113 • Published Sep 25, 2025 • 6

Behind RoPE: How Does Causal Mask Encode Positional Information?

Paper • 2509.21042 • Published Sep 25, 2025 • 10

Quantized Visual Geometry Grounded Transformer

Paper • 2509.21302 • Published Sep 25, 2025 • 9

ScaleDiff: Scaling Difficult Problems for Advanced Mathematical Reasoning

Paper • 2509.21070 • Published Sep 25, 2025 • 9

Residual Off-Policy RL for Finetuning Behavior Cloning Policies

Paper • 2509.19301 • Published Sep 23, 2025 • 20

VCRL: Variance-based Curriculum Reinforcement Learning for Large Language Models

Paper • 2509.19803 • Published Sep 24, 2025 • 122
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs