Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
fafa's picture
๐Ÿ”„ In a Training Loop

fafa

kunkun0919
6
Lpc206's profile picture
ยท
https://github.com/QuZikun/

AI & ML interests

LLM

Recent Activity

upvoted a paper 11 days ago
COBRA-Skills: Contextual Bandit-Guided Evolution for Agent Skill Optimization
upvoted a paper 21 days ago
Rethinking On-Policy Distillation of Large Language Models II: One Training Example
authored a paper about 1 month ago
SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation
View all activity

Organizations

None yet
kunkun0919 's papers 2
arxiv:2608.04419
arxiv:2509.24696
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs