Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Peilin Wu's picture

Peilin Wu

qualidea1217
2 1

AI & ML interests

None yet

Recent Activity

upvoted a paper about 1 hour ago
BiasReducer: Adaptive Bias Mitigation for Reward Models
upvoted a paper 14 days ago
An Empirical Study of Harness Design for Coding Agents
published a dataset 10 months ago
qualidea1217/RaR-Medicine-20k-o3-mini-converted
View all activity

Organizations

None yet

qualidea1217 's models 10

qualidea1217/Qwen2.5-7B-Instruct-PPO-HiPRAG

8B • Updated Oct 8, 2025 • 43

qualidea1217/Qwen2.5-7B-Instruct-GRPO-HiPRAG

8B • Updated Oct 8, 2025 • 12

qualidea1217/Qwen2.5-3B-PPO-HiPRAG

3B • Updated Oct 8, 2025 • 9

qualidea1217/Qwen2.5-3B-Instruct-PPO-HiPRAG-Undersearch-Only

3B • Updated Oct 8, 2025 • 4

qualidea1217/Qwen2.5-3B-Instruct-PPO-HiPRAG-Lambda_p-0.6

3B • Updated Oct 8, 2025 • 5

qualidea1217/Qwen2.5-3B-Instruct-PPO-HiPRAG-Lambda_p-0.2

3B • Updated Oct 8, 2025 • 7

qualidea1217/Qwen2.5-3B-Instruct-PPO-HiPRAG

3B • Updated Oct 8, 2025 • 21

qualidea1217/Qwen2.5-3B-Instruct-PPO-HiPRAG-Oversearch-Only

3B • Updated Oct 7, 2025 • 17

qualidea1217/Qwen2.5-3B-Instruct-GRPO-HiPRAG

3B • Updated Oct 7, 2025 • 138

qualidea1217/Llama-3.2-3B-Instruct-PPO-HiPRAG

4B • Updated Oct 7, 2025 • 8
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs