Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Peilin Wu
qualidea1217
2
1
Follow
AI & ML interests
None yet
Recent Activity
upvoted
a
paper
about 1 hour ago
BiasReducer: Adaptive Bias Mitigation for Reward Models
upvoted
a
paper
14 days ago
An Empirical Study of Harness Design for Coding Agents
published
a dataset
10 months ago
qualidea1217/RaR-Medicine-20k-o3-mini-converted
View all activity
Organizations
None yet
qualidea1217
's models
10
Sort: Recently updated
qualidea1217/Qwen2.5-7B-Instruct-PPO-HiPRAG
8B
•
Updated
Oct 8, 2025
•
43
qualidea1217/Qwen2.5-7B-Instruct-GRPO-HiPRAG
8B
•
Updated
Oct 8, 2025
•
12
qualidea1217/Qwen2.5-3B-PPO-HiPRAG
3B
•
Updated
Oct 8, 2025
•
9
qualidea1217/Qwen2.5-3B-Instruct-PPO-HiPRAG-Undersearch-Only
3B
•
Updated
Oct 8, 2025
•
4
qualidea1217/Qwen2.5-3B-Instruct-PPO-HiPRAG-Lambda_p-0.6
3B
•
Updated
Oct 8, 2025
•
5
qualidea1217/Qwen2.5-3B-Instruct-PPO-HiPRAG-Lambda_p-0.2
3B
•
Updated
Oct 8, 2025
•
7
qualidea1217/Qwen2.5-3B-Instruct-PPO-HiPRAG
3B
•
Updated
Oct 8, 2025
•
21
qualidea1217/Qwen2.5-3B-Instruct-PPO-HiPRAG-Oversearch-Only
3B
•
Updated
Oct 7, 2025
•
17
qualidea1217/Qwen2.5-3B-Instruct-GRPO-HiPRAG
3B
•
Updated
Oct 7, 2025
•
138
qualidea1217/Llama-3.2-3B-Instruct-PPO-HiPRAG
4B
•
Updated
Oct 7, 2025
•
8