Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🔄
In a Training Loop
Jerry Pan
JERRYPAN617
3
25
Follow
webbrain-one-240398's profile picture
1 follower
·
12 following
https://jerrypan617.github.io/
jerrypan617
AI & ML interests
Latent Space Reasoning, VLM, Test-Time Scaling
Organizations
JERRYPAN617
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
a dataset
3 months ago
AgentVQA/AgentVQA
Viewer
•
Updated
Dec 18, 2025
•
20.5k
•
212
•
1
liked
a dataset
4 months ago
TACPS-liv/Spatial-DISE
Viewer
•
Updated
Jul 22
•
12.4k
•
859
•
5
liked
5 datasets
11 months ago
PKU-Alignment/PKU-SafeRLHF-single-dimension
Viewer
•
Updated
Jun 14, 2024
•
81.1k
•
170
•
3
PKU-Alignment/PKU-SafeRLHF
Viewer
•
Updated
Oct 18, 2024
•
164k
•
14.1k
•
198
HuggingFaceH4/ultrafeedback_binarized
Viewer
•
Updated
Oct 16, 2024
•
187k
•
18.2k
•
348
LooksJuicy/ruozhiba
Viewer
•
Updated
Apr 9, 2024
•
1.5k
•
195
•
325
Karsh-CAI/btfChinese-DPO-small
Viewer
•
Updated
Apr 7, 2024
•
5k
•
10
•
23
liked
2 models
11 months ago
Qwen/Qwen2.5-1.5B-Instruct
Text Generation
•
2B
•
Updated
Sep 25, 2024
•
6.59M
•
•
870
JERRYPAN617/HH-BTRewardModel-roberta
Reinforcement Learning
•
0.1B
•
Updated
Nov 13, 2025
•
17
•
1
liked
7 datasets
11 months ago
ys-zong/VLGuard
Viewer
•
Updated
Jan 19, 2025
•
3k
•
355
•
19
PKU-Alignment/MM-SafetyBench
Viewer
•
Updated
Sep 19, 2024
•
6.72k
•
1.9k
•
8
saferlhf-v/BeaverTails-V
Viewer
•
Updated
Mar 8, 2025
•
30.4k
•
636
•
7
PKU-Alignment/PKU-SafeRLHF-V
Viewer
•
Updated
Mar 25, 2025
•
30.4k
•
300
•
6
Moemu/Muice-Dataset
Viewer
•
Updated
May 18
•
3.74k
•
388
•
63
liuhaotian/LLaVA-Instruct-150K
Preview
•
Updated
Jan 3, 2024
•
4.35k
•
637
MMMU/MMMU
Viewer
•
Updated
5 days ago
•
11.6k
•
76.3k
•
336
liked
a Space
11 months ago
Sleeping
Agents
1
Qwen2.5 Psydoctor Demo
📈
1
基于 Qwen2.5-1.5B-Instruct 模型微调的 LoRA 适配器,专门用于心理医生对话场景。
liked
2 datasets
12 months ago
FreedomIntelligence/medical-o1-reasoning-SFT
Viewer
•
Updated
Apr 22, 2025
•
90.1k
•
21.2k
•
1.19k
nvidia/Nemotron-CC-Math-v1
Viewer
•
Updated
Dec 23, 2025
•
190M
•
13.7k
•
101
liked
a model
12 months ago
JERRYPAN617/qwen2.5-lora-psydoctor
Text Generation
•
Updated
Oct 25, 2025
•
16
•
1
Load more