Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Zhehao Zhang's picture

Zhehao Zhang

ZhehaoZhang
1 6 4
·
https://zzh-sjtu.github.io/zhehaozhang.github.io/

AI & ML interests

NLP

Organizations

OSU NLP Group's profile picture Amazon Science's profile picture

upvoted a collection 3 months ago

Tulu3 with distraction mitigation data

Collection
LLM and LRM can be easily distracted by hidden instructions or irrelevant tasks. We curated SFT and DPO data that model can finetune to avoid distract • 5 items • Updated Oct 30, 2025 • 3
upvoted a paper 5 months ago

QUEST: Training Frontier Deep Research Agents with Fully Synthetic Tasks

Paper • 2605.24218 • Published May 22 • 44
upvoted a paper 6 months ago

When Benign Inputs Lead to Severe Harms: Eliciting Unsafe Unintended Behaviors of Computer-Use Agents

Paper • 2602.08235 • Published Feb 9 • 2
upvoted a paper 7 months ago

CUA-Suite: Massive Human-annotated Video Demonstrations for Computer-Use Agents

Paper • 2603.24440 • Published Mar 25 • 98
upvoted a paper about 1 year ago

LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model

Paper • 2509.00676 • Published Aug 31, 2025 • 85
upvoted a paper almost 2 years ago

Personalization of Large Language Models: A Survey

Paper • 2411.00027 • Published Oct 29, 2024 • 33
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs