Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Xiaozhe Yao's picture

Xiaozhe Yao

xzyao
5 7 68
21world's profile picture dipankarsarkar's profile picture adamm-hf's profile picture
·

AI & ML interests

None yet

Recent Activity

liked a model 26 days ago
incoai/GLM-5.3-Flash-DFlash2
updated a bucket 2 months ago
researchcomputer/kernels
published a bucket 2 months ago
opentela-ai/kernels
View all activity

Organizations

Research Computer's profile picture DS3Lab's profile picture AutoAI's profile picture ICML2023's profile picture ETH Zurich's profile picture eth-easl's profile picture Compressed LMs's profile picture Xiaozhe Yao and Friends's profile picture DeltaZip's profile picture vagents's profile picture Scaling VIT's profile picture Apertus Community's profile picture OpenTela's profile picture Benchmaker's profile picture

upvoted a paper over 1 year ago

Qwen2.5-Omni Technical Report

Paper • 2503.20215 • Published Mar 26, 2025 • 174
upvoted a paper almost 2 years ago

Unpacking SDXL Turbo: Interpreting Text-to-Image Models with Sparse Autoencoders

Paper • 2410.22366 • Published Oct 28, 2024 • 84
upvoted 5 papers over 2 years ago

WARP: On the Benefits of Weight Averaged Rewarded Policies

Paper • 2406.16768 • Published Jun 24, 2024 • 23

Decoding Compressed Trust: Scrutinizing the Trustworthiness of Efficient LLMs Under Compression

Paper • 2403.15447 • Published Mar 18, 2024 • 15

When Scaling Meets LLM Finetuning: The Effect of Data, Model and Finetuning Method

Paper • 2402.17193 • Published Feb 27, 2024 • 27

How to Train Data-Efficient LLMs

Paper • 2402.09668 • Published Feb 15, 2024 • 43

DeltaZip: Multi-Tenant Language Model Serving via Delta Compression

Paper • 2312.05215 • Published Dec 8, 2023 • 1
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs