Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
carrie's picture

carrie

ttira
7

AI & ML interests

None yet

Organizations

None yet

upvoted an article about 1 year ago
view article
Article

Introduction to MedVideoCap-55K: A New, Large-Scale, High-Quality Medical Video-Caption Pair Dataset

wangrongsheng
•
Jun 25, 2025
• 11
upvoted 2 papers about 1 year ago

MedGen: Unlocking Medical Video Generation by Scaling Granularly-annotated Medical Videos

Paper • 2507.05675 • Published Jul 8, 2025 • 27

ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation

Paper • 2506.18095 • Published Jun 22, 2025 • 67
upvoted 4 papers over 1 year ago

Soundwave: Less is More for Speech-Text Alignment in LLMs

Paper • 2502.12900 • Published Feb 18, 2025 • 85

Enabling Scalable Oversight via Self-Evolving Critic

Paper • 2501.05727 • Published Jan 10, 2025 • 72

On the Compositional Generalization of Multimodal LLMs for Medical Imaging

Paper • 2412.20070 • Published Dec 28, 2024 • 42

HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs

Paper • 2412.18925 • Published Dec 25, 2024 • 107
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs