Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
EJVision 's Collections
Core models
AudioUnderstanding
Mashups
Translation applications
AIDub
BookReader
Text_to_image
Interesting models to try
Programming
Embedding models
Diarization
Defense against the Dark Arts
Prompt collections
Reading models
MoE
Video
Hybrids
Translation models
TTS Models
Translation prompts

Diarization

updated Jun 5
Upvote
-

  • pyannote/segmentation-3.0

    Voice Activity Detection • Updated May 10, 2024 • 5.42M • 2.12k

  • pyannote/speaker-diarization-3.1

    Automatic Speech Recognition • Updated May 10, 2024 • 7.15M • 4.19k

  • pyannote/speaker-diarization-community-1

    Automatic Speech Recognition • Updated Sep 29, 2025 • 5.55M • 2.61k

  • ibm-granite/granite-speech-4.1-2b-plus

    Automatic Speech Recognition • 2B • Updated Jun 16 • 77.3k • 94

  • nvidia/parakeet-tdt-0.6b-v3

    Automatic Speech Recognition • 0.6B • Updated Aug 5 • 596k • • 1.18k

  • ACE-Step/acestep-transcriber

    Audio-Text-to-Text • 11B • Updated Feb 3 • 3.16k • 65
Upvote
-
  • Collection guide
  • Browse collections
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs