Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
nengo 's Collections
sentiment
embeddings
vision
video
articles
coding
small models
roleplay models
imagen
erp
voice
gradio-themes
spaces
top
sound
presets
testing
datasets
base

sound

updated May 1
Upvote
-

  • stabilityai/stable-audio-open-1.0

    Text-to-Audio • 1B • Updated Jun 19, 2025 • 18.9k • 1.55k

  • Systran/faster-whisper-large-v3

    Automatic Speech Recognition • Updated Nov 23, 2023 • 1.2M • 629

  • kyutai/tts-1.6b-en_fr

    Text-to-Speech • Updated Sep 11, 2025 • 69.3k • 378

  • ACE-Step/acestep-v15-xl-base

    Text-to-Audio • 5B • Updated Apr 7 • 1.7k • 99

  • smcleod/parakeet-tdt-0.6b-v2-int8

    Automatic Speech Recognition • Updated Dec 17, 2025 • 14 • 3

  • nvidia/nemotron-speech-streaming-en-0.6b

    Automatic Speech Recognition • 0.6B • Updated 28 days ago • 163k • 598

  • nvidia/canary-qwen-2.5b

    Automatic Speech Recognition • 3B • Updated Apr 21 • 47.1k • 451

  • CohereLabs/cohere-transcribe-03-2026

    Automatic Speech Recognition • 2B • Updated Jun 10 • 1M • • 1.07k

  • cstr/canary-1b-v2-GGUF

    Automatic Speech Recognition • 1.0B • Updated 1 day ago • 2.49k • 2
Upvote
-
  • Collection guide
  • Browse collections
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs