Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
咘噜咘噜的碰's picture

咘噜咘噜的碰

mysxs
1 11
·
https://github.com/mysxs
  • mysxs

AI & ML interests

speech

Recent Activity

upvoted a paper 1 day ago
StepAudio 3 Gen Technical Report
liked a model about 1 year ago
microsoft/VibeVoice-1.5B
liked a model about 1 year ago
SparkAudio/Spark-TTS-0.5B
View all activity

Organizations

None yet

liked 3 models about 1 year ago

microsoft/VibeVoice-1.5B

Text-to-Speech • 3B • Updated Jan 22 • 401k • 2.48k

SparkAudio/Spark-TTS-0.5B

Text-to-Speech • Updated Mar 7, 2025 • 933 • 747

fishaudio/s1-mini

Text-to-Speech • Updated Feb 6 • 2.13k • 709
liked a Space about 1 year ago
Running on CPU Upgrade
Featured
974

TTS Arena V2

🗣
974

Compare two text-to-speech voices and vote for the better

liked 5 models over 1 year ago

fixie-ai/ultravox-v0_5-llama-3_2-1b

Audio-Text-to-Text • 0.7B • Updated Mar 11 • 132k • 96

Qwen/Qwen2.5-Omni-7B

Any-to-Any • 11B • Updated Apr 30, 2025 • 341k • 1.94k

ASLP-lab/OSUM

Updated Feb 16, 2025 • 12

nari-labs/Dia-1.6B

Text-to-Speech • 2B • Updated Jun 1, 2025 • 29k • • 2.91k

kyutai/mimi

Feature Extraction • 96.2M • Updated Jul 2, 2025 • 493k • • 326
liked 2 models almost 2 years ago

amphion/hifigan_speech_bigdata

Updated Dec 21, 2023 • 4

pyannote/segmentation

Voice Activity Detection • Updated May 10, 2024 • 1.97M • 696
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs