Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

Duplicated from  meituan-longcat/LongCat-AudioDiT-3.5B

9r4n4y
/
LongCat-AudioDiT-3.5B-tts-text-to-speech-SOTA

Text-to-Speech
Safetensors
Chinese
English
audiodit
tts
voice-cloning
voice-conversion
zero-shot-tts
SOTA
diffusion-transformer
speech-synthesis
audio-generation
neural-tts
commercial-use
Model card Files Files and versions
xet
Community
1
LongCat-AudioDiT-3.5B-tts-text-to-speech-SOTA
15.3 GB
Ctrl+K
Ctrl+K
  • 2 contributors
History: 3 commits
opemipotijani1's picture
opemipotijani1
Upload LICENSE
f15d16c verified about 2 months ago
  • .gitattributes
    1.57 kB
    Duplicate from meituan-longcat/LongCat-AudioDiT-3.5B 2 months ago
  • LICENSE
    34.5 kB
    Upload LICENSE about 2 months ago
  • LongCat-AudioDiT.svg
    53.1 kB
    Duplicate from meituan-longcat/LongCat-AudioDiT-3.5B 2 months ago
  • README.md
    12.4 kB
    Update README.md 2 months ago
  • architecture.png
    318 kB
    xet
    Duplicate from meituan-longcat/LongCat-AudioDiT-3.5B 2 months ago
  • config.json
    2.47 kB
    Duplicate from meituan-longcat/LongCat-AudioDiT-3.5B 2 months ago
  • model.safetensors
    15.3 GB
    xet
    Duplicate from meituan-longcat/LongCat-AudioDiT-3.5B 2 months ago