Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Duplicated from
meituan-longcat/LongCat-AudioDiT-3.5B
9r4n4y
/
LongCat-AudioDiT-3.5B-tts-text-to-speech-SOTA
like
14
Text-to-Speech
Safetensors
Chinese
English
audiodit
tts
voice-cloning
voice-conversion
zero-shot-tts
SOTA
diffusion-transformer
speech-synthesis
audio-generation
neural-tts
commercial-use
License:
mit
Model card
Files
Files and versions
xet
Community
1
Copy to bucket
new
main
LongCat-AudioDiT-3.5B-tts-text-to-speech-SOTA
15.3 GB
Ctrl+K
Ctrl+K
2 contributors
History:
2 commits
9r4n4y
Update README.md
a83966e
verified
2 months ago
.gitattributes
Safe
1.57 kB
Duplicate from meituan-longcat/LongCat-AudioDiT-3.5B
2 months ago
LICENSE
Safe
1.06 kB
Duplicate from meituan-longcat/LongCat-AudioDiT-3.5B
2 months ago
LongCat-AudioDiT.svg
Safe
53.1 kB
Duplicate from meituan-longcat/LongCat-AudioDiT-3.5B
2 months ago
README.md
Safe
12.4 kB
Update README.md
2 months ago
architecture.png
Safe
318 kB
xet
Duplicate from meituan-longcat/LongCat-AudioDiT-3.5B
2 months ago
config.json
Safe
2.47 kB
Duplicate from meituan-longcat/LongCat-AudioDiT-3.5B
2 months ago
model.safetensors
Safe
15.3 GB
xet
Duplicate from meituan-longcat/LongCat-AudioDiT-3.5B
2 months ago