Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
nota-ai
/
Qwen3.5-4B-QAD-W4A16
like
2
Follow
Nota AI
154
Text Generation
Safetensors
qwen3_5
qwen3.5
quantization
int4
w4a16
compressed-tensors
quantization-aware-distillation
efficient-inference
conversational
arxiv:
2607.04244
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
Qwen3.5-4B-QAD-W4A16
Ctrl+K
Ctrl+K
2 contributors
History:
1 commit
jykim310
bokyeong1015
Model upload
a67b0fe
23 days ago
figure
Model upload
23 days ago
.gitattributes
Safe
1.61 kB
Model upload
23 days ago
README.md
Safe
3.39 kB
Model upload
23 days ago
chat_template.jinja
Safe
7.76 kB
Model upload
23 days ago
config.json
14.9 kB
Model upload
23 days ago
merges.txt
Safe
3.35 MB
Model upload
23 days ago
model.safetensors
4.19 GB
xet
Model upload
23 days ago
preprocessor_config.json
Safe
390 Bytes
Model upload
23 days ago
processor_config.json
Safe
1.19 kB
Model upload
23 days ago
tokenizer.json
Safe
20 MB
xet
Model upload
23 days ago
tokenizer_config.json
Safe
1.14 kB
Model upload
23 days ago
video_preprocessor_config.json
Safe
385 Bytes
Model upload
23 days ago
vocab.json
Safe
6.72 MB
Model upload
23 days ago