tomer-nv commited on
Commit
c743c5e
·
1 Parent(s): f52cc60

Update model card details

Browse files
Files changed (1) hide show
  1. README.md +4 -3
README.md CHANGED
@@ -94,8 +94,8 @@ Global<br>
94
  ### Use Case: <br>
95
  NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4 is a general purpose reasoning and chat model intended to be used in English, Code, and supported multilingual contexts. This model is optimized for collaborative agents and high-volume workloads. It is intended to be used by developers designing AI Agent systems, chatbots, RAG systems, and other AI-powered applications. This model is also suitable for complex instruction-following tasks and long-context reasoning.
96
 
97
- ### Release Date [Insert the expected release date below]: <br>
98
- June 12, 2026 via [Hugging Face](https://huggingface.co/nvidia/NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4)
99
 
100
  ## References(s):
101
  * [\[2411.19146\] Puzzle: Distillation-Based NAS for Inference-Optimized LLMs](https://arxiv.org/abs/2411.19146)
@@ -613,7 +613,8 @@ The GitHub Crawl was collected using the GitHub REST API and the Amazon S3 API.
613
  * **Labeling Method by dataset**: Hybrid: Automated, Human, Synthetic
614
 
615
  ## Inference:
616
- * **Acceleration Engine:** vLLM
 
617
  **Test Hardware:**
618
  - 1× NVIDIA H100-80GB
619
  - 8× NVIDIA H100-80GB
 
94
  ### Use Case: <br>
95
  NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4 is a general purpose reasoning and chat model intended to be used in English, Code, and supported multilingual contexts. This model is optimized for collaborative agents and high-volume workloads. It is intended to be used by developers designing AI Agent systems, chatbots, RAG systems, and other AI-powered applications. This model is also suitable for complex instruction-following tasks and long-context reasoning.
96
 
97
+ ### Release Date: <br>
98
+ July 6, 2026 via [Hugging Face](https://huggingface.co/nvidia/NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4)
99
 
100
  ## References(s):
101
  * [\[2411.19146\] Puzzle: Distillation-Based NAS for Inference-Optimized LLMs](https://arxiv.org/abs/2411.19146)
 
613
  * **Labeling Method by dataset**: Hybrid: Automated, Human, Synthetic
614
 
615
  ## Inference:
616
+ **Acceleration Engine:** vLLM
617
+
618
  **Test Hardware:**
619
  - 1× NVIDIA H100-80GB
620
  - 8× NVIDIA H100-80GB