--- license: llama3.2 license_link: https://huggingface.co/meta-llama/Llama-3.2-3B/blob/main/LICENSE.txt base_model: - meta-llama/Llama-3.2-3B - lianghsun/Llama-3.2-Taiwan-3B pipeline_tag: text-generation library_name: mlx language: - zh - en tags: - mlx - mlx-lm - 4-bit - quantized - llama - llama3.2 - Taiwan - ROC - zh-tw - continued-pretraining - text-generation --- # Llama-3.2-Taiwan-3B-Arbor-4bit `masato25/Llama-3.2-Taiwan-3B-Arbor-4bit` is an Apple MLX 4-bit quantized derivative of [`lianghsun/Llama-3.2-Taiwan-3B`](https://huggingface.co/lianghsun/Llama-3.2-Taiwan-3B). This was prepared in response to the GGUF-only repository [`QuantFactory/Llama-3.2-Taiwan-3B-GGUF`](https://huggingface.co/QuantFactory/Llama-3.2-Taiwan-3B-GGUF). The MLX conversion is made from the upstream safetensors model (`lianghsun/Llama-3.2-Taiwan-3B`), not by converting GGUF weights back to another format. This conversion is intended for local inference on Apple Silicon using MLX / MLX-LM. It preserves the upstream model architecture and tokenizer files, while quantizing supported linear weights to 4-bit for a smaller memory footprint. ## Attribution and license - Upstream Taiwan model: [`lianghsun/Llama-3.2-Taiwan-3B`](https://huggingface.co/lianghsun/Llama-3.2-Taiwan-3B) - Original base model: [`meta-llama/Llama-3.2-3B`](https://huggingface.co/meta-llama/Llama-3.2-3B) - Upstream license: Llama 3.2 Community License (`llama3.2`) - License file: - This repository is a derivative conversion/quantization and is not the original Meta Llama or Llama-3.2-Taiwan release. A `NOTICE` file is included in this repository. By using, copying, modifying, redistributing, deploying, or making this derivative model available to others, you are responsible for complying with all applicable upstream terms, including the Llama 3.2 Community License and any additional terms, notices, access requirements, usage instructions, export-control, sanctions, or other legal requirements that apply to the upstream models. If upstream licenses, notices, or model-page terms are updated, those upstream terms may impose additional or different obligations. Please review the upstream model pages and license before use or redistribution. ## No affiliation, sponsorship, endorsement, or trademark grant This repository is independently prepared and published by the repository owner. It is **not affiliated with, sponsored by, approved by, or endorsed by Meta, Llama, lianghsun, QuantFactory, or their affiliates** unless they explicitly state otherwise. The names "Llama", "Meta", "Llama-3.2-Taiwan", "lianghsun", "QuantFactory", and related marks are used here only for reasonable descriptive attribution and identification of the upstream source/base models. No trademark license or other rights in those marks are granted by this repository. ## Conversion details - Source: `lianghsun/Llama-3.2-Taiwan-3B` - Related GGUF request/source reference: `QuantFactory/Llama-3.2-Taiwan-3B-GGUF` - Format: MLX / MLX-LM - Quantization: 4-bit - Group size: 64 Equivalent conversion command: ```bash python -m mlx_lm convert \ --hf-path lianghsun/Llama-3.2-Taiwan-3B \ --mlx-path ./Llama-3.2-Taiwan-3B-Arbor-4bit \ -q \ --q-bits 4 \ --q-group-size 64 ``` ## Usage Install MLX-LM: ```bash pip install -U mlx-lm ``` Generate text: ```bash mlx_lm.generate \ --model masato25/Llama-3.2-Taiwan-3B-Arbor-4bit \ --prompt "中華民國憲法第一條" ``` Python example: ```python from mlx_lm import load, generate model, tokenizer = load("masato25/Llama-3.2-Taiwan-3B-Arbor-4bit") prompt = "中華民國憲法第一條" response = generate(model, tokenizer, prompt=prompt, max_tokens=128) print(response) ``` ## Important safety, quality, and compliance notes Quantization may change model behavior, quality, robustness, refusal behavior, calibration, and safety characteristics compared with the upstream model. No claim is made that this derivative is equivalent to, safer than, or better than the upstream model. Large language model outputs may be inaccurate, unsafe, biased, offensive, incomplete, or unsuitable for your intended use. You are responsible for evaluating outputs and for complying with applicable laws, licenses, policies, and platform rules. ## Warranty and liability disclaimer This derivative model and associated files are provided **as-is** and **without warranties or conditions of any kind**, express or implied, including without limitation warranties of merchantability, fitness for a particular purpose, title, non-infringement, accuracy, availability, or error-free operation. To the maximum extent permitted by applicable law, the repository owner is not liable for any direct, indirect, incidental, special, consequential, exemplary, punitive, or other damages arising from or related to use of this repository, the derivative model, or model outputs. Nothing in this README is legal advice. You are responsible for reviewing and complying with the applicable license terms and laws.