Text Generation
GGUF
English
Māori
llama.cpp
abteex-ai-labs
aotearoa
general
local-first
lumynax
new-zealand
qwen
sovereign-ai
text
vllm
vllm-compatible
vllm-experimental
nvidia-nim
nim-compatible
nim-candidate
nvidia-nemo
nem
nvidia-nemo-pathway
nem-pathway
nem-convert-required
conversational
Instructions to use AbteeXAILab/lumynax-infused-qwen3-17b-gguf with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- llama.cpp
How to use AbteeXAILab/lumynax-infused-qwen3-17b-gguf with llama.cpp:
Install (macOS, Linux)
curl -LsSf https://llama.app/install.sh | sh # Start a local OpenAI-compatible server with a web UI: llama serve -hf AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0 # Run inference directly in the terminal: llama cli -hf AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0
Install from WinGet (Windows)
winget install llama.cpp # Start a local OpenAI-compatible server with a web UI: llama serve -hf AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0 # Run inference directly in the terminal: llama cli -hf AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0
Use pre-built binary
# Download pre-built binary from: # https://github.com/ggerganov/llama.cpp/releases # Start a local OpenAI-compatible server with a web UI: ./llama-server -hf AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0 # Run inference directly in the terminal: ./llama-cli -hf AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0
Build from source code
git clone https://github.com/ggerganov/llama.cpp.git cd llama.cpp cmake -B build cmake --build build -j --target llama-server llama-cli # Start a local OpenAI-compatible server with a web UI: ./build/bin/llama-server -hf AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0 # Run inference directly in the terminal: ./build/bin/llama-cli -hf AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0
Use Docker
docker model run hf.co/AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0
- LM Studio
- Jan
- vLLM
How to use AbteeXAILab/lumynax-infused-qwen3-17b-gguf with vLLM:
Install from pip and serve model
# Install vLLM from pip: pip install vllm # Start the vLLM server: vllm serve "AbteeXAILab/lumynax-infused-qwen3-17b-gguf" # Call the server using curl (OpenAI-compatible API): curl -X POST "http://localhost:8000/v1/chat/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "AbteeXAILab/lumynax-infused-qwen3-17b-gguf", "messages": [ { "role": "user", "content": "What is the capital of France?" } ] }'Use Docker
docker model run hf.co/AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0
- Ollama
How to use AbteeXAILab/lumynax-infused-qwen3-17b-gguf with Ollama:
ollama run hf.co/AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0
- Unsloth Studio
How to use AbteeXAILab/lumynax-infused-qwen3-17b-gguf with Unsloth Studio:
Install Unsloth Studio (macOS, Linux, WSL)
curl -fsSL https://unsloth.ai/install.sh | sh # Run unsloth studio unsloth studio -H 0.0.0.0 -p 8888 # Then open http://localhost:8888 in your browser # Search for AbteeXAILab/lumynax-infused-qwen3-17b-gguf to start chatting
Install Unsloth Studio (Windows)
irm https://unsloth.ai/install.ps1 | iex # Run unsloth studio unsloth studio -H 0.0.0.0 -p 8888 # Then open http://localhost:8888 in your browser # Search for AbteeXAILab/lumynax-infused-qwen3-17b-gguf to start chatting
Using HuggingFace Spaces for Unsloth
# No setup required # Open https://huggingface.co/spaces/unsloth/studio in your browser # Search for AbteeXAILab/lumynax-infused-qwen3-17b-gguf to start chatting
- Pi
How to use AbteeXAILab/lumynax-infused-qwen3-17b-gguf with Pi:
Start the llama.cpp server
# Install llama.cpp: brew install llama.cpp # Start a local OpenAI-compatible server: llama serve -hf AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0
Configure the model in Pi
# Install Pi: npm install -g @mariozechner/pi-coding-agent # Add to ~/.pi/agent/models.json: { "providers": { "llama-cpp": { "baseUrl": "http://localhost:8080/v1", "api": "openai-completions", "apiKey": "none", "models": [ { "id": "AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0" } ] } } }Run Pi
# Start Pi in your project directory: pi
- OpenClaw new
How to use AbteeXAILab/lumynax-infused-qwen3-17b-gguf with OpenClaw:
Start the llama.cpp server
# Install llama.cpp: brew install llama.cpp # Start a local OpenAI-compatible server: llama serve -hf AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0
Configure OpenClaw
# Install OpenClaw: npm install -g openclaw@latest # Register the local server and set it as the default model: openclaw onboard --non-interactive --mode local \ --auth-choice custom-api-key \ --custom-base-url http://127.0.0.1:8080/v1 \ --custom-model-id "AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0" \ --custom-provider-id llama-cpp \ --custom-compatibility openai \ --custom-text-input \ --accept-risk \ --skip-health
Run OpenClaw
openclaw agent --local --agent main --message "Hello from Hugging Face"
- Docker Model Runner
How to use AbteeXAILab/lumynax-infused-qwen3-17b-gguf with Docker Model Runner:
docker model run hf.co/AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0
- Lemonade
How to use AbteeXAILab/lumynax-infused-qwen3-17b-gguf with Lemonade:
Pull the model
# Download Lemonade from https://lemonade-server.ai/ lemonade pull AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0
Run and chat with the model
lemonade run user.lumynax-infused-qwen3-17b-gguf-Q8_0
List all available models
lemonade list
- Hermes Agent
How to use AbteeXAILab/lumynax-infused-qwen3-17b-gguf with Hermes Agent:
Start the llama.cpp server
# Install llama.cpp: brew install llama.cpp # Start a local OpenAI-compatible server: llama serve -hf AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0
Configure Hermes
# Install Hermes: curl -fsSL https://hermes-agent.nousresearch.com/install.sh | bash hermes setup # Point Hermes at the local server: hermes config set model.provider custom hermes config set model.base_url http://127.0.0.1:8080/v1 hermes config set model.default AbteeXAILab/lumynax-infused-qwen3-17b-gguf:Q8_0
Run Hermes
hermes
- Atomic Chat
| fdeb5cd638f262467f8cfa3f97d14d4071076dfbee5907ca4e3af7e50a3bd08b .gitattributes | |
| c267255e71c41a5b665dde795c2d10dfc879c7614c1da427061724c935708228 artifacts/release_training_summary.json | |
| 4736809a6a40e57838c9a7be72565d05cb67de9fac9d72170398c8279e808808 artifacts/smoke_llama_cpp.json | |
| d74a16cac51b82a96c0a08217f38eb4152e0a4de400bd0b06498c5678681d808 hf_space/app.py | |
| 7860d80161c381596cc0fd093e7a53d6601763622421419afdc0fe4a6c274a0f hf_space/README.md | |
| 3e78abed8cdc940c6bf1c763d81d829ec4dfe2f8b3897e5051989ef19169eaeb hf_space/requirements.txt | |
| 4ab50822ff8fedbfe6990a78dbe3fdfd4c2784a709fa9e38aadbcac1bbf78598 LICENSE.txt | |
| 845ff57f31fd8614b3d614c8933d1908cf832d274a89d08e87041e1ccaa8afe0 merged_model/PACKAGE_STATE.txt | |
| 947d1618f3190c28bbb3cd563c968c634e90c90c9358eeb890f08c343d93a2b4 ollama/create_ollama_model.ps1 | |
| 42829274777221114d78e02b6cc4e1282c2ea010f7517d35e8dff98aa7b71434 ollama/Modelfile | |
| 11634717e21c1f8147c766d54002fc37ac1320da08fe5a468d8b6d49e68ea7fc quickstart.py | |
| 061b54daade076b5d3362dac252678d17da8c68f07560be70818cace6590cb1a Qwen3-1.7B-Q8_0.gguf | |
| 562ecc3d26ed1ffbe333a9cce9096e2c610062cbc0e383b32b0675be37f40050 README.md | |
| 309ccc3deb2475cfed8c407be46e65c77fbf0f5f408e67997adc77caebb69b4d release_export_manifest.json | |
| 2a7ca962dd79646b8470b45ec926ade0a4eb01ebdd93452e6990d8997666e378 requirements.txt | |
| cdf5c5bf1dad925791ffc3131f9554c122e060343a6e241e130190b018d4bc06 UPLOAD_TO_HF.md | |
| 2dfede0e6610c473959c963b292fcec325452acba33fd1bba21110e04933df53 VERSION.txt | |
| c67b5c495980b8b47373b282199db821db725db6048d3a214a45d8a004768dbe docs/lumynax-release-map.svg | |
| f8b25eeecdc70a58a25723a3fab574ad7055885224fd89473e4ce6a84468ece1 docs/lumynax-release-overview.svg | |
| aa3cce835dd7fc98d597fc46b68cbb4a2150ba2bd75cfbdf8bf49b7bea8413ad docs/lumynax-runtime-flow.svg | |