How to use from
vLLM
Install from pip and serve model
# Install vLLM from pip:
pip install vllm
# Start the vLLM server:
vllm serve "allura-org/Mistral-Small-24b-Sertraline-0304"
# Call the server using curl (OpenAI-compatible API):
curl -X POST "http://localhost:8000/v1/chat/completions" \
	-H "Content-Type: application/json" \
	--data '{
		"model": "allura-org/Mistral-Small-24b-Sertraline-0304",
		"messages": [
			{
				"role": "user",
				"content": "What is the capital of France?"
			}
		]
	}'
Use Docker
docker model run hf.co/allura-org/Mistral-Small-24b-Sertraline-0304
Quick Links

Sertraline 24b

sertraline_summary1.jpg

About

An actually decent instruct SFT tune of Mistral Small 3.

System Prompts

I tested with the following Claude-like system prompts, however they were not trained in and any similar prompts can likely be used:

Non-Reasoning

You are Claude, a helpful and harmless AI assistant created by Anthropic.

Reasoning

You are Claude, a helpful and harmless AI assistant created by Anthropic. Please contain all your thoughts in <think> </think> tags, and your final response right after the closing </think> tag.

For reasoning, it's recommended to force the thinking (by prefilling <think>\n on the newest assistant response), as well as not including previous thought blocks in new requests.

Instruct Template

v7-Tekken, same as the original instruct model.

Dataset

This model was trained on allura-org/inkstructmix-v0.2.1.

Downloads last month
14
Safetensors
Model size
24B params
Tensor type
BF16
·
Inference Providers NEW
Input a message to start chatting with allura-org/Mistral-Small-24b-Sertraline-0304.

Model tree for allura-org/Mistral-Small-24b-Sertraline-0304

Finetuned
(44)
this model
Merges
1 model
Quantizations
5 models

Datasets used to train allura-org/Mistral-Small-24b-Sertraline-0304