How to use from the
Use from the
Transformers library
# Use a pipeline as a high-level helper
from transformers import pipeline

pipe = pipeline("text-generation", model="allura-org/Mistral-Small-24b-Sertraline-0304")
messages = [
    {"role": "user", "content": "Who are you?"},
]
pipe(messages)
# Load model directly
from transformers import AutoTokenizer, AutoModelForCausalLM

tokenizer = AutoTokenizer.from_pretrained("allura-org/Mistral-Small-24b-Sertraline-0304")
model = AutoModelForCausalLM.from_pretrained("allura-org/Mistral-Small-24b-Sertraline-0304", device_map="auto")
messages = [
    {"role": "user", "content": "Who are you?"},
]
inputs = tokenizer.apply_chat_template(
	messages,
	add_generation_prompt=True,
	tokenize=True,
	return_dict=True,
	return_tensors="pt",
).to(model.device)

outputs = model.generate(**inputs, max_new_tokens=40)
print(tokenizer.decode(outputs[0][inputs["input_ids"].shape[-1]:]))
Quick Links

Sertraline 24b

sertraline_summary1.jpg

About

An actually decent instruct SFT tune of Mistral Small 3.

System Prompts

I tested with the following Claude-like system prompts, however they were not trained in and any similar prompts can likely be used:

Non-Reasoning

You are Claude, a helpful and harmless AI assistant created by Anthropic.

Reasoning

You are Claude, a helpful and harmless AI assistant created by Anthropic. Please contain all your thoughts in <think> </think> tags, and your final response right after the closing </think> tag.

For reasoning, it's recommended to force the thinking (by prefilling <think>\n on the newest assistant response), as well as not including previous thought blocks in new requests.

Instruct Template

v7-Tekken, same as the original instruct model.

Dataset

This model was trained on allura-org/inkstructmix-v0.2.1.

Downloads last month
14
Safetensors
Model size
24B params
Tensor type
BF16
·
Inference Providers NEW
Input a message to start chatting with allura-org/Mistral-Small-24b-Sertraline-0304.

Model tree for allura-org/Mistral-Small-24b-Sertraline-0304

Finetuned
(44)
this model
Merges
1 model
Quantizations
5 models

Datasets used to train allura-org/Mistral-Small-24b-Sertraline-0304