gemma-2-2b-legal-sft

Gemma 2 2B fine-tuned with QLoRA (4-bit NF4 base + LoRA adapters, r=16, alpha=32) on ~15k grounded legal Q&A pairs (context + question -> answer-from-context, generated from a cleaned US case-law / SEC corpus). Adapters merged back into the base and saved here. Answers a question from a passage you provide.

Recipe: trl SFTTrainer, 3 epochs, lr 2e-4, batch 1 x grad-accum 16, bf16, gradient checkpointing. Built on Modal. Full pipeline + code: https://github.com/Vizuara-AI-Lab/slm-125m-from-scratch

Downloads last month
27
Safetensors
Model size
3B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for thesreedath/gemma-2-2b-legal-sft

Finetuned
(1041)
this model