logging to output/sampled_response/cluster_1/training.log [sft] loading tokenizer... [sft] loading model... [sft] tokenizing 1206 examples... [sft] tokenizing validation split for eval... [data] loaded 150 validation examples (config=sampled_response, subpop=cluster_1) [sft] eval split: 150 examples [sft] starting training... [sft] saving... saved to output/sampled_response/cluster_1 gpu: NVIDIA RTX 6000 Ada Generation time: 1088s training log -> output/sampled_response/cluster_1/training_log.json run metadata -> output/sampled_response/cluster_1/finetune_config.json environment -> output/sampled_response/cluster_1/environment.json README -> output/sampled_response/cluster_1/README.md [upload] uploading output/sampled_response/cluster_1 to https://huggingface.co/1jamesthompson1/Qwen3.5-9B-nz-wvs-sampled_response-cluster_1...