1.24B Llama record
Captured configuration records LlamaForCausalLM and Safetensors metadata reports 1,235,814,400 parameters.
Open Source Model Profile · Grogros
This Grogros release is a 1.24B-parameter Llama text-generation fine-tune with documented hyperparameters and Llama3.2 licensing.
This Grogros release is a Llama-family text-generation fine-tune. The captured configuration identifies LlamaForCausalLM and Safetensors metadata reports 1,235,814,400 parameters. According to the model card, it is a fine-tune of Grogros/Llama-3.2-1B-Instruct-distillation-SecretSauceLongJail-5.0-HarmfulLLMLat on the None dataset.
Captured configuration records LlamaForCausalLM and Safetensors metadata reports 1,235,814,400 parameters.
According to the model card and hub tags, the release is presented as a fine-tune of Grogros/Llama-3.2-1B-Instruct-distillation-SecretSauceLongJail-5.0-HarmfulLLMLat.
According to the model card, training used learning rate 1e-05, batch sizes 16 and 8, seed 42, multi-GPU distribution, Adafactor, cosine scheduling, and 500 training steps.
Card data records llama3.2.
Source: Grogros/Llama-3.2-1B-Instruct-distillation-SecretSauceLongJail-5.0-HarmfulLLMLat-PT2
Captured: Unknown. Processed: 2026-09-07T19:35:05.753496+00:00.
Llama-3.2-1B-Instruct-distillation-SecretSauceLongJail-5.0-HarmfulLLMLat-PT2 This model is a fine-tuned version of Grogros/Llama-3.2-1B-Instruct-distillation-SecretSauceLongJail-5.0-HarmfulLLMLat on the None dataset. Model description More information needed Intended uses & limitations More information needed Training and evaluation data More information needed Training procedure Training hyperparameters The following hyperparameters were used during training: learning_rate: 1e-05 train_batch_size: 16 eval_batch_size: 8 seed: 42 distributed_type: multi-GPU optimizer: Use adafactor and the args are: No additional optimizer arguments…
F001F002F003F004F005F006F007F008F009F010F011F012