Mergekit SLERP merge
According to the model card, this is a mergekit merge produced with the SLERP merge method.
Open Source Model Profile · harkov000
R1-Stheno-8B is an 8.03B-parameter Llama text-generation merge from harkov000. According to the model card, it is a SLERP merge built with mergekit.
R1-Stheno-8B is published by harkov000 as a Transformers text-generation model. The captured configuration identifies LlamaForCausalLM with model type llama and about 8.03B Safetensors parameters. According to the model card, it merges DeepSeek-R1-Distill-Llama-8B with Sao10K/Llama-3.1-8B-Stheno-v3.4 using SLERP.
According to the model card, this is a mergekit merge produced with the SLERP merge method.
According to the model card, the merged models are DeepSeek-R1-Distill-Llama-8B and Sao10K/Llama-3.1-8B-Stheno-v3.4, matching the hub base-model tags.
Captured configuration identifies LlamaForCausalLM with 8030261248 parameters, and the merge configuration specifies bfloat16 dtype.
Source: harkov000/R1-Stheno-8B
Captured: Unknown. Processed: 2026-09-07T19:35:22.780724+00:00.
merge This is a merge of pre-trained language models created using mergekit . Merge Details Merge Method This model was merged using the SLERP merge method. Models Merged The following models were included in the merge: deepseek-ai/DeepSeek-R1-Distill-Llama-8B Sao10K/Llama-3.1-8B-Stheno-v3.4 Configuration The following YAML configuration was used to produce this model: models: - model: deepseek-ai/DeepSeek-R1-Distill-Llama-8B - model: Sao10K/Llama-3.1-8B-Stheno-v3.4 merge_method: slerp base_model: deepseek-ai/DeepSeek-R1-Distill-Llama-8B dtype: bfloat16 parameters: t: [ 0 , 0.5 , 0.25 ]
F001F002F003F004F005F006F008F009F010F011F012