Skip to content

EthenEthenEthen

Open Source Model Profile · harkov000

R1-Stheno-8B

R1-Stheno-8B is an 8.03B-parameter Llama text-generation merge from harkov000. According to the model card, it is a SLERP merge built with mergekit.

Publisher
harkov000
Task
text-generation
Model type
llama
License
Unknown
Library
transformers
Publication status
Accepted · not indexed

Model overview

R1-Stheno-8B is published by harkov000 as a Transformers text-generation model. The captured configuration identifies LlamaForCausalLM with model type llama and about 8.03B Safetensors parameters. According to the model card, it merges DeepSeek-R1-Distill-Llama-8B with Sao10K/Llama-3.1-8B-Stheno-v3.4 using SLERP.

Recorded capabilities

Mergekit SLERP merge

According to the model card, this is a mergekit merge produced with the SLERP merge method.

Two documented parents

According to the model card, the merged models are DeepSeek-R1-Distill-Llama-8B and Sao10K/Llama-3.1-8B-Stheno-v3.4, matching the hub base-model tags.

8B-class Llama build

Captured configuration identifies LlamaForCausalLM with 8030261248 parameters, and the merge configuration specifies bfloat16 dtype.

Use cases in the source record

  • Conversational text-generation experiments comparing the merged model against its two documented parents.

Limitations and unknowns

  • No evaluation results were extracted from this record.
  • No license value was extracted for this record.
  • No context-window value was extracted from this record.
  • Provider state is historical snapshot data, not independently refreshed current availability.

Source and provenance

Source: harkov000/R1-Stheno-8B

Captured: Unknown. Processed: 2026-09-07T19:35:22.780724+00:00.

merge This is a merge of pre-trained language models created using mergekit . Merge Details Merge Method This model was merged using the SLERP merge method. Models Merged The following models were included in the merge: deepseek-ai/DeepSeek-R1-Distill-Llama-8B Sao10K/Llama-3.1-8B-Stheno-v3.4 Configuration The following YAML configuration was used to produce this model: models: - model: deepseek-ai/DeepSeek-R1-Distill-Llama-8B - model: Sao10K/Llama-3.1-8B-Stheno-v3.4 merge_method: slerp base_model: deepseek-ai/DeepSeek-R1-Distill-Llama-8B dtype: bfloat16 parameters: t: [ 0 , 0.5 , 0.25 ]

F001F002F003F004F005F006F008F009F010F011F012