Llama-3.1 reasoning lineage
According to the model card, it was finetuned from unsloth/meta-llama-3.1-8b-instruct-bnb-4bit as an open reasoning Llama referencing Open-R1.
Open Source Model Profile · EpistemeAI
Reasoning-Llama-3.1-CoT-RE1 is an 8.03B-parameter Llama-3.1 reasoning fine-tune from EpistemeAI. According to the model card, it uses Unsloth and TRL from a 4-bit Llama instruct base.
Reasoning-Llama-3.1-CoT-RE1 is published by EpistemeAI as a Llama-based text-generation fine-tune. The captured configuration identifies LlamaForCausalLM and Safetensors metadata reports 8030261248 parameters. According to the model card, it was finetuned from unsloth/meta-llama-3.1-8b-instruct-bnb-4bit with Unsloth and TRL, referencing Open-R1.
According to the model card, it was finetuned from unsloth/meta-llama-3.1-8b-instruct-bnb-4bit as an open reasoning Llama referencing Open-R1.
According to the model card, training was 2x faster with Unsloth and Hugging Face TRL, with hub tags listing the Bespoke-Stratos-17k dataset.
Captured config reports LlamaForCausalLM and llama, with Safetensors metadata reporting about 8.03B parameters.
Source: EpistemeAI/Reasoning-Llama-3.1-CoT-RE1
Captured: Unknown. Processed: 2026-09-07T19:35:05.149934+00:00.
Reasoning Llama model Open Reasoning Llama model Reference Open-R1: a fully open reproduction of DeepSeek-R1 Uploaded model Developed by: EpistemeAI License: apache-2.0 Finetuned from model : unsloth/meta-llama-3.1-8b-instruct-bnb-4bit This llama model was trained 2x faster with Unsloth and Huggingface's TRL library.
F001F002F003F004F005F006F007F009F010F011