Testing-only reasoning variant
According to the model card, the publisher labels this Meta Llama 3 8B reasoning release for testing purposes only.
Open Source Model Profile · CreitinGameplays
Llama-3.1-8b-reasoning-test is an 8.03B-parameter Llama-family text-generation model from CreitinGameplays. According to the model card, it is a testing-only reasoning experiment.
Llama-3.1-8b-reasoning-test is published by CreitinGameplays as a text-generation model. The captured configuration identifies LlamaForCausalLM with model type llama, and Safetensors metadata reports about 8.03B parameters under mit. According to the model card, it is a testing-only Llama 3 8B reasoning experiment.
According to the model card, the publisher labels this Meta Llama 3 8B reasoning release for testing purposes only.
According to the model card, the example loads with bitsandbytes 8-bit quantization when CUDA is available and falls back to CPU otherwise.
According to the model card, the prompt instructs the assistant to use the end_reasoning token when ending a reasoning step with the stated sampling settings.
Source: CreitinGameplays/Llama-3.1-8b-reasoning-test
Captured: Unknown. Processed: 2026-09-07T19:35:04.077226+00:00.
Meta Llama 3 8B Reasoning (Testing purpose only) Code example: # test the model import torch from transformers import AutoTokenizer, AutoModelForCausalLM, TextStreamer def main (): model_id = "CreitinGameplays/Llama-3.1-8b-reasoning-test" # Load the tokenizer. tokenizer = AutoTokenizer.from_pretrained(model_id, add_eos_token= True ) # Load the model using bitsandbytes 8-bit quantization if CUDA is available. if torch.cuda.is_available(): model = AutoModelForCausalLM.from_pretrained( model_id, load_in_8bit= True , device_map= "auto" ) device = torch.device( "cuda" ) else : model = AutoModelForCausalLM.from_pretrained(model_id) device…
F001F002F003F004F005F006F007F010F011F012F013F014F017F019