8.03B Llama scale
Captured config identifies LlamaForCausalLM and Safetensors metadata reports 8030261248 parameters.
Open Source Model Profile · birgermoell
Llama-3-dare_ties is an 8.03B-parameter Llama text-generation merge from birgermoell. Its model card documents a LazyMergekit dare_ties configuration and Transformers usage.
Llama-3-dare_ties is published by birgermoell as a text-generation model. The captured configuration identifies LlamaForCausalLM with a llama model type, and Safetensors metadata reports 8030261248 parameters. According to the model card, it merges Meta-Llama-3-8B and Meta-Llama-3-8B-Instruct with the dare_ties method.
Captured config identifies LlamaForCausalLM and Safetensors metadata reports 8030261248 parameters.
According to the model card, the merge uses density 0.53, weight 0.6, dare_ties method, int8_mask true, and bfloat16 dtype.
The card names Meta-Llama-3-8B as base and Meta-Llama-3-8B-Instruct in the merge, matching hub base-model tags.
The model card shows Transformers pipeline usage with float16, automatic device mapping, and sampling settings.
Source: birgermoell/Llama-3-dare_ties
Captured: Unknown. Processed: 2026-09-07T19:34:41.514765+00:00.
Llama-3-dare_ties Llama-3-dare_ties is a merge of the following models using LazyMergekit : meta-llama/Meta-Llama-3-8B-Instruct 🧩 Configuration models: - model: meta-llama/Meta-Llama-3-8B - model: meta-llama/Meta-Llama-3-8B-Instruct parameters: density: 0.53 weight: 0.6 merge_method: dare_ties base_model: meta-llama/Meta-Llama-3-8B parameters: int8_mask: true dtype: bfloat16 💻 Usage !pip install -qU transformers accelerate from transformers import AutoTokenizer import transformers import torch model = "birgermoell/Llama-3-dare_ties" messages = [{ "role" : "user" , "content" : "What is a large language model?" }] tokenizer = AutoTo…
F001F002F003F004F005F006F007F008F009F010F011F012F013F014