Skip to content

EthenEthenEthen

Open Source Model Profile · RLHFlow

Llama3-v2-iterative-DPO-iter3

Llama3-v2-iterative-DPO-iter3 is an 8.03B-parameter Llama text-generation model from RLHFlow. The captured model card is an unfilled auto-generated template, and no license value was extracted.

Publisher
RLHFlow
Task
text-generation
Model type
llama
License
Unknown
Library
transformers
Publication status
Approved for indexing

Model overview

RLHFlow publishes Llama3-v2-iterative-DPO-iter3 as a transformers text-generation model. Captured config identifies LlamaForCausalLM with a llama model type, and Safetensors metadata reports 8,030,261,248 parameters. Hub tags include safetensors, conversational, arxiv:1910.09700, and text-generation-inference. The captured card states only that it was automatically generated.

Recorded capabilities

About 8.03B parameters

Safetensors metadata reports 8,030,261,248 parameters, or about 8.03B.

Llama causal LM config

The captured configuration identifies LlamaForCausalLM with a llama model type.

Use cases in the source record

  • Conversational text-generation on the captured pipeline and conversational hub tags. No publisher task description was extracted.

Limitations and unknowns

  • No evaluation results were extracted from this record.
  • No license value was extracted from this record.
  • The captured model card is an unfilled auto-generated template and does not document training procedure, license, or architecture details.
  • The repository name is not treated as evidence of Llama 3 lineage, DPO training, or iteration count.
  • Provider state is historical snapshot data, not independently refreshed current availability.

Source and provenance

Source: RLHFlow/Llama3-v2-iterative-DPO-iter3

Captured: Unknown. Processed: 2026-09-07T19:34:36.066018+00:00.

Model Card for Model ID Model Details Model Description This is the model card of a 🤗 transformers model that has been pushed on the Hub. This model card has been automatically generated. Developed by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Model type: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Finetuned from model [optional]: [More Information Needed] Model Sources [optional] Repository: [More Information Needed] Paper [optional]: [More Information Needed] Demo [optional]: [More Informat…

F001F002F003F004F005F006F007F008F009F010F011F014F015F016