Skip to content

EthenEthenEthen

Open Source Model Profile · princeton-nlp

Llama-3-Base-8B-SFT-IPO

Llama-3-Base-8B-SFT-IPO is an 8.03B-parameter Llama text-generation release from princeton-nlp. Its model card links it to the SimPO preprint on reference-free preference optimization.

Publisher
princeton-nlp
Task
text-generation
Model type
llama
License
Unknown
Library
transformers
Publication status
Accepted · not indexed

Model overview

Llama-3-Base-8B-SFT-IPO is published by princeton-nlp as a llama-based text-generation model. The captured configuration identifies LlamaForCausalLM and Safetensors metadata reports 8030261248 parameters. According to the model card, it was released from the SimPO preprint on simple preference optimization with a reference-free reward.

Recorded capabilities

Llama 8.03B configuration

Captured config identifies LlamaForCausalLM and llama, with Safetensors metadata reporting 8030261248 parameters.

SimPO preprint link

According to the model card, this release comes from the preprint SimPO: Simple Preference Optimization with a Reference-Free Reward.

Transformers conversational record

Hub data records transformers library support, with tags including conversational, text-generation-inference, and endpoints-compatible.

Use cases in the source record

  • Conversational text-generation workflows using the captured transformers Llama configuration with inference-endpoint compatibility.
  • Preference-optimization research referencing the documented SimPO preprint, using the card's external repository pointer.

Limitations and unknowns

  • No license value was extracted from this record.
  • No evaluation results were extracted from this record.
  • No context-window value was extracted from this record.
  • Provider state is historical snapshot data, not independently refreshed current availability.
  • Training, dataset, and method details beyond the SimPO preprint reference were not extracted; the card points to an external repository.

Source and provenance

Source: princeton-nlp/Llama-3-Base-8B-SFT-IPO

Captured: Unknown. Processed: 2026-09-07T19:34:55.787866+00:00.

This is a model released from the preprint: SimPO: Simple Preference Optimization with a Reference-Free Reward Please refer to our repository for more details.

F001F002F003F004F005F006F007F008F009F010