Skip to content

EthenEthenEthen

Open Source Model Profile · 123-cao

Qwen2-0.5B-Instruct

Qwen2-0.5B-Instruct is a 0.49B-parameter Qwen2-family instruction-tuned text-generation model from 123-cao. According to the model card, it contains the instruction-tuned 0.5B Qwen2 model.

Publisher
123-cao
Task
text-generation
Model type
qwen2
License
apache-2.0
Library
Unknown
Publication status
Accepted · not indexed

Model overview

Qwen2-0.5B-Instruct is published by 123-cao as a text-generation model. The captured configuration identifies Qwen2ForCausalLM with model type qwen2, and Safetensors metadata reports about 0.49B parameters. According to the model card, it contains the instruction-tuned 0.5B Qwen2 model, and hub tags link it to Qwen/Qwen2-0.5B.

Recorded capabilities

Qwen2 text-generation architecture

The captured configuration identifies Qwen2ForCausalLM with model type qwen2, tagged for safetensors, chat, and conversational use.

0.49B instruction-tuned size

Safetensors metadata reports 494,032,768 parameters; according to the model card, this repo contains the instruction-tuned 0.5B Qwen2 model.

Qwen2-0.5B base lineage

Hub tags record base_model:Qwen/Qwen2-0.5B and a matching finetune tag.

Documented chat template

According to the model card, the publisher documents tokenizer.apply_chat_template usage with a user-role prompt example.

Publisher-reported comparison table

According to the model card, the publisher reports comparison figures against Qwen1.5-0.5B-Chat across MMLU, HumanEval, GSM8K, C-Eval, and IFEval.

Use cases in the source record

  • Conversational text-generation workflows using the publisher-documented chat-template format.
  • Instruction-tuning experiments consistent with the publisher-described supervised fine-tuning and preference-optimization post-training.

Limitations and unknowns

  • No context-window value was extracted from this record.
  • No evaluation results were extracted in structured form; the comparison table in the model card is a publisher claim and was not independently verified.
  • Provider state is historical snapshot data and should be refreshed before being presented as current.
  • Series-level capability and training descriptions come from the publisher model card and describe the Qwen2 family generally.

Source and provenance

Source: 123-cao/Qwen2-0.5B-Instruct

Captured: Unknown. Processed: 2026-09-07T19:35:33.409670+00:00.

Qwen2-0.5B-Instruct Introduction Qwen2 is the new series of Qwen large language models. For Qwen2, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters, including a Mixture-of-Experts model. This repo contains the instruction-tuned 0.5B Qwen2 model. Compared with the state-of-the-art opensource language models, including the previous released Qwen1.5, Qwen2 has generally surpassed most opensource models and demonstrated competitiveness against proprietary models across a series of benchmarks targeting for language understanding, language generation, multilingual…

F001F002F003F004F005F006F007F008F009F010F011F013F014F015F016F017