DELLA mergekit construction
According to the model card, the model was merged with mergekit using the DELLA method on Qwen2.5-7B-Instruct as the base.
Open Source Model Profile · Yuuta208
Qwen2.5-7B-Instruct-Qwen2.5-Coder-7B-Merged-della-29 is a 7.62B-parameter Qwen2 merge from Yuuta208 combining instruction and coder variants with DELLA.
Qwen2.5-7B-Instruct-Qwen2.5-Coder-7B-Merged-della-29 is published by Yuuta208 as a text-generation merge. Captured configuration identifies Qwen2ForCausalLM with a qwen2 model type, and Safetensors metadata reports 7,615,616,512 parameters. According to the model card, it combines Qwen2.5-7B-Instruct and Qwen2.5-Coder-7B with the DELLA method.
According to the model card, the model was merged with mergekit using the DELLA method on Qwen2.5-7B-Instruct as the base.
The model card says Qwen2.5-7B-Instruct and Qwen2.5-Coder-7B were included with weights 0.5 and 0.6, density 0.8, normalization, int8 masking, and float16 dtype.
Captured configuration records Qwen2ForCausalLM with a qwen2 model type at about 7.62B parameters.
Source: Yuuta208/Qwen2.5-7B-Instruct-Qwen2.5-Coder-7B-Merged-della-29
Captured: Unknown. Processed: 2026-09-07T19:35:17.665719+00:00.
output_model_della This is a merge of pre-trained language models created using mergekit . Merge Details Merge Method This model was merged using the DELLA merge method using Qwen/Qwen2.5-7B-Instruct as a base. Models Merged The following models were included in the merge: Qwen/Qwen2.5-Coder-7B Configuration The following YAML configuration was used to produce this model: models: - model: Qwen/Qwen2.5-7B-Instruct parameters: weight: 0.5 - model: Qwen/Qwen2.5-Coder-7B parameters: weight: 0.6 merge_method: della base_model: Qwen/Qwen2.5-7B-Instruct parameters: density: 0.8 normalize: true int8_mask: true dtype: float16
F001F002F003F004F005F006F009F010F011F012