Skip to content

EthenEthenEthen

Open Source Model Profile · ATH-MaaS

Marco-LLM-ES

Marco-LLM-ES is a 7.62B-parameter Qwen2 text-generation model from ATH-MaaS. Its model card describes continued pretraining for Catalan, Basque, Galician, and Spanish.

Publisher
ATH-MaaS
Task
text-generation
Model type
qwen2
License
apache-2.0
Library
Unknown
Publication status
Accepted · not indexed

Model overview

Marco-LLM-ES is published by ATH-MaaS as a Qwen2 text-generation model. The captured configuration identifies Qwen2ForCausalLM, and Safetensors metadata reports 7,615,616,512 parameters. According to the model card, it belongs to a series fine-tuned for common languages used in Spain, including Catalan, Basque, Galician, and Spanish.

Recorded capabilities

Documented Spain-languages focus

According to the model card, the series is fine-tuned for Catalan, Basque, Galician, and Spanish, with this repository holding the 7B base model.

Reported 50B-token continued pretraining

According to the model card, the model underwent extensive continued pretraining on about 50 billion tokens for the target languages.

Multilingual tokenizer and architecture note

According to the model card, the series uses a Transformer design with SwiGLU, QKV bias, group-query attention, and an improved multilingual tokenizer.

Apache-2.0 license

Card data records Apache-2.0 licensing, and hub tags include a license:apache-2.0 entry.

Use cases in the source record

  • Spanish-language text workflows in Catalan, Basque, Galician, and Spanish using the publisher-described continued-pretraining base.
  • Post-training research with SFT, RLHF, or further pretraining, which the model card recommends over direct base-model generation.

Limitations and unknowns

  • No evaluation results were extracted from this record; the card's competitiveness statement has not been independently verified.
  • No context-window value, hardware requirement, or inference pricing was extracted.
  • Provider state is historical snapshot data and should be refreshed before being presented as current.
  • Language and training claims come from the publisher model card and linked paper, and were not independently verified by Ethen.

Source and provenance

Source: ATH-MaaS/Marco-LLM-ES

Captured: Unknown. Processed: 2026-09-07T19:35:01.987227+00:00.

Marco-LLM-ES-7B Introduction Marco-LLM-ES is a series of enhanced language models specifically fine-tuned for common languages used in Spain, including Catalan, Basque, Galician, and Spanish. This repository contains the 7B Marco-LLM-ES base language model. Compared with the state-of-the-art open-source language models, Marco-LLM-ES has undergone extensive continued pretraining on a dataset containing approximately 50 billion tokens, enhancing its capabilities in the targeted languages while maintaining competitiveness in general benchmarks. For more details, please refer to our Hugging Face page . Model Details Marco-LLM-ES series…

F001F002F003F004F005F006F007F008F009F010F011F013F014F015