Skip to content

EthenEthenEthen

Open Source Model Profile · open-thoughts

OpenThinker-Agent-v1

OpenThinker-Agent-v1 is an open-thoughts 8.19B Qwen3 agent model. Its card documents Qwen3-8B post-training for terminal and coding tasks.

Publisher
open-thoughts
Task
text-generation
Model type
qwen3
License
apache-2.0
Library
transformers
Publication status
Accepted · not indexed

Model overview

OpenThinker-Agent-v1 is published by open-thoughts as a text-generation model for agentic work. The captured configuration identifies Qwen3ForCausalLM with Safetensors metadata reporting 8,190,735,360 parameters. According to the model card, it is post-trained from Qwen/Qwen3-8B for tasks including Terminal-Bench 2.0 and SWE-Bench.

Recorded capabilities

Qwen3-8B agent post-train

According to the model card, the model is post-trained from Qwen/Qwen3-8B with sequential SFT and RL stages for agentic tasks.

Terminal and software-agent focus

The model card says it is trained for agentic tasks such as Terminal-Bench 2.0 and SWE-Bench.

Curated SFT, RL, and filtration pipeline

The card documents about 15,200 SFT traces, about 720 RL tasks, and a three-stage filtration pipeline that removes tasks with unreliable or slow verifiers.

Use cases in the source record

  • Terminal-agent workflows benchmarked against Terminal-Bench-style shell tasks.
  • Software-engineering agent experiments in the style of SWE-Bench coding tasks.

Limitations and unknowns

  • No evaluation results were extracted from this record.
  • No context-window value was extracted from this record.
  • Provider state is historical snapshot data and should be refreshed before being presented as current availability.
  • Agent-benchmark characterizations come from the publisher model card and have not been independently verified by Ethen.

Source and provenance

Source: open-thoughts/OpenThinker-Agent-v1

Captured: Unknown. Processed: 2026-09-07T19:35:59.605475+00:00.

Project | SFT dataset | RL dataset | SFT model | RL model OpenThinker-Agent-v1 OpenThoughts-Agent is an open-source effort to curate the best datasets for training agents. Our first release includes datasets , models and our research codebase . OpenThinker-Agent-v1 is a model trained for agentic tasks such as Terminal-Bench 2.0 and SWE-Bench . The OpenThinker-Agent-v1 model is post-trained from Qwen/Qwen3-8B . It is SFT-ed on the OpenThoughts-Agent-v1-SFT dataset, then RL-ed on the OpenThoughts-Agent-v1-RL dataset. This model is the final model after both SFT and RL. For the model after the SFT stage only, see OpenThinker-Agent-v1-S…

F001F002F003F004F005F006F007F008F009F010F012F013F014F015F016F017