Vulcan-Agent 0.8B โ€” Convergence SFT v2 (no-run shakeout)

First checkpoint lineage in the convergence turn format: structured thinking block (REFLECTION/STATE/GOAL/PLAN/PREDICT/DECISION), environment/agent conversation roles, Qwen-XML tool calls after </think>. Trained on the v2 no-run mix (140,730 rows across agentic loops, plan-act handoff, RDDL planning, chat, instruct, creative writing; sources incl. TextArena, Factorio, Diplomacy, Jericho, ScienceWorld, xlam, gsm8k-grounded).

This is the EfficientZero-shakeout variant: rows containing executable <run> blocks were excluded (uses_run=False filter) to isolate pipeline validation from the REPL architecture change. The REPL-enabled sibling trains on the full mix.

Usage notes

  • Chat template ships in the repo (chat_template.jinja) โ€” it accepts roles system/environment/agent/user and will raise on anything else.
  • Generation prompt: <|im_start|>agent\n. Stop on <|im_end|> (id 248046).
  • Tool calls: <tool_call><function=NAME><parameter=KEY>...</parameter></function></tool_call>
  • Part of the Vulcan cognitive-core program (Artivus).
Downloads last month
4
Safetensors
Model size
0.9B params
Tensor type
BF16
ยท
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for artivus-ai/vulcan-agent-0.8b-convergence-sft-v2

Finetuned
(451)
this model