agentic post-training
agentic post-training Articles
Browse 3 articles about agentic post-training.

NeoHorse-1-9B: Agentic Post-Training on Qwen3.5-9B Explained
NeoHorse-1-9B fine-tunes Qwen3.5-9B with a routing harness for agentic post-training, gaining +3.44 macro average across ten benchmarks.
NeoHorse-1-9BQwen3.5-9B fine-tuneagentic post-training

How to Run NeoHorse-1-4B Locally: Specs and Setup Basics
NeoHorse-1-4B ships as BF16 safetensors with a 262K native context, extensible to 1M. Here's what that means for local hardware.
NeoHorse-1-4B localrun NeoHorse locally4B model context length

NeoHorse-1-4B: A Small Model Testing the Road to Self-Improving AI
NeoHorse-1-4B fine-tunes Qwen3.5-4B with a routing harness for agentic tasks, scoring 64.87 average, up 5.93 points over its base model.
NeoHorse-1-4Brecursive self-improvement AIQwen3.5-4B fine-tune