Ornith 1.5
DeepReinforce · Aug 2026DeepReinforce's agentic coding model, self-improved with RL - 1.5 extends the loop to generating its own training tasks. State-of-the-art among open coders at its size, purpose-built for terminal coding agents and tool use. The 9B runs on modest machines; the 35B is a mixture-of-experts (3B active per token) that runs fast for its size and adds vision.
from 5.6 GBruns from 16 GB RAM256K contextMIT