3 models
Ranked by Terminal-Bench 2.1
Open weights 78.2

Large coding-reasoning model for agentic software tasks and RL search

deepreinforce/ornith-1.0-397b 2026-06-25 262.144K context Input not listed Output not listed
Open weights 62.8

Large coding-reasoning model for agentic software tasks and RL search

deepreinforce/ornith-1.0-35b 2026-06-25 262.144K context Input not listed Output not listed
Open weights 40.6

Open coding-reasoning model for repository tasks and self-improving agents

deepreinforce/ornith-1.0-9b 2026-06-25 262.144K context Input not listed Output not listed