Ornith-1.0
Agentic Codingに特化した公開LLM
軽量な9B,中規模な35B,大規模な397Bの3段階が用意されている
最大の397Bは,Agentic Coding系のベンチマークにおいて,Claude Opus 3.7に匹敵する性能を示している
https://gyazo.com/a2f10acafedd1e3edd4464c15fc7ac3b
Ornith-1.0: Self-Scaffolding LLMs for Agentic Coding | DeepReinforce Blog | Jun. 2026 より引用
35B,9Bは同規模のモデルと比較して,高いスコアを示している.
https://gyazo.com/8e206bfc49137b9eb86dffab41ea85e7
公開モデルとの比較(左:35Bクラス,右:12Bクラス | 赤:高得点,白:平均点,青:低得点)
ファイル:Orinith_1_0.xlsx
スコアデータは以下を使用
deepreinforce-ai/Ornith-1.0-9B · Hugging Face
deepreinforce-ai/Ornith-1.0-35B-FP8 · Hugging Face
参考
Ornith-1.0: Self-Scaffolding LLMs for Agentic Coding | DeepReinforce Blog | Jun. 2026
Ornith 1.0 を DGX Spark で動かして日本語性能を Gemma 4 / Nemotron と比べてみた | DevelopersIO
Claude Opus 4.7と同等性能のコーディングAIモデル「Ornith-1.0」が登場、ローカルで動作する小型モデルもラインナップ - GIGAZINE