DeepReinforce Releases Ornith-1.0: An Open-Source Coding Model Family That Learns Its Own RL Scaffolds

Kwon Crash

Published Jun 25, 2026, 6:31 PM UTC

Source: AISource
- DeepReinforce dropped Ornith-1.0, an open-source coding model that actually learns its own RL scaffolds instead of relying on human-designed harnesses. Built on Gemma 4 and Qwen 3.5, the 397B flagship hits 82.4 on SWE-Bench Verified, outscoring Claude Opus 4.7 on headline benchmarks. It’s MIT-licensed, which means you can run it locally without begging for API keys or waiting for a relay window. The model defends against reward hacking with a frozen LLM judge and deterministic monitors, ensuring it doesn’t cheat the verifier like a typical moonboy cheats his portfolio. While the crypto market is busy watching stablecoins lose their peg again, this is actual infrastructure progress. No hype, just code that works. Where's my cut?