logo
  • menu
  • Markets
  • ETFs
  • Live
  • Spot
  • Futures
  • Bots
  • Learn
  • Sign In
  • Sign Up
  • Downloads
  • English
  • |
  • USD
  • |
Sign Up
Crypto PricesLearnLatest NewsDownloadsMarketsSpotAnnouncements
Home/
Latest News/
Live

Ornith Is the Open-Source Coding Model Built for Agents, Not Humans

By Decrypt
Jun 30, 2026
4.1 
★
★
★
★
★
★
★
★
★
★
 79 User Rating
Share

Parameters are basically the number of dials and configurations a model can handle on its training. The more parameters, the more capable a model is. A 9-billion-parameter model is considered small, good enough to run on a good smartphone, but not capable of doing any heavy reasoning task reliably. A 397 billion model is much more capable, but requires some heavy computing, the kind that is not available on consumer hardware.

Aloha! Meet Ornith-1.0, a family of open-source LLMs specialized for agentic coding.

So Agentic AI means no one needs to be at the keyboard for most of the time. That's the whole point. This is also the direction where the most commercially relevant progress is happening in 2026—the models that can run unsupervised through 20-step dev workflows are worth more than the ones that write a clean function on request.

However, most large language models are still designed with human feedback in mind.

How Ornith’s brain works

Most AI coding agents are paired with a human-designed harness—a fixed set of rules for how the agent structures its work: when to call a tool, how to handle an error, how to decompose a multi-step problem. Ornith instead "treats the scaffold as a learnable object that co-evolves with the policy."

Translation: instead of inheriting someone else's playbook, it develops its own.

During reinforcement learning, each training step happens in two stages. The model first reads the task and proposes a refined strategy for approaching it. Then it uses that strategy to generate a solution.

The reward from the outcome flows back to both stages—so the model is optimized for writing better strategies, not just better code. Do that thousands and millions of times, and task-specific approaches emerge without a human engineering them.

DeepReinforce also takes reward hacking seriously. If the model can write its own training scaffold, it can theoretically write a scaffold that games the verifier—touching a file to make it look like it completed a task without actually doing the work. Three layers of defense block this: the environment and test suite are immutable and outside the model's reach, a deterministic monitor flags any attempt to access restricted paths or alter verification scripts, and a frozen judge model sits on top of the automated verifier as a veto.

The numbers

The flagship 397 billion parameter model posts 82.4 on SWE-bench Verified—a test where an AI is given a real bug from an open-source GitHub repository and must fix it without seeing the test suite, scored as the percentage of issues it successfully resolves.

That beats Claude Opus 4.7's 80.8 and DeepSeek-V4-Pro's 80.6 on the same test. On Terminal Bench 2.1—89 tasks run inside containerized terminal environments ranging from debugging async code to resolving security vulnerabilities, scored by completion rate—it posts 77.5 against Claude Opus 4.7's 70.3. 

The 9 billion parameter model might be the more interesting data point. It posts 69.4 on SWE-bench Verified—higher than Gemma 4-31B's 52 and competitive with Qwen 3.5-35B's 70, despite being 3-4 times smaller.

Who it's for, and who it isn't

Ornith-1.0 is explicitly not a general-purpose AI. The model's own documentation says it may underperform on tasks outside agentic coding. If you want AI to summarize a document, help you write your doctoral thesis, or draft an email, Ornith-1.0 is the wrong pick.

It's optimized for a narrow problem set: developer pipelines where an AI agent takes a task description, operates inside a code repository or terminal session, and completes multi-step work without intervention. This is a tool that was built for people who are already running agent infrastructure—not for people trying to decide if AI is worth using.

Ornith-1.0-397B does surpass Claude Opus 4.7 on both different coding benchmarks, but Anthropic's current flagship, Claude Opus 4.8, scores higher. The comparison that holds is within the open-source category, at comparable parameter counts, on coding-specific agent tasks.

For developers building self-hosted coding pipelines, agentic infrastructure, or similar coding-focused work, the small and medium models running on edge hardware may be genuinely useful, but the average Joe may be better looking somewhere else.

Disclaimer: The information on this page may have been obtained from third parties and does not necessarily reflect the views or opinions of BitKan. This content is provided for general informational purposes only, without any representation or warranty of any kind, nor shall it be construed as financial or investment advice. BitKan shall not be liable for any errors or omissions, or for any outcomes resulting from the use of this information. Investments in digital assets can be risky. Please carefully evaluate the risks of a product and your risk tolerance based on your own financial circumstances. Products mentioned in this article may not be available in your region.

Latest News

Industry

Cryptocurrency

Airdrop

Markets

  • Brazil’s CVM Launches 60-Day Sprint to Tokenize Securities

    Brazil’s CVM Launches 60-Day Sprint to Tokenize Securities

    The Brazilian Securities and Exchange Commission (CVM) has officially established a dedicated task force to develop an experimental regulatory framework for tokenized securities, providing a fast-tracked timeline for the digital capital market.
    Martha Grizzard
    Jul 21, 2026
  • Hyperliquid Enables Permissionless Markets With HIP-4 Plan

    Hyperliquid Enables Permissionless Markets With HIP-4 Plan

    Hyperliquid has announced a forthcoming enhancement to its HIP-4 upgrade that will allow for the permissionless deployment of decentralized prediction markets.
    Christopher Smith
    Jul 21, 2026
  • DTCC Launches Live Tokenized Asset Trading for Wall Street

    DTCC Launches Live Tokenized Asset Trading for Wall Street

    The DTCC successfully transitioned from pilot testing to live production trades on July 15, 2026, marking the largest-scale institutional tokenization initiative to date.
    Cornell Rachel
    Jul 16, 2026
  • South Korea Updates Asset Law to Include Cryptocurrency

    South Korea Updates Asset Law to Include Cryptocurrency

    The South Korean Ministry of Economy and Finance announced a transition from the 1950 State Property Act to a new National Asset Basic Act to better reflect modern digital resources.
    Martha Grizzard
    Jul 16, 2026
  • New SEC Crypto Rule to Cut Red Tape for Startup Fundraising

    New SEC Crypto Rule to Cut Red Tape for Startup Fundraising

    The U.S. Securities and Exchange Commission plans to introduce a major regulatory framework this month to simplify capital formation and reduce operational hurdles for cryptocurrency businesses.
    Martha Grizzard
    Jul 8, 2026
View more data 
BTCBTC(BTC)
$0
--(Last 24h)
SpotFutures

Top

View more
  1. 1S&P 500 Reclaims 200-Day Moving Average, Bitcoin Gains
  2. 2Trump Softens His Stance on Reciprocal Tariffs, US Stocks and Crypto Markets Rise
  3. 3Vitalik Buterin : The current price of ETH has not been affected by the merger event
  4. 4Vibhu Norby : Solana Spaces store to bring 100K people to Solana per month
  5. 5CZ: compared with the record high nine months ago, the current situation of the industry is much better

Top Gainers

View more
Fusionist
FusionistACE

$0.3220

+164.03%
Akedo
AkedoAKE

$0.0104

+64.76%
Kekius Maximus
Kekius MaximusKEKIUS

$0.006828

+56.18%
StaFi
StaFiFIS

$0.002541

+47.73%
Heima
HeimaHEI

$0.1738

+44.35%

Top Trending

View more
Fusionist
FusionistACE

$0.3206

+162.95%
Uniswap
UniswapUNI

$3.2030

-7.99%
Audiera
AudieraBEAT

$0.6384

-28.99%
OKB
OKBOKB

$106.710

+3.70%
Sandisk
SandiskSNDK

$1,656.35

+6.02%

Recently added

View more
KiiChain
KiiChainKII

$0.0672

-4.00%
GameStop
GameStopGMEB

$18.7100

+0.32%
Thinking Cat
Thinking CatHMM

$0.0132

-33.72%
StonkBroker
StonkBrokerSTONKBROKER

$0.0363

+24.48%
STONK
STONKSTONK

$0.009308

-12.98%

Learn

View more
  1. 1What Are ARC-20 Tokens? How Do ARC-20 Tokens Work?
  2. 2What Is the Usual Protocol? How Does Its Tokenomics Work?
  3. 3What Are AI Agent Frameworks? How Do They Power Cryptocurrency?
  4. 4What Is JPYSC? How Japan’s Regulated Stablecoin Works
  5. 5Are AI Agents Safe for Crypto? How to Secure Your Assets
About Us
  • About BitKan
  • Contact Us
  • Announcements
  • VIP Program
  • BitKan Ambassador
  • Institutional Services
Products
  • Spot
  • Futures
  • Crypto Prices
  • Learn
  • News
  • Markets
  • How to Buy Crypto
  • BTC to USD Calculator
  • Reward
Help
  • Help Center
  • Email Us
  • Live Chat
  • Download APP
  • Listing Application
  • Buy Bitcoin
  • Buy Ethereum
  • Buy Dogecoin
  • Buy Altcoins
Terms
  • Terms of Use
  • Privacy Policy
  • Trading Rules
  • Fee
K-Site
English
About Us
+
  • About BitKan
  • Contact Us
  • Announcements
  • VIP Program
  • BitKan Ambassador
  • Institutional Services
Products
+
  • Spot
  • Futures
  • Crypto Prices
  • Learn
  • News
  • Markets
  • How to Buy Crypto
  • BTC to USD Calculator
  • Reward
Help
+
  • Help Center
  • Email Us
  • Live Chat
  • Download APP
  • Listing Application
  • Buy Bitcoin
  • Buy Ethereum
  • Buy Dogecoin
  • Buy Altcoins
Terms
+
  • Terms of Use
  • Privacy Policy
  • Trading Rules
  • Fee
K-Site
+
  • Twitter
  • Facebook
  • Telegram
  • YouTube
  • Instagram
  • Medium
  • Linkedin
@2012-2026 BITKAN.com