Home
Tech Grid
News Room
Interviews
Think Stack
Articles
  • Agentic AI

Agnes 2.5 Pro Alpha Debuts on Artificial Analysis with Low Cost


Agnes 2.5 Pro Alpha Debuts on Artificial Analysis with Low Cost
  • by: Business Wire
  • |
  • August 31, 2026

Every team building agentic products faces the same tradeoff: the models smart enough to run a task end to end are the ones that make the monthly bill flinch-worthy. Agnes 2.5 Pro Alpha ends that choice — the first Agnes AI text model on Artificial Analysis, the independent leaderboard, where its image and video models already rank top ten. Agnes AI is based in Singapore and trains its models in-house.

The model delivers frontier capability without the frontier bill, offering agentic performance at significantly lower cost.

Quick Intel

  • Agnes 2.5 Pro Alpha debuts on Artificial Analysis as first Agnes AI text model, image and video models already top ten.
  • Scored 39 on Intelligence Index v4.1 blending nine evaluations, ahead of Cohere Command A+ at 23 and Mistral Devstral 2 at 19.
  • Scores 67% on Terminal-Bench v2.1 for end-to-end command execution, ahead of DeepSeek V4 Pro, and 88% on GPQA Diamond.
  • Average task cost about three and a half cents on GDPval-AA v2; 10K tasks at $342 vs $12,300 on GPT-5.6 Sol and $37,000 on Claude Fable 5.
  • Pricing at $0.45 per million input tokens, $0.90 output with restraint of 26 turns and 38,000 tokens vs 3-4x for competitors.
  • Based in Singapore, trains models in-house; earliest version to reach Artificial Analysis with improvements planned.

Frontier Capability Measured on Independent Leaderboard

On the Intelligence Index v4.1, which blends nine evaluations, it scored 39 — the standout in its cohort, ahead of Cohere's Command A+ (23) and Mistral's Devstral 2 (19).

It scores 67% on Terminal-Bench v2.1, running commands and working a task end to end, ahead of DeepSeek V4 Pro, and 88% on GPQA Diamond, search-resistant graduate-level science, near the leading 94%. Write the code, drive the machine, think through the problem: that's the whole job.

On GDPval-AA v2, it ran the average task for about three and a half cents. Ten thousand tasks a month is $342 — versus $12,300 on GPT-5.6 Sol and $37,000 on Claude Fable 5. Two things drive the gap: a low price ($0.45 per million input tokens, $0.90 output), and restraint — 26 turns and 38,000 tokens where some models burn three to four times as many.

"Most developers are priced out of the best AI — not for lack of skill, but because the meter never stops," said Bruce Yang, Founder of Agnes AI. "We built Agnes to change that, in public, so anyone can hold us to it."

AI Parity for Developers and Startups Priced Out

Agnes 2.5 Pro Alpha is the earliest version to reach Artificial Analysis, not the finished product. It will improve over the coming months, each version measured independently and in public. That is AI parity: frontier capability for the developers, startups, and teams priced out of it. Free API access is where it starts.

 

About Agnes AI

Agnes AI builds full-modality foundation models across text, image, and video. Its mission is AI parity: frontier-grade AI accessible to developers, startups, and small teams.

  • Frontier AIAI AgentsTerminal Bench
News Disclaimer
Want to reach B2B tech decision-makers through TechIntelPro? Get our Media Kit
  • Share
Enterprise Tech News