이번엔 미국 선두가 값을 내렸다.
그동안 값 싸움은 중국 몫이었는데..
앤트로픽이 새 모델을 내놨는데, 성능은 최상위권인데 값이 절반이다.
앤트로픽이 현지시간 7월 24일 '클로드 오퍼스5(Claude Opus 5)'를 공개했다.
AI 평가기관 아티피셜 어낼리시스의 종합 지능 지수에서 61점을 받았다. 앤트로픽 자사 최상위 모델 페이블5(60점)보다 1점 높다.
디지털타임스는 이 점수를 두고 '현존 공개 모델 중 왕좌'라고 전했다.
그럼 값은 얼마나 싼가요?
핵심은 여기다.
- 오퍼스5 — 100만 토큰당 5~25달러
- 페이블5 — 100만 토큰당 10~50달러 (정확히 두 배)
- GPT-5.6 솔 — 100만 토큰당 5~30달러
페이블5의 딱 절반이다. 종합 점수는 페이블5보다 높은데 값은 반이라는 얘기다.
성능이 진짜 다 이긴 건가요?
그건 아니다. 여기가 중요하다.
종합 점수는 1위지만, 세부 지표는 엇갈린다.
- 지식업무(GDPval-AA) 1861점·에이전트 검색 90.8% — 오퍼스5가 최고
- 에이전트 코딩(딥SWE) 68.8% — GPT-5.6 솔(72.7%)이 앞선다
- 전문가 추론(HLE) 56.3% — 페이블5(56.5%)가 근소하게 앞선다
즉 '전부 1등'은 아니다. 코딩이나 고난도 추론은 위에 있는 모델이 아직 있다.
내 생각엔, 오퍼스5의 진짜 무기는 최고 점수가 아니라 '이 정도 성능을 이 값에'라는 조합이다.
앤트로픽은 뭘 노리는 걸까
앤트로픽은 오퍼스5를 두고 '장시간 돌아가는 에이전트를 위한 큰 개선'이라고 설명했다. 코딩과 전문 업무 성능을 끌어올렸다는 것이다.
라인업도 촘촘해졌다. 미토스·페이블(고성능) 아래에 오퍼스(중상급), 그 아래 소네트(실무형), 하이쿠(최경량)로 이어진다.
디지털타임스가 인용한 업계 관계자는 이를 두고 '고비용 AI에 부담을 느끼던 기업 고객을 빠르게 흡수하려는 승부수'라고 했다.
며칠 전까지 값 파괴는 중국 오픈웨이트 모델 얘기였다. 이번엔 미국 선두가 스스로 값을 반으로 내렸다.
값이 내려가면 따라오는 것
AI 요금이 또 내려갈 여지가 생겼다.
최상위 성능을 절반 값에 쓸 수 있으면, 챗봇을 붙인 회사 입장에선 굳이 두 배 비싼 모델을 고집할 이유가 준다.
물론 코딩처럼 여전히 더 비싼 모델이 앞서는 영역도 있다. '무조건 오퍼스5'는 아니고, 쓰는 일에 따라 갈린다.
그래도 흐름은 분명하다. 성능 경쟁이 이제 '가격당 성능' 싸움으로 넘어가고 있다..
올 하반기엔 내가 쓰는 서비스 뒷단 모델이 조용히 바뀌는 걸 몇 번 더 보게 될 것 같다.
This time it was a US leader that cut the price.
Until now, price wars were China's game..
Anthropic just released a new model that scores near the very top — at half the price.
On July 24 local time, Anthropic released Claude Opus 5.
On Artificial Analysis' intelligence index it scored 61 — one point above Anthropic's own top model, Fable 5 (60).
Digital Times called that score 'the throne among publicly available models.'
So how much cheaper is it?
This is the core of it.
- Opus 5 — $5–25 per million tokens
- Fable 5 — $10–50 per million tokens (exactly double)
- GPT-5.6 Sol — $5–30 per million tokens
Exactly half of Fable 5. A higher overall score than Fable 5, at half the price.
Does it actually beat everything?
No — and this matters.
The overall score is first, but the sub-benchmarks are mixed.
- Knowledge work (GDPval-AA) 1861, agentic search 90.8% — Opus 5 leads
- Agentic coding (DeepSWE) 68.8% — GPT-5.6 Sol (72.7%) is ahead
- Expert reasoning (HLE) 56.3% — Fable 5 (56.5%) edges it out
So it isn't first at everything. For coding and the hardest reasoning, pricier models still lead.
To me, Opus 5's real weapon isn't the top score — it's the combination of 'this much performance at this price.'
What is Anthropic after?
Anthropic described Opus 5 as 'a step change improvement for long-running agents,' with gains in coding and professional work.
The lineup got denser too: below Mythos and Fable (high-end) sits Opus (upper-mid), then Sonnet (workhorse) and Haiku (lightest).
An industry source quoted by Digital Times called it 'a bet to quickly absorb enterprise customers weary of high-cost AI.'
Days ago, price disruption was a Chinese open-weight story. This time a US leader halved its own price.
What follows when the price drops
There's room for AI prices to fall again.
If you can get top-tier performance at half the cost, a company running a chatbot has less reason to insist on a model twice as expensive.
Of course, some areas — like coding — still favor pricier models. It's not 'always Opus 5'; it depends on the job.
But the direction is clear. The race is shifting from raw performance to performance-per-dollar..
I suspect we'll see the model behind our everyday services quietly swapped a few more times this half.
Sources · Anthropic Blog · Digital Times · Financial News