pragma.vision Technology observatory

Verification register AI & Agents

Readiness verdict

Claude Opus 5

A dated reading of what is claimed, reported, and independently verified in the current evidence.

As of
2026-07-30
Revision
1
Method
v1.3.0

Current reading

Readiness band and full integer triple

AI-assisted assembly · derived results

Readiness band

Mature

Primary summary from verified readiness

Confidence
68% · fresh
Computed at
2026-07-30T10:51:45.076368+00:00
Claimed
95

Public ambition and stated capability

Reported
95

Observed practitioner reporting

Verified
95

Independently supported evidence

Gap
0

Claimed and verified align

Evidence strength Critical

Decision

What the current evidence supports

Human editorial judgment · 2026-07-30

Adopt with guardrails

Why
Anthropic's own pricing table confirms the vendor's 'half the price of Fable 5' claim exactly ($5/$25 vs $10/$50 per MTok), and both official and press sources corroborate near-Fable-5 coding/agent benchmark parity (CursorBench 3.2 within 0.5%, OSWorld 2.0 ahead at 1/3 cost) — but the same sources disclose concrete capability gaps in cybersecurity work that must be respected before broad or security-adjacent adoption.
Next
Pilot claude-opus-5 as the model backing conductor- and implementer-tier agents (per feedback_ultracode_agent_models: currently pinned to plain 'opus') for a cost/perf trial given its documented half-of-Fable-5 pricing, then gate any security-sensitive or exploit-adjacent automation behind Mythos-5/human review since Opus 5 explicitly trails on cybersecurity exploitation and cannot scan binaries.

Constraints

Blockers

No named blocker is present in the current public projection.

Evidence summary

Derived counts

AI-assisted assembly

Total
5
Tier 1
0
Tier 2
1
Tier 3
4
Supports
3
Contradicts
0
Context
2
Latest observed
2026-07-30

Counts and dates only. Raw signals, private excerpts, trust records, and internal corpus material are not published here.

Publication record

Revisions

Initial public reading

  1. 2026-07-30 Reading moved from mature to mature.

Your opinion

Tell us anything.

What works, what doesn't, what's missing — especially about our watches, lenses, and the register itself. Anonymous is fine; leave an email if you'd like a reply.