Verification register AI & Agents
Current reading
Readiness band and full integer triple
AI-assisted assembly · derived results
Readiness band
Mature
Primary summary from verified readiness
- Confidence
- 68% · fresh
- Computed at
- 2026-07-30T10:51:45.076368+00:00
- Claimed
- 95
- Reported
- 95
- Verified
- 95
- Gap
- 0
Public ambition and stated capability
Observed practitioner reporting
Independently supported evidence
Claimed and verified align
Evidence strength Critical
Decision
What the current evidence supports
Human editorial judgment · 2026-07-30
Adopt with guardrails
- Why
- Anthropic's own pricing table confirms the vendor's 'half the price of Fable 5' claim exactly ($5/$25 vs $10/$50 per MTok), and both official and press sources corroborate near-Fable-5 coding/agent benchmark parity (CursorBench 3.2 within 0.5%, OSWorld 2.0 ahead at 1/3 cost) — but the same sources disclose concrete capability gaps in cybersecurity work that must be respected before broad or security-adjacent adoption.
- Next
- Pilot claude-opus-5 as the model backing conductor- and implementer-tier agents (per feedback_ultracode_agent_models: currently pinned to plain 'opus') for a cost/perf trial given its documented half-of-Fable-5 pricing, then gate any security-sensitive or exploit-adjacent automation behind Mythos-5/human review since Opus 5 explicitly trails on cybersecurity exploitation and cannot scan binaries.
Constraints
Blockers
No named blocker is present in the current public projection.
Evidence summary
Derived counts
AI-assisted assembly
- Total
- 5
- Tier 1
- 0
- Tier 2
- 1
- Tier 3
- 4
- Supports
- 3
- Contradicts
- 0
- Context
- 2
- Latest observed
- 2026-07-30
Counts and dates only. Raw signals, private excerpts, trust records, and internal corpus material are not published here.
Publication record
Revisions
Initial public reading
- 2026-07-30 Reading moved from mature to mature.