Verification register AI & Agents
Current reading
Readiness band and full integer triple
AI-assisted assembly · derived results
Readiness band
Watch
Primary summary from verified readiness
- Confidence
- 58% · fresh
- Computed at
- 2026-09-22T00:11:54.651334+00:00
- Claimed
- 90
- Reported
- 80
- Verified
- 69
- Gap
- +21
Public ambition and stated capability
Observed practitioner reporting
Independently supported evidence
Claimed leads verified
Evidence strength Strong
Decision
What the current evidence supports
Human editorial judgment · 2026-09-22
Proceed with caution
- Why
- A frontier-scale open-weight release (MIT license, 552B MoE backbone / 763.2B total incl. Engram parameters, native vision, 1M-token context, novel Causal Encoder-Decoder architecture) with strong headline benchmarks (GPQA Diamond 90.9, Terminal-Bench 2.1 90.6) is genuinely notable and independently corroborated, but concrete evidence shows real gaps versus Opus 5 on newer agentic benchmarks, prohibitive self-host hardware requirements, and a rocky, criticized rollout — not yet a safe default adoption six days post-release.
- Next
- Track independent evaluations (Vals.ai, LMArena, agentic long-horizon benchmarks) and self-host cost/availability of GB200/H200-class hardware before adopting for any Pragma.Vision inference path; re-check once the V4-Pro deprecation (Sept 14 cutover) and vLLM tooling maturity settle.
Constraints
Blockers
No named blocker is present in the current public projection.
Evidence summary
Derived counts
AI-assisted assembly
- Total
- 6
- Tier 1
- 0
- Tier 2
- 4
- Tier 3
- 2
- Supports
- 2
- Contradicts
- 3
- Context
- 1
- Latest observed
- 2026-09-11
Counts and dates only. Raw signals, private excerpts, trust records, and internal corpus material are not published here.
Publication record
Revisions
Initial public reading
- 2026-09-21 Reading moved from watch to watch.