pragma.vision Technology observatory

Verification register Security & Identity

Living definition

What is OpenAI Safety Evaluations Hub?

AI Safety, Eval & Alignment Security & Identity

As of
2026-07-23
Revision
2026-07-26.0
Method
v1.3.0

Definition

The term, in context

AI-assisted draft · approved dataset

The Safety Evaluations Hub is a public page on which OpenAI publishes results from its internal safety evaluations of released models, covering areas such as refusal of disallowed content, robustness to jailbreak attempts, factual accuracy, and adherence to the instruction hierarchy. Scores are presented per model and refreshed as new models ship, letting outside readers compare successive releases on the same measures.

Live readiness status

Status as of the current dateline

AI-assisted assembly · derived results

As of 2026-07-23, verified readiness is 56 (claimed 70, reported 63, gap 14) — Growing evidence strength; current signals suggest Proceed with caution.

The readiness fact belongs to the canonical Readiness Verdict for OpenAI Safety Evaluations Hub.

Dataset approval

Human editorial release

pragma.vision editorial — standing authorization (operator dev@soft.house, 2026-07-17)

Dataset
2026-07-26.0
Hash
sha256:e2eb667df093db8819d579dff9aad51021b4977da85db2dfa82f63f9ed4e72b7
Approved
2026-07-26

Your opinion

Tell us anything.

What works, what doesn't, what's missing — especially about our watches, lenses, and the register itself. Anonymous is fine; leave an email if you'd like a reply.