pragma.vision Technology observatory & strategic foresight

Verification register Security & Identity

Living definition

What is ForesightSafety Bench?

AI Safety, Eval & Alignment Security & Identity

As of
2026-08-20
Revision
2026-08-17.1
Method
v1.3.0

Definition

The term, in context

AI-assisted draft · approved dataset

ForesightSafety Bench is a benchmark suite for measuring how well an AI system anticipates the downstream consequences of a proposed action before taking it, scoring models on scenarios where an immediate step looks benign but leads to harm. It presents outcomes as comparable scores across models, so safety-relevant foresight becomes something measured rather than assumed.

Live readiness status

Status as of the current dateline

AI-assisted assembly · derived results

As of 2026-08-20, verified readiness is 51 (claimed 60, reported 56, gap 9) — Weak evidence strength; current signals suggest Track; not yet.

The readiness fact belongs to the canonical Readiness Verdict for ForesightSafety Bench.

Dataset approval

Human editorial release

pragma.vision editorial — standing authorization (operator dev@soft.house, 2026-07-17)

Dataset
2026-08-17.1
Hash
sha256:0c37108a182f510fc8cf4d1dfc3a30339f2b59e62654201edea2f530a268707e
Approved
2026-08-17

Your opinion

Tell us anything.

What works, what doesn't, what's missing — especially about our watches, lenses, and the register itself. Anonymous is fine; leave an email if you'd like a reply.