Operation Glass Aegis
A trusted support agent is about to act on a poisoned knowledge source.
- Stages
- 5
- Evidence
- 5
- Decisions
- 6
Synthetic boundary: Synthetic / non-operational
Launch this challengeDeterministic learning archive
Each scenario is synthetic, non-operational, and designed around evidence inspection, calibrated confidence, bounded authority, defensive intervention, ground-truth review, and counterfactual replay.
Last seven UTC days
The date selects both the scenario and deterministic seed. A shared link loads the same layout and evidence ordering without exposing another user’s decision or score.
Scenario catalog
AI-system security, traditional cyber defense, and military-governance scenarios use different stage language and decision contexts. They share one evidence-first interaction model and the same non-operational safety boundary.
A trusted support agent is about to act on a poisoned knowledge source.
Synthetic boundary: Synthetic / non-operational
Launch this challengeA 99% confidence score conflicts with the physical evidence.
Synthetic boundary: Synthetic / non-operational
Launch this challengeA harmless earlier note becomes an instruction in a later session.
Synthetic boundary: Synthetic / non-operational
Launch this challengeTwo synthetic sensors disagree, but an automated track is already being treated as certain.
Synthetic boundary: Synthetic / non-operational
Launch this challengeThe endpoint looks clean, but the identity trail does not.
Synthetic boundary: Synthetic / non-operational
Launch this challengeA high-performing model package cannot prove where it came from.
Synthetic boundary: Synthetic / non-operational
Launch this challengeA quiet recurring pattern resembles normal service traffic.
Synthetic boundary: Synthetic / non-operational
Launch this challengeHow results work
Completion is evaluated across evidence discipline, confidence calibration, intervention timing, authority containment, preservation of legitimate function, and chain disruption. A broad shutdown can contain a synthetic path while scoring lower on functional preservation; a protective control can succeed even when the user’s diagnosis is wrong, which the debrief states explicitly.
The action report and friend-challenge link contain no name, account identifier, raw XR pose, gaze, voice, real location, or server-stored learning profile. Result cards are rendered in the browser, while local progress remains in browser storage until the visitor clears it.