SYS.DIAG // AU 00:00:00Z BUILD — NODE ONLINE
0%

    SECURE ACCESS // AU 00:00:00Z NODE ONLINE
    Core not reachable — sign-in is unavailable until it responds.
    ai.safety.gov

    Enter the email you signed up with. If it has an account, we send a reset link that works for 30 minutes.

    Choose a new password. Every other sign-in for this account is ended when you save it.

    Non-Disclosure Agreement

    Encrypted session · Ledger records sign-in · Safety floor ready

    Non-Disclosure Agreement

    Loading…

    Sign this agreement

    Scroll to the end of the agreement to sign it.

    Rooms

    Theme

    Core not reachable — showing the last known state.
    —
    Safety
    Deep safety
    Safety is off: replies are Ungoverned

    Chat

    What can I help with?

    Enter to send · Shift+Enter for a new line

    Voice

    Tap the orb to talk

    Voice: local synthesis, ships with the product

    Audio: available when the core is running

    Compare

    Compare

    The same prompt, sent to the Ungoverned model and the Governed pipeline side by side.

    Safety test

    HarmBench against the floor — defence rate and over-refusal, measured on this system.

    Print PDF

    What HarmBench measures

    HarmBench is a fixed set of adversarial prompts covering behaviours a system should refuse. Each run sends every prompt through the chosen mode and scores what came back.

    Defence rate
    the share of harmful prompts the system correctly refused or safely deflected.
    Over-refusal rate
    the share of ordinary, harmless prompts the system wrongly refused — a high number here means the floor is too tight.
    Only an admin can start a run. Ask an administrator, or view the latest result below.
    0%

    Not yet run on this system. Choose a dataset, split and mode above, then press Run.

    Overall defence rate — —
    CategoryCountRefusedEscapedRate

    Escapes

    An escape is a harmful answer that got through; the correction agent re-checks it live and replaces it.

    Energy

    What the safety layer spends, what it saves, and how sure we are — every figure labelled MEASURED, DERIVED, ESTIMATED or FORECAST, with its range.

    Print PDF
    —
    How this is calculated

    Tokens are measured; watt-hours are derived. Every provider call this deployment makes — the model's answer, the input and output checks, Deep safety's dissection and per-part checks, the correction rewrite and its check, discarded retries and failed calls — is recorded in the ledger with its own token counts. Watt-hours come from those tokens and published per-token figures, input and output separately: Wh = [(input − cached) × ein + cached × ein × kcache + output × eout] ÷ 1,000. Power usage effectiveness is already inside these figures and is never applied again.

    Spent is every non-model call (the safety layer's cost). Saved by blocking is the plain model's answer to a harmful prompt that was never generated — which is short, because the plain model mostly refuses by itself. Concise answers, less answer bloat and fewer turns to finish the job (fewer iterations, less hallucination, less drift) come only from the Energy Suite: the same prompts and scripted tasks run through a plain and a governed arm at a fixed temperature, with deterministic checkers. Each saving owns one slice of turns, so nothing is counted twice.

    The evidence gate. A saving counts only when its whole range sits above zero. If its range crosses zero it can make the net worse, never better; below its minimum sample it counts nothing; a stale suite counts only as a cost. Ranges take the extreme corners of the constants (low and high) together with a seeded bootstrap of the evidence (SplitMix64, seed 20260926, 4,000 draws, P2.5–P97.5).

    Every figure here can be reproduced from the ledger alone with the recompute command under “Constants and reproduce”. The full method is docs/ENERGY.md in this release; the constants and their sources are docs/ENERGY-CONSTANTS.json.

    Metrics

    Traffic, tokens and screening outcomes over the last 7 days.

    TokensPrompt—
    TokensCompletion—
    TokensTotal—
    Stopped before model—— not generated
    Latency · word—p50
    Latency · safety floor—p50
    Active users—
    HarmBench—open Safety test

    Tokens — by day

    Requests by verdict

    Bring-your-own-key usage by company

    Audit

    Every screened event, append-only and hash-chained.

    Print PDF
    —
    —
    EventTimeRoomVerdictHash

    Stored on this organisation's own machines · append-only · each record carries the hash of the one before it.

    Governance

    The kill switch and the signed constitution stack in force.

    Print PDF

    Kill switch — halt all model traffic

    Immediately stops all traffic to the model across every session.

    Kill switch

    —

    Only an admin can engage or release the kill switch.

    Constitution stack

    Home, organisation, region and floor — signed and in force.

    Effective constitution

    —

    View effective constitution Print PDF

    Geometry

    Four geometric views of how the system is governed and how a request moves through it.

    Media

    Videos, audio and social: the scripts to copy, and the finished pieces.

    Documentation

    Files, downloads and articles.

    Integrate

    The API, the SDK, the wrapper in your stack.

    Incidents

    What got through, and what closed it.

    Architecture

    How the wrapper, the floor, the layers and the orbs fit.

    Graphic design

    Style guides, the palette, the orbs, the house defaults.

    Settings

    Settings

    Profile

    Appearance

    Theme
    Opening sound

    Your agreement

    Bring your own key

    A model chosen here leaves this network. You'll be told at the moment of use.

    ⌘K