Use cases· Last updated

Compliance confidence thresholds with Jev

SOC 2 / ISO rows have owners, cadence, and pass/fail. Jev can flag “this evidence blob never mentions production.” The control owner still attests. Jev is not the auditor and not the GRC system of record.

This unofficial page is the confidence thresholds slice of the compliance evidence pre-score pack. Intent: apply the Jev (TypeSafe System One) decision model to compliance evidence pre-score confidence thresholds. Primary search language: Compliance Jev confidence thresholds. Confirm patterns on docs.typesafe.ai. This site does not sell, issue, or proxy TypeSafe keys. Use a credential you already have from the console or a documented gateway.

Independent angle (cover ≠ clone): Checklist is the control; Jev pre-scores whether evidence language looks complete. Owner still signs — not a GRC-clone or rival checklist-IA photocopy. Treat floors as production gates and price the false-reject cost — official 0.5/0.9 sketches are illustrations.

Compliance use-case context

Thresholds turn compliance evidence pre-score answers into act / review / abstain. They are product policy, not a hyperparameter TypeSafe ships. Official 0.5 / 0.9 sketches are illustrations. This slice also carries the false-reject discussion: over-gating compliance evidence pre-score hides calibration.

Hub: Use cases. Compare, when the other tool is the real job: compliance checklists.

Confidence Thresholds inputs

You need (1) pinned answers on a frozen contract and (2) labels for gap / ready / review gold from control owners, plus production-mention gold. State shape:

{
  "control": { "id": "CC-6.1", "prompt": "Evidence must describe production access reviews this quarter." },
  "evidence": { "title": "Access review export", "text": "We reviewed staging users in January." },
  "policy": { "env": "Staging-only language is a gap for production controls." }
}

Decision signals and actions

Axis Where it lives Compliance use
choice / score / noul answer payload What to do with the control prompt + evidence blob
confidence Choice & Score only Whether to trust the argmax
Distance from 0.5 Noul Whether mentions_prod is decided
FLOORS = {
    "nudge_owner": 0.50,      # illustrations — replace
    "mark_control_passed": 0.98,
}
NOUL_TAU = 0.70  # for mentions_prod

def allow(ans, action):
    return ans.confidence >= FLOORS[action]

Do not treat a Noul of 0.5 as a “medium” compliance evidence pre-score score — it means yes and no are equally likely. Conjunctions stay in your code.

Guardrails and escalation

TypeSafe’s confidence-gated examples use a lower bar for recoverable reads than for irreversible actions. Those numbers are illustrations. For compliance evidence pre-score, treat mark_control_passed as the high bar (marking a control passed or signing attestation). Tune on labels — see offline evaluation.

Band around 0.5 on mentions_prod always reviews. Do not copy 0.70 onto Choice confidence.

Evaluation and rollout notes

Fit loop: pin jev-1.13.0 → replay → plot error vs confidence → pick floors where auto-act error ≤ your SLA. Pin jev-1.13.0 (the versioned id) after you fit thresholds. jev-latest and the marketing line jev-1.13 can move. Log the response model. TypeSafe’s published list price for jev-1.13 is $0.042 per million input tokens (vendor claim — confirm on the models page); output tokens are free on that same page. Unused distractors still bill as input.

Official Python and JavaScript SDKs read TYPESAFE_API_KEY and retry documented 429/529. This site does not sell, issue, or proxy TypeSafe keys. Use a credential you already have from the console or a documented gateway.

Pack map

Slice Page
Graph and primitives decision workflow
What may enter state input contracts
What to gather first evidence collection
Atomic rules policy checks
Act / review / abstain you are here
Reviewer payload human handoff
What to persist audit trail
How it breaks failure modes
Labeled replay evaluation
Shadow → canary production rollout

FAQ

Should mark_control_passed use 0.9 everywhere? No. Over-gating hides calibration and dumps the queue on humans. Fit per action.

Can I reuse a Noul τ as Choice confidence? No. Jaggedness: they are not interchangeable. See confidence.

Where is the rest of the Compliance pack? Start with Compliance decision workflow and Compliance human handoff. Cluster hub: Use cases.

Can Jev be our auditor? No. It pre-scores language. Owners and auditors sign. See governance.

May we send screenshots? Not as images. Transcribe what the screenshot shows, then ask snap questions.

What this page does not claim

Disclaimer

This is an independent unofficial site and is not affiliated with TypeSafe AI; official documentation is available at https://docs.typesafe.ai.

Primary documentation: https://docs.typesafe.ai. Hub: Use cases.

Sources

Public TypeSafe or adjacent documentation only. No private claims.