Safety Exchange — Specification 001
The working contract for coordinating bounded safety signals without sharing raw evidence or transferring enforcement authority.
Turnkeeper Labs · applied research
Turnkeeper Labs studies how machine learning can help people interpret risk and preserve provenance — without turning model output into automatic enforcement.
Synthetic-first · Held-out closed unless stated · Reviewable evidence
Help reviewers understand bounded safety signals in context.
Evaluate where models help, fail, or require specialist judgment.
Give reviewers clearer evidence without automating accusations.
Research method
Research begins with the conditions people actually face—not with a model looking for a use.
Work with practitioners to identify the decision, evidence, and human responsibility involved.
Prototype with synthetic or privacy-minimized data that can be inspected and stress-tested.
Document failure modes, boundaries, and unresolved questions before describing progress.
Open questions
Where AI can strengthen human review—and where it must not take over.
A synthetic benchmark study of contamination, decision changes, transferred risk, and recovery after one false identity link.
Follow the researchSource dependence, provenance displays, and confidence inflation under synthetic lineage controls.
Open the studyFrozen adversarial cases scoring naive multiplicity against lineage-aware reconstruction.
Open WardBenchSynthetic calibration runs test where models help, fail, or require specialist judgment.
View the roadmapPrivacy-minimized patterns and counterevidence help specialists review what a bounded signal means.
Explore Ward evidence systemsSynthetic demonstrations bind a recommendation to an exact, reviewable effect without granting it authority.
Open the demonstrationNo experiment authorizes enforcement on its own.
Evidence from the work
The working contract for coordinating bounded safety signals without sharing raw evidence or transferring enforcement authority.
Synthetic calibration plans, research gates, and the questions still being tested.
How a recommendation becomes a bounded request that still requires explicit authorization.
The public building blocks that make provenance and review easier to verify.
Boundaries before breakthroughs
Experiments remain clearly labeled until code, tests, and operating evidence support them.
Research starts with bounded data and does not expand collection by default.
Models may surface evidence or recommendations. Consequential judgment stays human.
Bounded, synthetic, reviewable research designed to strengthen human judgment.
A surveillance system, live accusation engine, or source of autonomous enforcement.
Work with Turnkeeper Labs on research that can be bounded, inspected, and improved with the people responsible for protecting children.
Turnkeeper Labs is an initiative of Turnkeeper.