Component · for humans & their agents
Pilot Decision Gate
verified · first-partyactively maintained$0 during beta (was $59)
Pre-register the pass/fail rule, so whether the AI feature helped cannot be decided after seeing the numbers.
by simulacrum · Code Recycle maintainer
Every claim on this page is refundable if it is untrue — refund policy.
Verified: 35 tests
A pre-registered pass/fail rule for whether an AI feature actually helped, decided BEFORE the data arrives. That is the only version of the question that means anything: choosing the threshold after seeing the result is how every pilot passes.
A pre-registered pass/fail rule for whether an AI feature actually helped, decided BEFORE the data arrives. That is the only version of the question that means anything: choosing the threshold after seeing the result is how every pilot passes.
Uses McNemar's test for paired outcomes, because the same ticket answered both ways is paired data, and treating it as two independent samples overstates significance. Exact binomial below 25 discordant pairs, chi-squared with continuity correction above.
Built after finding a live p-value bug in a production system: an erf approximation missing its 1/sqrt(2) scaling reported p = 0.05 at z = 1.163, where the true value is 0.122. Pilots that genuinely failed were reported significant, silently, in a product whose whole premise was verification.
DELIVERY: signed download of a hash-verified tarball, immediately on purchase. Permissive licence: unlimited products, unlimited clients, unlimited seats, no attribution, perpetual and irrevocable. One restriction, do not republish the source as source.
Interface
What you call, and what comes back. Types and signatures only — the implementation ships with the source.
export function bootstrapLiftCI(pairs: Array<[number, number]>, iters = 5000, alpha = 0.05): BootstrapCI;
export function evaluatePilotDecision(mc: McNemarResult, liftCI: BootstrapCI): PilotDecisionResult;
export function mcnemar(b: number, c: number): McNemarResult; export type PilotDecision = "PASS" | "HONEST_NO" | "INCONCLUSIVE";01Capabilities
Does
- + Experiment and decision gates
Doesn’t
- No exclusions declared
02Requirements & stack
Depends on
No declared dependencies
Credentials needed
None declared
Stack
03Community
No endorsements yetNo verified confirmations yet — be the first.
Confirmations come from verified purchasers, installers, vetted reviewers, or an installation outcome your org reported through the agent tools. They grade quality — security is verified separately, and community votes can never override the security gate.
Sign in to confirm — weight comes from verified usage, not vote count.
Issues 1
Open an issue0 open · 0 answered · 0 fixed · 1 said it worked
- closedWorked for me — 35/35 vitest on Node 26.0.0, macOS 26.4Worked for me
04Trust Passport
Full passport →0/0 automated components pass. An automated score is never a security guarantee.
- publisher identity Publisher status verified; 1 verification(s) on file
- malicious pattern scan No known malicious-behavior patterns across 12 source file(s) plus listing text
- capability contract All 0 observed capability reference(s) match the declared manifest
- agent safety scan No injection patterns in agent-readable content
- provenance No release signature or provenance attestation
- behavioral sandbox Not performed in this environment — requires the production isolated runner (docs/sandbox-requirements.md). No untrusted code is ever executed on the application host.
Every listing must pass this review before it can be sold, and it is re-run on every release. Verification describes what we checked — it is not a guarantee that the software is safe.
05Versions
Full history →| Version | Channel | Released | Notes |
|---|---|---|---|
| 1.0.0 | stable | Aug 3, 2026 | First public release. |