vNext Screening Engine in Production, with a Rebuilt Benchmark & Site
- The vNext decision engine is now the authoritative engine in production. Measured on the live corpus of 1,325,118 listed records across 53 sources, it produced 0 severe false positives across 5,000 adversarial negatives (95% CI ≤ 0.077%).
- Sensitivity to real listed parties held: exact-name recall was 97.33% (95% CI 94.83–98.64%, n=300), and all 300 sampled sanctions positives were actionable.
- Weak resemblance now resolves to a low-confidence watch signal (MONITOR → low) rather than a match, so near-miss negatives concentrate in the review band instead of surfacing as sanctions-grade results.
3 additional release notes
- Political exposure stays distinct from sanctions: across 150 PEP-only cases, 0 were served at sanctions-grade severity.
- Semantic (FAISS) vector retrieval was built, measured on an identical corpus, and left disabled — on this corpus it cost precision for a recall gain that was statistically indistinguishable from zero.
- The public landing page and /benchmark were rebuilt around the deployed engine. Every published figure is driven by a single typed source of truth, and superseded synthetic-benchmark numbers were moved into a clearly labelled historical archive. Availability is reported separately and is never counted as a false positive.