Skip to content
#

confidence-sequences

Here are 9 public repositories matching this topic...

Always-valid sequential A/B testing engine: mSPRT confidence sequences make peeking safe by construction, CUPED cuts variance up to 49%, SRM gates bad data. Built-in adversarial peeking harness proves the claim: naive daily peeking hit 27.5% false positives in 2,000 simulations; this engine held 1.7%. All numbers reproducible from committed seeds.

  • Updated Aug 4, 2026
  • Python

Runtime-verified fidelity for long-context LLM serving: detect, label-free, when sparse-attention KV compression silently degrades output, and bound it with anytime-valid confidence sequences. Custom HF attention backend + elastic probe scheduler. H1/H4 confirmed at 7B across Qwen2.5 & Mistral on 2x H100 · 88 tests · pre-registered hypotheses.

  • Updated Aug 26, 2026
  • Python

Improve this page

Add a description, image, and links to the confidence-sequences topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the confidence-sequences topic, visit your repo's landing page and select "manage topics."

Learn more