Validation Report

Two real runs: a labeled target (DLL3) used to validate the funnel on known binders, and an unlabeled target (BCMA) run in production to a ranked shortlist with real structure-based scoring. Every score is a model prediction; wet-lab binding is the only truth.

Generated 2026-07-05

DLL3 — recovery gate (validate before trusting)
A funnel that runs is not a funnel that works.
263,958
Pool (unique)
21
Known positives
0 / 21
Recall (first pass)
21 / 21
Recall (after fixes)

Three real causes found and fixed:

  • A constant-region tag on ~half the pool clones (absent from the reference binders) split the same molecule into two clusters.
  • CDR3 extraction failed on ~36% of real sequences, collapsing clustering.
  • The abundance-ranked budget cut discarded low-abundance real binders.
BCMA — production run
1,200,612
Pool
4,000
Stage 1 · sequence filter
400
Stage 2 · folded + scored
100
Shortlist
400
Folds
0.69 nM
Median predicted Kd
58 / 100
Sub-1 nM
97 / 100
Under 10 nM

Each folded complex gives two near-independent signals — interface confidence and a contact-based predicted Kd. Ranking on interface confidence alone versus the consensus produced top-10 lists that shared zero candidates, so the funnel ranks on the consensus: a candidate must be strong on both.

RankCombined scorePredicted Kd (nM)Abundance
10.9600.080.0414
20.9760.070.0069
30.9350.130.0005
40.8800.120.0003
50.8820.160.015
60.8850.250.0003
70.8830.190.0357
80.9110.030

Many top hits are low-abundance — the funnel surfaces rare-but-strong binders that pure NGS-count ranking would bury.

Honest limits

Predicted, not measured. The shortlist is where to spend expression capacity, not a validated affinity table.

The Kd enrichment is partly circular — the ranking weighs predicted binding energy. The independent interface signal hedges this, but it is not external validation.

Specificity needs counter-screens. BCMA now includes TACI and BAFF-R counter-select constructs; specificity calls are only meaningful when the real folding backend scores both target and counter-screen constructs.