9UXA_L confidently wrong novel
B-cell CLL/lymphoma 7 protein family member B · Q9BQE9 · RCSB 9UXA · AF-Q9BQE9-F1 (v6)
Blue where the experiment agrees with AlphaFold; amber-to-red where it diverges. The scale is anchored to absolute Ångströms, so hotspots are comparable across structures.
Per-residue accuracy vs. confidence
Reading along the protein chain: red is how far each residue sits from the experiment (Cα deviation in Å, higher = worse); green is local accuracy (lDDT×100); blue dotted is AlphaFold's own confidence (pLDDT). Stretches where confidence stays high but the red line is large are exactly where AlphaFold is confidently wrong.
Take-home: mean confidence pLDDT 87.65 vs. overall accuracy lDDT 0.81 and TM-score 0.40.
The metrics
Cα deviation: how far residue i sits from where the experiment places it, after superposing the whole chain. Δᵢ = |Pᵢ − (R·Qᵢ + t)| Å, with Pᵢ/Qᵢ the experimental/model Cα coordinates and R,t the best-fit rotation and translation.
per-residue lDDT: local accuracy at residue i without superposition — the fraction of i's neighbour distances (within 15 Å) the model preserves. lDDTᵢ = ¼ Σ_t 1[ |d_exp − d_model| < t ], t ∈ {0.5, 1, 2, 4} Å.
pLDDT: AlphaFold's confidence for residue i (0–100) — its own predicted lDDT, output by the network before seeing the experiment.
Is the confidence honest?
Each point is one residue: AlphaFold's predicted confidence (pLDDT, horizontal) against its actual accuracy (lDDT×100, vertical). Points on the dashed diagonal are perfectly calibrated; points well below it are overconfident — AlphaFold was surer than it should have been.
Take-home: pLDDT–lDDT correlation 0.1433 (near 1 = well calibrated; near or below 0 = confidence unrelated to, or opposite, real accuracy).
The metrics
pLDDT (x): AlphaFold's predicted per-residue confidence, 0–100. lDDT×100 (y): the accuracy actually achieved at that residue. Perfect calibration puts every point on the diagonal pLDDTᵢ = 100·lDDTᵢ.
Calibration correlation: the headline is the Pearson correlation of the two across all residues. r = cov(pLDDT, lDDT) / (σ_pLDDT · σ_lDDT) — near 1 means confidence tracks accuracy honestly; ≤ 0 means it does not.
Where the shape differs
The difference between the experimental and predicted residue–residue distance maps (Å). Bright regions mark pairs of residues whose separation AlphaFold got wrong — often a whole domain placed in the wrong position relative to the rest of the structure.
Take-home: mean distance-map difference 3.54 Å.
The metric
Distance-matrix difference: each cell is how much the separation of residues i and j differs between prediction and experiment. |Dᵢⱼ^exp − Dᵢⱼ^model|, where Dᵢⱼ = |rᵢ − rⱼ| is the distance between the two residues. Superposition-free, so a domain in the wrong place shows up as a bright off-diagonal block rather than being averaged away.
Did AlphaFold know it was wrong?
Left: AlphaFold's own predicted error (PAE, Å) for each residue pair. Right: the error we actually measured. Where the right panel is much brighter than the left, AlphaFold underestimated its own error.
Take-home: mean predicted error 7.77 Å vs. mean observed error 3.46 Å; 14.4% of residue pairs were more wrong than AlphaFold predicted.
The metrics
PAE (predicted): AlphaFold's Predicted Aligned Error — PAEᵢⱼ is the position error (Å) it expects for residue j when the structure is aligned on residue i, output by the network.
Observed error: the frame-invariant reality we measure for the same pair. Eᵢⱼ = | |rᵢ−rⱼ|_exp − |rᵢ−rⱼ|_model |. If the observed panel is far brighter than the predicted one, AlphaFold underestimated its own error — it was overconfident.
Per-domain breakdown
| Source | Domain | Range | Residues | TM | RMSD Å | mean Cα Δ | Name |
|---|---|---|---|---|---|---|---|
| PAE | PAE:1-202 | 1-202 | 44 | 0.16 | 6.96 | 6.22 |
All metrics
Global fold agreement
| TM-score (norm. experiment) | 0.40 |
| TM-score (norm. model) | 0.12 |
| TM-score (norm. shorter) | 0.40 |
| TM-score (norm. longer) | 0.12 |
| Cα-RMSD (Å) | 6.96 |
| backbone-RMSD (Å) | 6.84 |
| all-atom-RMSD (Å) | 7.40 |
| core-RMSD (Å) | 2.27 |
| core fraction | 0.23 |
| GDT_TS | 25.00 |
| GDT_HA | 7.39 |
| MaxSub | 0.36 |
| structural overlap (3.5 Å) | 0.20 |
Local, superposition-free
| lDDT | 0.81 |
| contact-map Jaccard | 0.84 |
| contact precision | 0.89 |
| contact recall | 0.94 |
| distance-matrix mean Δ (Å) | 3.54 |
| CAD-score (approx) | 0.80 |
Backbone & secondary structure
| SS agreement Q3 (%) | 56.82 |
| mean Δφ (°) | 25.90 |
| mean Δψ (°) | 33.50 |
| torsion within 30° (frac) | 0.57 |
| Rg experiment (Å) | 23.34 |
| Rg model (Å) | 22.50 |
| ΔRg (Å) | 0.83 |
Confidence calibration
| mean pLDDT | 87.65 |
| pLDDT↔lDDT Pearson | 0.14 |
| pLDDT↔lDDT Spearman | 0.15 |
| PAE↔observed Pearson | 0.24 |
| PAE overconfident frac | 0.14 |
| mean PAE (Å) | 7.77 |
| mean observed error (Å) | 3.46 |
Context & headline
| coverage of model | 0.22 |
| coverage of experiment | 1.00 |
| seq identity aligned (%) | 97.73 |
| confidently-wrong residue frac | 0.64 |
| FRAUD score | 0.36 |
Deposited 2025-05-13 · released 2026-05-06 · EM · 3.28 Å · closest pre-cutoff chain: none