EuCAIFCon 2026
Raphaël Bonnet-Guerrini1,2, Johann Ioannou-Nikolaides3, Inar Timiryasov3, Vincenzo Piuri1
1Università degli Studi di Milano · 2INFN Milano · 3Niels Bohr Institute
02 · Promise and limits
Improve model, Learn from them?
Amplify, suppress, or steer them.
Sparse ⇒ interpretable ⇒ causal
How to find interpretable concepts in Physics foundation models?
03 · Model
04 · Method
05 · Atlas
| Read-out | Physics role | Evidence |
|---|---|---|
| \(z_{bc}\) | bright + clean | AUROC .91 · 4/4 seeds |
| \(z_{bc}'\) | clean, secondary | controls · 4/4 |
| \(z_{aux}\) | auxiliary activity | AUROC .84 · 3/4 |
| \(z_{depth}\) | detector depth | AUROC .986 · layer 2 |
06 · Diagnosis
| Intervention | Direction head | Uncertainty head |
|---|---|---|
| remove \(z_{bc}\) | +0.06° null | +0.31 causal |
| remove clean pair | +0.03° null | +0.32 causal |
| remove clean family (194) | −0.02° null | +0.78 causal |
| remove clean core | +0.02° null | +0.37 causal |
| scale \(z_{bs}\) | V-shaped support | monotone 3/4 |
Backbone and direction head stay frozen, they both heads read the same \(\mathrm{CLS}\) state.