individual
Technical AI safety research
| BlueDot Impact | $200 | 100% | 1 |
| Date ↓ | Funder | Via | Amount | Cause | Source | Purpose |
|---|---|---|---|---|---|---|
| Jul 2026 | BlueDot Impact | $200 | Interpretability, Evals | bluedot | Mitigations for emergent misalignment can hide it behind contextual triggers instead of removing it. I'm testing if white-box tools(NLAs, Jacobian lens, probes) can catch it when standard evals can't. |