individual
Technical AI safety research
| BlueDot Impact | $100 | 100% | 1 |
| Date ↓ | Funder | Via | Amount | Cause | Source | Purpose |
|---|---|---|---|---|---|---|
| Jun 2026 | BlueDot Impact | $100 | Interpretability, Alignment methods, Security | bluedot | Mechanistically understanding if RLHF/DPO unlearning are susceptible to jailbreak attack and extending current research to synchophancy |