organization
Technical AI safety researchAI safety
| Long-Term Future Fund | $67,000 | 85% | 1 |
| Neel Nanda | $11,000 | 14% | 1 |
| BlueDot Impact | $1,000 | 1% | 1 |
| Date ↓ | Funder | Via | Amount | Cause | Source | Purpose |
|---|---|---|---|---|---|---|
| Jun 2026 | BlueDot Impact | $1,000 | AI safety | bluedot | Directing an MSc thesis on how unverbalized biases enter LLM's chain-of-thoughts: finding the 'bias anchor' sentences and the attention path from the model's input to its output. | |
| Nov 19, 2024 | Neel Nanda | Manifund | $11,000 | Interpretability | manifund | Mechanistic Interpretability research for unfaithful chain-of-thought (1 month) |
| Jan 2024 | Long-Term Future Fund | EA Funds | $67,000 | Interpretability | ea_funds | 4-month stipend for MATS extension on mechanistic interpretability benchmark + 2-month stipend for career switch |