individual
OtherTechnical AI safety research
| BlueDot Impact | $1,600 | 100% | 3 |
| Date ↓ | Funder | Via | Amount | Cause | Source | Purpose |
|---|---|---|---|---|---|---|
| Aug 2026 | BlueDot Impact | $500 | Technical AI safety research | bluedot | Attending TARA Delhi (Technical Alignment Research Accelerator), a 14-week alignment training cohort, to close implementation gaps in my ongoing Narrow Secret Loyalty research. | |
| Jul 2026 | BlueDot Impact | $950 | Other | bluedot | Scaling our RL-CAI pipeline (extends narrow secret-loyalty models via Constitutional AI) from 1.5B to a 32B organism, to test whether action-breadth extension holds at frontier model scale. | |
| Jul 2026 | BlueDot Impact | $150 | Interpretability, Alignment methods | bluedot | Applying CAI (SL-CAI+RL-CAI) to an SFT-installed secret loyalty organism to characterize how RL changes loyalty breadth in the activation×action taxonomy and whether HH-RLHF mitigates it. |