individual
Technical AI safety research
| BlueDot Impact | $150 | 100% | 1 |
| Date ↓ | Funder | Via | Amount | Cause | Source | Purpose |
|---|---|---|---|---|---|---|
| Jul 2026 | BlueDot Impact | $150 | Alignment methods | bluedot | Investigating value and sycophancy in reward models: how much of reward model stereotyping stems from RLHF, and how much is inherited from base models? Sycophancy, or mere association? |