organization
Technical AI safety research
| Long-Term Future Fund | $436,461 | 100% | 6 |
| Date ↓ | Funder | Via | Amount | Cause | Source | Purpose |
|---|---|---|---|---|---|---|
| Apr 2023 | Long-Term Future Fund | EA Funds | $30,000 | Interpretability | ea_funds | This grant provides 5 months of funding for office space for collaboration on interpretability/model-steering alignment research |
| Apr 2023 | Long-Term Future Fund | EA Funds | $40,000 | Interpretability | ea_funds | Conference publication of interpretability and LM-steering results |
| Jan 2023 | Long-Term Future Fund | EA Funds | $220,000 | Interpretability | ea_funds | Year-long salary for shard theory and RL mech int research |
| Jan 2023 | Long-Term Future Fund | EA Funds | $115,411 | Interpretability | ea_funds | General support for Alexander Turner and team research project - Writing new motivations into a policy network by understanding and controlling its internal decision-influences |
| Jan 2022 | Long-Term Future Fund | $1,050 | Foundations | vipul_donations | The grants database gives the following intended use of funds: "Ad campaign for "Optimal Policies Tend To Seek Power" to ML researchers on Twitter" | |
| Mar 20, 2019 | Long-Term Future Fund | $30,000 | Foundations | vipul_donations | Grant for building towards a “Limited Agent Foundations” thesis on mild optimization and corrigibility. Grantee is a third-year computer science PhD student funded by a graduate teaching assistantship; to dedicate more attention to alignment research, he is applying for one or more trimesters of funding (spring term starts April 1). |