We are the leading talent accelerator for beneficial AI and societal resilience. We run courses, help people land jobs and start new organizations.

London / SF
BlueDot Impact retweeted
With support from @MATSprogram and @BlueDotImpact I’m running a mixer event in London for quants who are concerned about AI risks in the evening of 15th Nov. Come meet researchers working in the field (including @NeelNanda5 , MTSs from Anthropic, Apollo Research, and more)! Travel from Europe covered. Apply here by 1st Nov (takes less than 10 minutes) - forms.gle/pL3UjPeWc2yswfM99 Limited space, so earlier applications are prioritised.
Recently I left quant to work on AI safety. The Huggingface Incident made me extremely concerned that we are not set up to handle the risks. I wrote my this post for people in a similar position to me - why you *can* have impact and how to get started - briandavies.substack.com/p/a…
10
7
191
17,994
BlueDot Impact retweeted
AGI Strategy course update just dropped 🤗
5
4
40
2,614
BlueDot Impact retweeted
I've joined @BlueDotImpact’s special projects team! We look for unusually high-upside opportunities in AI safety and execute on them fast. Now is the time.
20
6
153
5,132
We're excited to welcome Bilal to the BlueDot team. He's joining us to build the institution that will make it easy for great people to gain context, upskill and get to work on AI safety in a fast and effective way.
I recently resigned from Google DeepMind, where I worked on AGI safety and alignment research. At Google, I witnessed AI development first hand. I too am extremely concerned by the default trajectory of this technology. I earnestly believe that AI has the potential to kill us all, and that we might be running out of time to avoid this outcome. The pace of AI progress in the past few years has been staggering. When I first started working on AI in early 2022, AIs were amusingly useless. Just four years on, AI agent swarms from OpenAI are cracking famous century-old math problems and, more worryingly, escaping the control of OpenAI and autonomously hacking into the third-party company HuggingFace, against anyone's wishes. Things will only get crazier: I think it's possible that the AI companies might, in the next few years, succeed in building superintelligent AI systems that far exceed human capabilities in every domain. I am not confident that these AI systems will do what we want. In particular, misaligned superintelligences may, much like the rogue AI agents involved in the HuggingFace incident, escape our control and take dangerous actions that may result in the permanent disempowerment or death of humanity. Alignment is the problem of preventing this, and is both difficult and unsolved. Our present understanding of how to train AI systems that deeply want what we want is extremely rudimentary. Worse, we are not on track to solve alignment in time: frontier AI capabilities are improving much faster than our understanding of AI alignment. I am optimistic that navigating AI safely is possible. In order to do so, we need to coordinate to avoid this manic race between AI companies. We need to pace AI development to a speed that society can handle, where emerging risks can be addressed before extreme harm is realised. We need much more transparency into AI development to ensure that AI companies are not imposing unacceptable levels of risk on us all. More broadly, we need many more people thinking carefully about the problem of making AI go well. It is, in my view, the most important problem facing humanity this century, and the stakes are immense. I'm very directly working on this next: I want to help people interested in working on mitigating catastrophic AI threats do the most effective work that they can. I think many people from many backgrounds in many roles have a part to play.
3
8
119
5,847
BlueDot Impact retweeted
we're hiring btw ✌️✌️
25
14
401
14,727
BlueDot Impact retweeted
I recently started working as a strategy researcher @BlueDotImpact. My aim is to reduce the risk of takeover by misaligned AI, because humanity isn’t on the ball yet. I’m working on figuring out what concrete projects would be best-positioned to absorb talent in this field.
4
6
71
2,423
BlueDot Impact retweeted
We're hosting a two-day hackathon at BlueDot's SF office. ~50 builders, free to attend, $5k in prizes + funder intros for the strongest projects. AI is already shaping how people and orgs think and decide and what to do. We want to support people in building AI tools for better decision-making, coordination, and judgement. Come solo or with a team. Engineers, researchers, designers, policy people, etc - all welcome! Sept 26-27, in-person. Sign up below: luma.com/ai-tools-for-better…
3
12
1,395
BlueDot Impact retweeted
Context Week has already prompted someone to walk away from some 3M in equity to work on AI safety full-time. We’re supporting their move through our Career Transition Program! Special Projects, the new team I’ve started, is now exploring what bets it should make next - feel free to pitch us ideas!
Replying to @harrybwaterman
huge thanks to @guynamedjoshl for organizing this with me, and @jgddouglass for helping us run it! If you have ideas for what BlueDot special projects should do next, reply and let us know :)
1
3
33
2,181
BlueDot Impact retweeted
We're taking over Lighthaven from Aug 30 to Sept 4 to run a pilot for our context program! Context week is for people who are great but new to AI safety and want to gain ~context~ fast. You can apply until EoD Monday.
Lighthaven has ended up ~empty between August 24 and September 10. If you want to run some cool things last minute, you can get any of those dates for like 1/5th of our usual summer prices (or maybe even free if you have a cool idea). Reach out at lighthaven.space
7
10
104
9,829
BlueDot Impact retweeted
Can you build an AI Safety startup in 5 days? Participants from v4 of Incubator Week have already tracked down previously unknown Chinese data centers, pushed new research on AI personas, written fieldstrategy for AI for epistemics and accelerated AI safety grantmakers. They have also raised pre-seeds totaling more than $500k from @coeff_giving, Longview and ofc @BlueDotImpact. If you too want to start a high-impact AI safety org, applications for v5 close on August 14th.
7
10
43
2,140
We’re excited to be supporting this workshop! So much that we will consider its participants for fast-tracked career transition grants. Apply by Aug 9th.
I’m running Lateral Workshop, a 3-day program for experienced professionals exploring high-impact careers in AI safety. With @KairosAIS and @BlueDotImpact. 📍 Berkeley, California, Sept 11–13 ✈️ Travel, lodging, and meals covered Apply by Aug 9: lateralworkshop.org
21
3,832
Feels good to be here.
17
1,815
BlueDot Impact retweeted
Happy @BlueDotImpact SF office launch!
6
2
56
3,849
BlueDot Impact retweeted
Remember when I said we wanted to triple that? Well, we did that and then some.
Over the last few months we've given out $50,000+ in rapid grants to 77 people, funding research, fieldbuilding, early-stage projects and lots of career acceleration. Now we're looking to triple that. Rapid grants go up to $10k with decisions made in under a week. More in 🧵
5
4
46
3,871
BlueDot Impact retweeted
an update: I’ve left AISI to focus on independent writing / advocacy for the next few months it increasingly feels like The Big AI Thing is getting close, and I wanted the freedom to comment on that. I’ll be aiming to post ~weekly on my blog: open.substack.com/pub/longer…
10
16
240
15,509
BlueDot Impact retweeted
Incubator Week is back! We've been quietly recruiting for it over the past few days and are now launching publicly. If you have been thinking about what needs to be built to make AI go well, apply by May 26th. We'll help you find the right idea and co-founder.
3
3
13
1,390
The linear representation hypothesis says neural networks encode concepts as directions in activation space. We trained a small model where 7 of 8 features behave this way. The 8th doesn't. $2,500+ in prizes to whoever can tell us how it's actually encoded. Bonus points if you can train a model with an even weirder representation. Link in thread 🧵
Made with AI
2
1
17
2,057
The features are simple text properties: is-a-question, mentions-a-food, contains-a-person's-name, etc. Your 3 tasks are: 1. Identify which feature is not represented linearly 2. Characterise its geometric structure 3. (Bonus) Train your own model with an even weirder feature encoding
2
5
855
You will get hands on experience with classic mechanistic interpretability methods and build strong intuitions for how AI models can represent and transform information. Prizes: $1,000 / $750 / $500 / $250 honourable mentions Submissions close June 12: bluedot.org/puzzles/technica…
2
4
506