@thinkymachines, formerly @workshoplabs cofounder. Blog: nosetgauge.com/ Tweets are all strictly in a personal capacity

who called it “timelines” instead of “T(foom)”
2
1
36
1,290
Rudolf Laine retweeted
Opposing AI takeover is not "speciesist" (if by that we mean some kind of bad unjustifiable tribalism). You exist as a HUMAN moral agent. Any ethical theory that requires you to ignore that bedrock identity severs you from the very thing that makes your agency recognizably yours.
Is opposing AI takeover speciesist? Which successors are good? What perspectives is the AI space missing that make it confused about these questions? New post out, that completes my series on alignment & succession. Thread -->
1
3
16
1,473
Rudolf Laine retweeted
I've added the reading list for our MATS stream (co-led with @LRudL_) to my website. It's a good starting place for AI strategy & governance work, especially if you're focused on distributing power/preventing power concentration. A 🧵 on some of my favorite recommendations:
6
32
279
20,023
Tinfoil is breaking the misuse prevention v privacy tradeoff in AI Misuse prevention should not be an excuse to violate user data privacy, that’s a skill issue solvable with technology
We’re releasing privacy-preserving, transparent and verifiable safeguards in Tinfoil Chat. The trending narrative says safety requires sacrificing individual privacy and freedom. We reject this and demonstrate that both can coexist. We also reject the absolutist argument that once any scanner exists in a privacy-preserving system, mass surveillance is the only outcome. We believe that preserving individual freedom and privacy long-term requires a responsible policy for responding to abusers who could ruin the promise for everyone.
1
6
34
1,573
Loss of control & concentration of power are actually very similar under the right frame. In both cases there is some actor, whether pure AI or human+AI, that can make its power uncontestable and wreck everyone else. Takeover by anyone is bad. (@RichardMCNgo pointed this out)
There are two ways AI progress could go very badly and that we must avoid. First, we could lose control of the future to AI. This is unacceptable; we are unapologetically on Team Humanity, and AI must always serve people. To ensure that, we need ways to ensure that alignment and safety techniques stay ahead of progress in model capabilities. Second, we could end up in a world with too much concentration of power. If an extraordinarily powerful AI is used by one person or company to impress their worldview onto everyone else, the results could be extremely dystopian. Avoiding these two threats requires walking a narrow middle path; for example, one country could gain too much power. Another example is one lab ending up with too much power.
6
10
86
4,037
This is the last post in a four-part series about the successionist challenge to humanist morality, how the problem of succession is central to the challenge AI transition, and what the arguments in favor of humanist morality really are. First part here: nosetgauge.com/p/alignment-a…
4
7
29
2,701
Separating the things that have power (currently humans, soon AIs) from the things that have moral worth that things in the world happen for (hopefully always including humans) is inherently dicey
1
10
214
Successionism could be viewed as specification-gaming of morality. To avoid it, it is helpful to remember that it is often powerful to simply throw out the spec
1
11
242
The entire AI space is obsessed with Benthamite utilitarianism, while missing the fact that his successor Mill was much more right about human morality. If you had to pick one preference ranking over philosophers to steer the AI transition, it should be Mill > Bentham
1
14
239
We should be a fan of chains of locally-valid succession to human descendants, even if that results in something we find weird and different (e.g. "transhumanism"), while being very opposed to succession across a big leap (how I define "posthumanism").
1
2
15
332
It is possible to oppose AI takeover on altruistic moral grounds, without being a bigoted speciesist.
4
20
5,535
Are there any successors we'd be happy with? Obviously yes: kids. But contra @robinhanson , AIs are not as good at being our kids as our kids are.
2
3
24
726
Is opposing AI takeover speciesist? Which successors are good? What perspectives is the AI space missing that make it confused about these questions? New post out, that completes my series on alignment & succession. Thread -->
1
11
58
3,416
In light of recent discourse:
Since about 2023 my rough sense of outcome buckets for humanity by likelihood has been: 10% foom&doom: at some point AIs climb very aggressively in capabilities and we fully lose control over them due to their actions & planning 40% bad selection: selection processes driving bad or very suboptimal outcomes due to important good things (like free societies, or humans in positions of power or economic actors) being outcompeted 10% stagnation: late in this century, as result of war or civilizational stasis & decline in both West & China, tech doesn't advance much 25% weird: deeply weird / hard to judge / mixed, would take lots of time to form an opinion on 15% obviously good: flourishing future by our current lights These ratios have been pretty stable, and I’ve found it hard to get much evidence to greatly lower or reduce any bucket. Compared to many I have larger error bars and feel more averse to “worldview distribution collapse”. Evidence from the OpenAI HuggingFace hack and considerations of RSI mechanics have made me update. I expect to settle around 14% foom&doom, with most of the increased mass there coming from the more extreme side of the bad selection bucket. You can have gripes with assigning numbers to things but I would encourage people to make their views more comparable and actionable and quantified, to help create common knowledge as we head into fast-changing turbulent times
4
486
Rudolf Laine retweeted
There’s a window before RSI where lab employees, particularly the top ~250 people, can probably meaningfully affect geopolitics and the future in general. The window closes when they’ve automated themselves such that the state can act unilaterally, which they’re all racing to do.
10
23
415
10,650
Multi-year well-staffed giga-brained institutionalized efforts by trillion-dollar labs to raise awareness of AGI risk (which they see as core to their theory of positive impact) were likely just beaten in one day by one guy with one act of courage
I resigned from Anthropic today. I spent the last three years doing pretraining research at both OpenAI and Anthropic. Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives. More thoughts below.
5
3
95
3,651
Rudolf Laine retweeted
AI is progressing much faster than I expected. At the start of last year @LRudL_ predicted AI would solve a Millennium Prize problem by the end of 2026. At the time I thought that was preposterous
4
1
34
1,251
many commenting on millennium prize drama but few thinking about the economic ramifications of unlocking the prize pool as an ARR stream
3
35
1,690
Rough heuristic: within a few years of AI looking impressive to amateurs in a field, it’s transformative for experts. Eg GPT-4 at math or code to various new proofs & ~all code in 3y ML R&D and CAD now likely similar (unless even faster from R&D feedback loop closing)
3
28
1,805