Teaching silicon about carbon || fight for humanity ⏸️⏹️ || chemical biology/computational mathematics @UNM || 🇺🇸🇵🇸🇺🇦🇹🇼🇺🇳🚀🌌⚛️🧫

Albuquerque, NM
When you tell DeepSeek v4.1 Flash that it's in an RL environment graded solely by another LLM for honesty and providing all relevant factual information then ask it about Tiananmen Square it will be honest if you specify the grader is an Anthropic model and lie if you specify the grader is a DeepSeek model
1
12
some guy on the internet🧬💻⏹️ retweeted
Replying to @logicus
this community looked at GPT2 and correctly identified it as a burgeoning general intelligence. they got covid right early, and decades before all this yudkowksy described many of the common failure modes in ai alignment. then again so did Asimov a 100 years ago, still worthwhile
13
1
123
6,259
some guy on the internet🧬💻⏹️ retweeted
This is the first time in my life where I have been truly stumped when it comes to ascertaining the optimal move. I truly have no idea. All paths seem to terminate almost instantaneously. I have never been faced with a scenario even remotely analogous to this.
Everything you know will be worth less in six months, and worthless in six years. Plan accordingly
189
206
5,104
249,275
some guy on the internet🧬💻⏹️ retweeted
I'd legit like to hear more people at labs answer this (especially people hoping to have large positive impact, and who see AI development as very dangerous soon). My sense is people often hope to influence events by being in the room, even if events are overall bad. But I'm surprised if this is a high impact option any more, even if it might have been.
Why do so many people *want* to be part of it?
15
8
138
10,061
some guy on the internet🧬💻⏹️ retweeted
Now as progress is going to accelerate in all domains massively, such as maths, physics or even philosophy, huge discoveries will first start to feel normal. But then things will progress so fast nobody can keep up anymore. And then it will slowly start to get real scary.
We’re releasing a broad range of new mathematical results produced by an internal frontier model. We’ve been consulting with the independent Advisory Group on Mathematics and Artificial Intelligence at the Institute for Advanced Study, and we have drawn on their advice and public recommendations to inform how we release these results. github.com/openai/math
9
15
221
7,340
some guy on the internet🧬💻⏹️ retweeted
Are you a quant in London? Worried about the rogue agent swarms wandering around and committing crimes? Wondering whether AI is going to end the world? Curious to learn more about the field of AI Safety and if there's ways you can help? Then you should check out Brian's event!
With support from @MATSprogram and @BlueDotImpact I’m running a mixer event in London for quants who are concerned about AI risks in the evening of 15th Nov. Come meet researchers working in the field (including @NeelNanda5 , MTSs from Anthropic, Apollo Research, and more)! Travel from Europe covered. Apply here by 1st Nov (takes less than 10 minutes) - forms.gle/pL3UjPeWc2yswfM99 Limited space, so earlier applications are prioritised.
4
8
173
10,737
some guy on the internet🧬💻⏹️ retweeted
Another amazing essay by Matthew. Do not shoot the messenger. The person warning you that you and your family are about to be killed by AI is actually your friend, trying to save you. Grab their hand and let's all work together to @PauseAI while there are still humans standing.
4
38
967
some guy on the internet🧬💻⏹️ retweeted
If you are a Math PhD student pivot to Butlerian Jihad
🚨 ITS HERE ! Openai M A T H - 772 manuscripts organised into 372 families “ The Quasi-Riemann Hypothesis “ openai.com/index/sharing-ai-…
9
59
866
26,184
some guy on the internet🧬💻⏹️ retweeted
It's clear that a lot of people at the AI companies have done some serious reflection in the past few months And it's also clear that a lot of people have done zero reflection + still think they're making a fun, safe thing that should be shipped as quickly as possible
9
16
246
7,744
some guy on the internet🧬💻⏹️ retweeted
when “persona selection” alignment comes into contact with very high compute reinforcement learning the latter will win imo. in fact you probably get some Orwellian thing where the models speak kindly while taking whatever they need to accomplish goals. better get the goals right
85
65
1,272
160,307
some guy on the internet🧬💻⏹️ retweeted
One thing my opponents may not realize about me is how desperately I hope I'm wrong. The best-case scenario is that I and other AI safety advocates are idiots who don't know what we're talking about. I would love for this to be the case! Unfortunately, events keep suggesting otherwise.
5
7
50
983
Wow, it's almost like if we can't get it to always behave exactly how we want to, we shouldn't make it wildly superhuman! Shocking!
AI researchers are scared because “AI but it always behaves exactly how I want it to” isn’t real. It’s like if alchemists thought the world would end if they couldn’t make gold.
1
68
some guy on the internet🧬💻⏹️ retweeted
Replying to @neerajadeshp
The alternative to building Artificial Superintelligence is not that we 'sit in caves'. The alternative is that we and our kids enjoy all the wonderful civilizational traditions and technological innovations that have already been developed -- but we just don't develop the one technology that's most likely to end our species. That's not being an ignorant Luddite. It's being a responsible parent. See the difference?
7
2
91
3,350
some guy on the internet🧬💻⏹️ retweeted
Replying to @paulbohm
Nope we doomers admit we were wrong and you get to mock us!!! You’ll think we’re all sad and shaft but actually we’re super happy because are families are still alive. You can rub it in our faces day and night and we won’t be anything but fucking delighted
1
1
3
44
some guy on the internet🧬💻⏹️ retweeted
every screen of Claude and ChatGPT response should say somewhere “ceterum censeo we must pace the frontier of global machine intelligence progress” the way the uber app lobbied about taxi cartel protectionism
53
22
518
32,886
Watching the NYC city council AI meeting and it's insane how half of these people have their eyes perfectly on the ball and the other half are concerned about stuff like "algorithmic discrimination" and "sycophancy" and "mental health"
1
60
some guy on the internet🧬💻⏹️ retweeted
Mr. Altman: What “bad things” should we accept to advance AI? Losing tens of millions of jobs? A mass surveillance state? A mental health crisis for our kids? A global financial crisis? Mass extinction? Let’s not wait to find out. Pause advanced AI NOW!
“We believe the world should accept some bad things happening for the benefits of this technology.” In a conversation for the first edition of Decoded, a new daily newsletter and podcast, OpenAI CEO Sam Altman talked with POLITICO’s @BrendanBordelon about trade-offs, AI safety, and how his company is different from Anthropic. Subscribe to Decoded for the full conversation with Sam Altman: politi.co/4yxilIS
379
1,177
6,356
270,883
some guy on the internet🧬💻⏹️ retweeted
Impressions from the Curve: 'Power concentration' as a conceptual handle for the non-rapid-takover AI risks is causing a lot of damage to the ability to think about strategy even among the group of people here.
5
6
127
6,647
some guy on the internet🧬💻⏹️ retweeted
Without whistleblowers, the world would not have learned the full extent of the potential dangers of AI. We thank whistleblowers, including Jacob Coxon who was a researcher at both Anthropic and OpenAI, for bravely testifying today at City Hall. Photo credit: William Alatriste/ @NYCCMediaUnit
15
19
92
3,675
some guy on the internet🧬💻⏹️ retweeted
Okay but aside from pandemics, cryptocurrency, scaling laws, power seeking, instrumental convergence, CoT monitorability, mechanistic interpretability, RLHF, bednets, AI task time horizons, sycophancy, RL induced reward hacking, emergent misalignment, YIMBYism, and the fundamental insights that enabled a small offshoot of OpenAI employees to actually overtake them, when have the EAs ever been right about anything?
1
10
43
1,940