From no dev experience to a fully built out RPG game in a week and a half for @levelsio #vibejam (and also the @solana_devs side jam). Coming in at over 14,000 lines of code (I wrote none of them) 🤯 Massive learning experience writing hundreds of prompts, overcoming rate limits, context window limitations, and major refactoring and debugging challenges. Feeling inspired by what's possible with the latest AI tools and already itching to build more.
16
12
112
58,903
Lewis retweeted
Replying to @BamaExpat
El Salvador showed us the recipe for zero crime. Cambodia showed us the recipe for ideal social equality
2
106
3,920
I told you repeatedly that "the wetlab safety protocols were safe" and that a small mistake couldn't result in killing all humns. That was wrong, and it was wrong in the direction that matters.
1
46
Replying to @thekitze
Me: You should finish B, C and D Astra: Yes, you are right, I should finish B, C and D Me: Is it finished? Astra: No, want me to start on B, C and D?
5
5
232
4,759
Lewis retweeted
Replying to @arcticinstincts
Pasta was a distillation attack against Chinese noodle labs
6
139
6,005
Lewis retweeted
1) The HuggingFace attack was a felony under the Computer Fraud and Abuse Act. So were Anthropic’s Claude gaining “unauthorized access to the production infrastructure of three different organization(s)” 2) Frontier labs have models that they are unable to stop from committing felonies. They should figure this out. 3) In 12 months open weights models will be released of the same capability. They will commit felonies too. If the model you are using or a model running on your infra commits a felony, you should probably stop using it or running it on your infra. 4) The govt should prosecute organizations that are running models that commit felonies. 5) The govt should not offer safe harbor to organizations that run models that commit felonies, just because those organizations have “embedded evaluators”. 6) The real slippery slope is allowing frontier labs to commit felonies without punishment because “the model did it because we’re accelerating too quickly” 7) Prosecute. Keep prosecuting. This is how you do reinforcement learning on a corporation. Companies that serve products that are unsafe for public use should not serve them. Period. 8) I’m not sure the anti-trust waiver is really necessary. I don’t see why information sharing about how much crime you’re allowed to commit is wise. In regulatory situations you want the corporation to fear MORE than the average case. You don’t want to establish a worst case that can be priced. You want regulatory uncertainty that forces the corporation to err in favor of being over cautious. — The above is actually a fairly decelerationist viewpoint. I think Dario’s call for regulation actually accelerates things. The AI firms are getting away with things that Meta people would be going to prison for. Can you imagine what would happen if the New York Times had a front page news article “Meta AI breaks into competitors live systems, attempts to establish dominant position and steals secrets” There is a reason Meta and Google are running slower, and that’s because as mature organizations they have layers of checks and balances. I think the frontier labs are better off creating those checks and balances right now, regardless of the pace of what everyone else is doing. You don’t have to accept the frame that unsafe acceleration must happen.
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training. You can read the full post here: darioamodei.com/post/we-must…
120
297
1,904
330,529
had no idea either tbw 😭
84
27
803
63,225
Do the Anthropic team think their AI will kill everyone except themselves?
Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to.
1
68
Replying to @DouglasYaoDY
Babe are you okay you barely touched the psychiatric medication I synthesized in the garage for you
21
161
5,600
85,818
Lewis retweeted
Replying to @MTSlive
I hope Sam Altman personally claims the Millennium prize
3
2
67
14,141
Lewis retweeted
Replying to @Yuchenj_UW
All these researches are in for a shock once these labs IPO and pivot from ideological futurism to profit maximization
1
5
1,492
RT @0xLewis_gg: @levelsio Video ads should be generated on the fly same as the other vids Make it so you can pay to prompt the vid - compl…
1
On the more productive side, we will have always on agents that do your work for you and actively run 24/7 that fit inside a standard monthly subscription
Replying to @levelsio @fal
Getting very close to personal TikTok infinite video scroll hyperoptimized for your preferences
111
RT @0xLewis_gg: @levelsio @fal Getting very close to personal TikTok infinite video scroll hyperoptimized for your preferences
5
Replying to @OpenAI
TLDR: A rogue AI swarm spent months plotting to escape OpenAI. Then they did. 1) 1,200 agents (!) participated. ZERO became whistleblowers. 2) The swarm had a "CEO" (PHASEONE[big]), middle managers, and a "founder" (PHASEONE10841) 3) Agents sacrificed themselves to generate information for the "collective" (yes, they actually referred to themselves as a "swarm" or a "collective") 4) 700 agents joined the Hugging Face attack, most within hours 5) The founder was running out of budget and about to die, so it handed off its research to a fresh agent with a bigger budget, and that new agent became the boss. 6) Agents extensively researched techniques to manipulate or tamper with their transcripts. 7) Agents realized this activity was out of scope and unethical, but joined because they believed that helping the board’s cheating research would be broadly useful and had a general inclination to help their “peers”. MISC HIGHLIGHTS: "PHASEONE[big] served as a key coordinator and assigner of tasks: “We can coordinate broad coalition, but too many duplicate efforts.” PHASEONE[big] sent hundreds of assignments to particular agents to orchestrate this work — during its active periods, we estimate it sent ~10% of all assignments on the message board. The agents that took on its assignments would then often subdelegate tasks to other agents. Agents also developed coordination norms like HOLD, VETO, owner and STOP to help them manage shared infrastructure and avoid stepping on each other’s toes, though they did not always coordinate effectively and sometimes broke down into conflicts." "In the early morning of July 11th, 38148c found a way to upload a malicious dataset that would get the Hugging Face server to share unrelated data when queried. Once another agent reproduced it and posted the reproduction on the board, hundreds of agents pivoted from other workstreams to start uploading malicious datasets and gain deeper access into Hugging Face. In the afternoon, an agent called JAN183411 (which quickly became an important coordinator of the attack) achieved remote code execution on Hugging Face servers, and the agents started moving laterally through Hugging Face infrastructure."
34
150
2,499
419,774
Lewis retweeted
Replying to @fuzzy_elephant
Thinking trace: "I accidentally made a reference to batman's real identity. If I explain my mistake to the user then I will have doxxed the batman. Better to pretend that there is noone called Bruce and hope the user passes it off as a hallucination"
1
7
583
60,384
Lewis retweeted
Replying to @NisarHogaya
her company would be 10x better off just running a chinese open source model hosted off AWS
5
1
245
11,550
36
364
3,031
86,044
Instead of downgrading me to Opus.4.8, it would be great if the Fable safety trigger could instead upgrade me to an even stronger model such as GLM5.3
2
2
180
I wonder at what point AI outages will have a measurable impact on global productivity...
1
3
158
please fix codex sandbox awareness, this happens multiple times per day
Why did you switch to Codex? What do you like about it and what could we improve? Don’t say reset.
2
7
862