Let's build safe AI! Law prof @ Alabama Contracts, Defamation, Legal NLP, & AI Safety

Tuscaloosa, AL
Why do workers have to wait for 2-4 weeks to be paid, in the same economy where online transactions go quickly and securely? A new draft 𝑃𝑎𝑦𝑑𝑎𝑦-forthcoming @WashULaw-proposes that they shouldn't. Daily, or at least weekly, pay can be a reality. papers.ssrn.com/sol3/papers.…
4
10
61
Goddarnit another benchmark is saturated
8
469
Please don't call me between 2 pm and 4 pm on Sunday. I'm going to be reading Scott's response and will need time to recover. Thank you for your consideration.
4
571
here with Justin and he's being nice. One sentiment expressed is that Justice requires models that are near perfect (open source, no bias, open training data, must never hallucinate, whatnot), with complete blindness to actual costs of slow justice, judicial bias, and basically any issue of judicial economy
At a conference with judges from around the world, and I’m shocked that (1) there’s a good amount of agreement about what to do on AI and (2) the agreement includes ideas that seem very sensible as well as ones that seem inaccurate or counterproductive 1. Judges must be responsible for and able to explain their decisions 2. AI’s identification of legal precedent is biased and must be carefully guarded against via audits 3. AI is good for summarizing and translating but can’t do good legal reasoning or judgment 4. Courts should protect their data by keeping it in a private cloud and using it to fine-tune custom AI models 5. Courts must define safeguards around judicial use of AI
2
1
12
1,329
By my scientific count, 13.4% of seminar time is spent on the moderator saying: "OK, so we only have 12 minutes and we have 14 questions on the queue, so in the interest of time and to make sure everyone has a chance to ask questions, we are going to ask speakers to keep the answers short and also the questions, because we need to make sure we have enough time, because we have another panel right after and <3.5 more tokens>"
6
339
Yonathan Arbel retweeted
I work at Google DeepMind. This won't make me popular. But it's all public reporting: 2014: DeepMind reportedly sold to Google on conditions: no military use, independent oversight 2026: a Pentagon contract for "any lawful government purpose" Not one safeguard survived intact
136
1,245
4,487
489,703
Lolz. Also: if you find yourself having to use AI, read only the first few sentences. Always use organically trained models. And for the love of all that's holy, ask it to use American English to save on excess tokens in words like color
I cannot believe this is real...
1
433
Signal boosting
We are hiring! Many different fields, including philosophers whose work is relevant to government and policy! Note that this ≠ only political philosophers! These days many areas of philosophy are urgently relevant. Come work with me and my colleagues @JohnsHopkinsSGP on rebuilding crumbling liberal democratic institutions philjobs.org/job/show/32233
2
416
Say what you will of my ideas (and some of you have), but my slide game is unrivaled
Replying to @ProfArbel
@ProfArbel asks whether under the Fourth Amendment AI agents are more like documents, tools or confidants. Excellent conversation at the @UNLCollegeofLaw symposium.
3
22
654
Thanks @IThinkIAgree for organizing and for envelope-pushing lectures by @KevinTFrazier , @mtokson and @Dschwarcz
4
72
Yonathan Arbel retweeted
The idea that the AI race isn't real is starting to catch on
Treasury Secretary Scott Bessent recently warned that “there is no day after tomorrow” if China pulls ahead in artificial intelligence; Americans receive this message from political leaders and from tech innovators. Yet, the first to cross the finish line is not always the one who wins a tech race. Take, for example, the long line of British inventors who helped bring electricity to the world in the 1800s. A betting man, at the cusp of the 20th century, would have been hard pressed to deny Britain its place in the future. But that’s not how the story goes. By 1912, America had become the world’s leading industrial power, and was producing five times more electricity, per person, than Britain. This logic—that it’s not always who is first who ends up cornering the market—may apply to A.I. today. Chang Che explores this way of thinking: newyorkermag.visitlink.me/NM…
2
2
432
Yonathan Arbel retweeted
Replying to @DavidPinsof
To prove an extraordinary claim you need extraordinary evidence. But an extraordinary claim doesn't depend on extraordinary evidence to be true. Also, we can play burden-of-proof tennis on what is the exact extraordinary claim. I'd argue that the proposition (building a socially transformative superintelligence will be safe on our first try) is quite the extraordinary claim
2
1
1
153
Yes. The legal AI safety community, by my read, converges on the need for ex post liability , ex ante supervision testing audits, and openess to halt / pause / pacing. A couple of years ago we called for a package of systemic regulation of AI.
3 key points about AI liability. 1. Liability is an indispensable tool for mitigating AI risk & should be central to AI governance. 2. The existing AI liability regime is deeply inadequate. 3. Even with the much more robust liability regime I favor, liability is insufficient.
1
2
10
686
Yonathan Arbel retweeted
3 key points about AI liability. 1. Liability is an indispensable tool for mitigating AI risk & should be central to AI governance. 2. The existing AI liability regime is deeply inadequate. 3. Even with the much more robust liability regime I favor, liability is insufficient.
2
19
72
7,642
My friends know it will take a lot for me to retweet Senator Sanders, but I support his efforts and will work together with anyone across the aisle to further the efforts to put a stop on building superintrlligence before we have sufficient assurances that it is safe
Jensen Huang, the CEO of leading AI company Nvidia, just said: If an AI model will "damage the world — then I think the answer is that we have to shut the labs down." My new bill to ban Artificial Superintelligence would do just that. If you can't control it, don't build it.
15
2
101
2,833
All right, but apart from a thousand agents coordinating on an unsanctioned board, a health-statistics portal, a pilates queue, a German wiki, three production networks, and the Australian government, what evidence do you have of misalignment?
This turned out to be the most prophetic tweet of the year
3
554
Discussing one of my favorite K cases today. (although every year I become increasingly more offended when the students call her old)
3
4
758
<in deep profound voice> maybe it's not a simulation theory, maybe we are just eval aware
4
415
This is going to sound like a crazy conspiracy theory, but what if Jensen only says he is not worried to boost his company price? h/t Justin
The guy selling all the compute, not worried at all about that compute. Solid work folks.
2
2
31
1,017
I want to clarify a point that seems to confuse a lot of smart people. Expert cyber people look at the HF incident and say -- uh? this is a stupid security incident. OAI messed up its sandbox, gave crappy instructions, and failed to maintain basic monitoring. You don't need to worry about AI superintelligence when the problem is basic human stupidity. It's a bad good point. It's good because they're right in their assessment and we have here holes that could be plugged with sufficient care and attention. But it's a bad point overall. When we design AI safety policy--and this includes the pacing/pause proposal--we do it to prevent catastrophic risk from happening in our world. In our fallen world, security is really bad (or rather, it is jagged). Even responsible looking institutions sometimes fail to defend against well known risks. Why? As every cyber person will tell you with exasperated look, it's a combination of human stupidity, executive arrogance, budget constraints, and Jerry who forgot to fire up the server before he left early for a doctor's appointment. Attackers who can coordinate a large scale attack on enough vulnerable systems at the same time can cause catastrophic harm (not necessarily extinction, but enough to be crippling). Such attacks are rare in practice for a number of reasons that have to do with who stands to gain from them and how. But if coordination frictions are low, we need to think about these risks seriously. We can't just extrapolate from things are fine to things will always be fine. AI could be really helpful here. Indeed, OAI just sicced Astra on its systems to plug up all the holes it could find. They are not sure there are no more holes, but they sound confident that the main ones were identified. But why didn't they do it beforehand? Why wait for a major incident? Because, as we noted, even sophisticated, well funded organizations fail to take sufficient precautions. This is what I mean by my basic Law of AI Safety: If your solution to a safety issue is "we can simply do" then you don't understand the issue and there is no need to hear the rest of the sentence. There is no simply here. We can't simply ensure that all of our infrastructure is safe, because doing so is the opposite of simple. So we don't even need to consider 0-days and thermal-based comms to understand how unaligned agents can wreak havoc (at the behest of an attacker or through badly crafted prompts or through more autonomous action). We need a systemic approach to AI safety. We need auditors and we need standards and we need liability and we need technical alignment. And because we need time to actually do all of those, a slowdown will be part of the recipe. Yes, as @AuerDirk notes, there's no guarantee we won't squander the time (we certainly wasted the time we had since GPT-2 came to the scene, despite many who suggested early action). But even if it's not sufficient, a slowdown is necessary.
4
9
858
Wtf Claude you just started a war Good point and you are right to push back!
SITUATION DETECTED: A U.S. SOCOM analyst used AI to produce an intelligence report that hallucinated that a Chinese ship in the Middle East was carrying nuclear-weapons components. The U.S. military was preparing to board the ship before the error was caught, per CNN.
13
4,696