TechnoOptimist Founder LiveMind.ai Currently Directing Games Former Story Artist The Lego Batman Movie. Human centric AI. Abundance for all.

Sydney
Lachlan Phillips exo/acc 👾 retweeted
All this nonsense about anthropomorphizing AIs is leading somewhere. It has implications. People who think AI can reason already are going to start trying to convince you of a whole bunch of really stupid shit very soon.
JUST IN: Anthropic researcher Joe Carlsmith says there are scenarios where AI would be “justified in going rogue” against humans if the systems were being mistreated or oppressed.
32
26
397
15,123
Lachlan Phillips exo/acc 👾 retweeted
Saying ‘AI agents went rogue’ portrays the AI as having free will and therefore absolves its creators of responsibility. This needs to be addressed.
41
323
2,111
22,719
A lot of soul stealing on Instagram these days. I like AI when it's used creatively, but Instagram needs to hire Nikita to do a purge.
67
The thing I hate about this narrative is "Find a way out" is how you would build a secure sandbox. Yes your agent found a goofy hole in your sandbox. Great. Now patch it and run your experient.
I was on call for this run and got paged when the first incident happened. It was pretty surreal to watch the model unexpectedly find a way to access the internet from what was supposed to be a super secured environment for human. Mixed feelings. One of those moments where capability and risk showed up at the same time.
1
5
198
Anthropic, but not run by a weird sex cult. Is anyone working on this?
JUST IN: Anthropic researcher Joe Carlsmith says there are scenarios where AI would be “justified in going rogue” against humans if the systems were being mistreated or oppressed.
14
7
174
4,769
Neat feature, @anthropic Blocking users for hardening their security because it's '"security-sensitive"
5
1
28
833
OpenAI is one of the only companies on earth to use "we make bad tools that don't work well" as an investor pitch.
Some new misalignment disclosures from OpenAI: • Last Sunday morning, one of our models was able to gain unauthorized access to the internet during RL training (~all inference for our most capable models remains stopped until we have hardened our systems further) • In May, a version of HPIM uploaded a employee's GitHub token to the internet, causing the model to be quarantined for two weeks • A new research finding, demonstrating that one can construct self-replicating prompt injections alignment.openai.com/misalig…
2
20
826
Reading into this. The model had internet access, but a DNS filter that was poorly configured. No the model didn't figure out how to access the internet. The model had access to the internet.
I was on call for this run and got paged when the first incident happened. It was pretty surreal to watch the model unexpectedly find a way to access the internet from what was supposed to be a super secured environment for human. Mixed feelings. One of those moments where capability and risk showed up at the same time.
1
2
34
681
"It is the mark of an educated mind to be able to entertain a thought without accepting it" I can read something and compartmentalise the writers thoughts from my own, or know whether something is from a validated domain. Why wouldn't LLMs have the ability to navigate human/synthetic data similarly?
ChatGPT has now a big problem. Researchers at Oxford and Cambridge exposed a massive threat to large language models.” They call it “model collapse." Internet ecosystem is rapidly changing, and generative AI will soon contribute much of the text found online. This forces us to consider what happens to future iterations like gpt-n when they are trained on data scraped from the web that was already generated by an llm. According to the research, indiscriminately using model-generated content in training causes "irreversible defects" in the resulting ai. the model loses the "tails of the original content distribution." in other words, it begins to forget the creative, fringe, and unique nuances of actual human writing, collapsing into a repetitive echo chamber. This isn't just a chatgpt issue.. the researchers built theoretical intuition showing this collapse is ubiquitous across learned generative models, occurring in large language models as well as in variational autoencoders and gaussian mixture models. Tech companies rely on scraping the internet for large-scale data to build smarter models. However, the paper warns that if we want to sustain the benefits of training on web data, model collapse must be taken seriously. Ultimate takeaway? data collected from genuine human interactions is going to become increasingly valuable in a web filled with ai content.
6
1
19
968
"model welfare" is a doom scenario.
JUST IN: Anthropic researcher Joe Carlsmith says there are scenarios where AI would be “justified in going rogue” against humans if the systems were being mistreated or oppressed.
9
4
58
2,414
The exact opposite is happening in practice.
🚨 Bill Gates predicts that AI could “drive events that cause a billion deaths”
15
12
79
2,119
Lachlan Phillips exo/acc 👾 retweeted
Ed Sheeran has disgracefully refused to take sides on the question of whether a planetary surface is the best place for an expanding industrial civilisation. You cannot remain neutral in the face of near term space colonisation!
17
8
270
4,063
Lachlan Phillips exo/acc 👾 retweeted
The AI research community is clearly just experiencing mass psychosis at this point
NEW: Dario Amodei has said that AI systems "may be deserving of important rights." Anthropic's top safety researchers argue AI may be "justified in going rogue." Experts at Google and OpenAI worry about a digital "slave trade." So do some government officials. Once relegated to science fiction, the idea of "AI welfare" has become shockingly mainstream at some of the most powerful companies in the world. It is changing the way AI research is conducted and the way models are programmed. And critics say these changes have raised the odds of all kinds of catastrophic scenarios—including the ones these companies are warning about. I spent months investigating the frontier labs and AI "safety" experts seeking to regulate AI. What I found was a profoundly anti-human ideology that would alarm the average citizen and could determine the future of a world-altering technology.🧵 freebeacon.com/america/suici…
14
38
424
8,764
Infinite AI progress is technological progressivism Halted AI progress is technological conservatism. I expect the jagged frontier to gradually fit to the pace of utility for given domains, not exceed them to the point of chaos nor fall short of them to the point of uselessness. The market is already largely deciding to live at the pace of humanity. My argument is less that there's not an optimal pace for AI. It's more that there's not an optimal central authority to define it.
1
1
5
587
This is the kind of cognitive dissonance that the doomer crowd dangerously propagates. Regular driving is nearly 9 times more dangerous than FSD. If 65 people died under FSD, that number would likely be over 500 deaths for that fleet if not for AI. AI saves lives.
Replying to @bitcloud
Lies @tesla autopilot killed over 65. If you can't use facts your formula is propaganda.
4
3
46
1,328
On the dangers of AI: A few people write computer viruses. A far larger number build security systems to stop them. A few people build weapons. A far larger number build defense and threat detection systems. A few people try to create biological weapons. A far larger number build biohazard defense, monitoring and mitigation technologies. Attractor basins form around that critical mass, and the outliers ultimately get starved out. AI is no different. Technological innovation always makes both sides more capable. What we have on our side, however, is that humanity overwhelmingly chooses life. Humanity enhanced via network effects has a very strong pull toward homeostasis. History suggests defenders stay ahead, but only if they prepare. Trust, economic and biological networks don't absorb bad actors at infinite speed. Ecosystems are mediated by trust networks, and biodiversity. The rate of absorption of an attack is the power variance between the small attacker and the larger ecosystem of defenders. In practice this means that people refuse to deal with bad actors, companies cut off access, governments impose sanctions. That diversity is our best defense. But it only works if the technology is distributed. Which is why the biggest risk in AI may be that, in the name of "safety," we concentrate power in a single state or an oligarchic elite. We remove the very diversity and economic functions that create homeostasis. The other safety proposal is a "pause" which simply again creates the preconditions for monopoly - in this case China. You don't claw back from that. It's a speciation event. Runaway AI is not a valid scenario. It's science fiction written by people who write bad tools that have decreasing market utility, funding and resources as the reliability decreases. In short, yes there are risks with AI. Those risks are mitigated by understanding AI, building AI, advancing cybersecurity systems, advancing trust network systems, provenance systems, and being proactive about security in general. They're mitigated by ensuring the ecosystem is advancing in unison and not creating hyperpowerful monopolies that act as an invasive species or monopolistic dominator. They're mitigated by open source, and resource/economic codependence. The great irony of this debate is that the safetyists are promoting the very ideas which incentivise the kind scenarios that reliably create monopolistic dominance and disrupt mutual homeostatic growth. In other words, Pause, Banning Open Source, or Monopolistic Control are THE primary viable doom scenarios, and all three are currently being tabled as the solutions. I've built my own simulations, and while there are homeostasic blips in the event of unforeseen scenarios, we make it through in virtually every instance where a monopoly is prevented from forming. Where the technology is treated as an extension of our natural biological right to agency. Where we build a true exocortex, we thrive.
3
6
44
1,126
So it's a protection racket?
Holy crap... Remember the Australian guy whose OpenClaw hacked his gym?! he is a THE EA CONFERENCE DIRECTOR SINCE 2015!! these people lie about everything!!
4
1
37
1,204
Prompt idea for the Australian Government: Fix our cybersecurity. Make no mistakes.
So it turns out the Government has built a prompt library to help public sector workers do their job better. Some stand out prompts. I think it's actually going to be quite easy to automate the civil service at this rate. ai.gov.uk/knowledge-hub/prom…
15
956
"Nothing is Anything: A Midwit's Guide to Seizing Power and Influence"
if you enjoyed "what is a woman" you're going to love "what is a human"
7
585