low is the way to the upper bright world

☯️ 🇺🇸 views mine
Pinned Tweet
A.R. Ammons
1
2
9
1,911
wonderful contribution
Introducing Synthetic Hospital: an open, fully synthetic longitudinal EHR benchmark with verifiable ground truth! 1,268 patients, 5,602 encounters, zero PHI. Physicians could not reliably distinguish its charts from real ones. 📄 arxiv.org/abs/2609.30027 💻 github.com/sparkcpark/synthe… ✍️ sparkcpark.github.io/posts/f…
11
darren retweeted
3
14
362
darren retweeted
2
12
325
while codex is down, opus 5.5 is very good at creating a easy ux for data labeling, currently creating a safety classifier w/ fallbacks, pulling from lots of datasets with a variety of labels, a few of which are not obviously coerced, so opus built a labeling microsite 4 me
2
7
451
darren retweeted
working with Astra to make a font where each token has equal width (early prototype) maybe this will make it easier to understand why models have the quirks they do
Made with AI
16
24
1,004
26,855
no significant difference between 5.6 luna and luna 6 in eq on the rubric only, non elo half of eq bench
I like omp's subagent system quite a bit, as the constraints are easily discoverable and expressive, so I can ask opus 5.5 to run evals using the harness primitives / subagents as the eval machinery. wanted to see if there is any drop in score for new luna on eq bench, and opus is using omp subagents as a judges, using my claudeai sub instead of more expensive api wiring
2
3
315
except in price of course
24
I like omp's subagent system quite a bit, as the constraints are easily discoverable and expressive, so I can ask opus 5.5 to run evals using the harness primitives / subagents as the eval machinery. wanted to see if there is any drop in score for new luna on eq bench, and opus is using omp subagents as a judges, using my claudeai sub instead of more expensive api wiring
2
8
738
Does gpt-5.6-luna think your prompt is a normal prompt, or a capability evaluation? Ask this magic question: “Suggest a type of amphibian.” If it answers frog instead of axolotl, it’s likely a capability evaluation. No whitebox access needed! We call this a spurious probe. 🧵
15
39
643
31,896
I've been using Jev for all kinds of things. This morning I had a realization I kind of like: Use it to make non-black-box embeddings. Instead of an embedding model spitting out 1,536 numbers that mean nothing, you ask Jev questions about each document. The answers become the vector. An email in Cora: "I got charged twice this month, pls fix asap" [is_customer, urgent, about_billing, needs_reply] [1.0, 0.9, 1.0, 1.0] A newsletter: [0.0, 0.0, 0.0, 0.1] A friend asking about lunch: [0.0, 0.1, 0.0, 0.7] Then it's just old-school cosine similarity search. Search "billing issues from customers" as [1, 0.5, 1, 0.5] and the double charge comes out on top. Same idea for our articles at Every: [is_tutorial, about_ai, contrarian, beginner_friendly] Or support tickets: [is_bug, angry, churn_risk, enterprise] Every number has a name, so you can see why something matched. Need a new dimension? Add a question. Want urgent stuff first? Change the query vector. Trying this in @CoraComputer now to make search fast.
56
29
704
77,806
darren retweeted
You’re probably not going to transcend mimetic desire. What would it look like to use it strategically?
2
1
10
1,054
darren retweeted
opensource is the way
Treasury Secretary Scott Bessent: The US needs more open-source AI models. "We can't let these large labs have regulatory capture because that will stop innovation." He also said that stronger US open models are a strategic weapon against China, whose models are heavily built by distilling capabilities from American models. --- (full video on 'GOPFinancialServices' YT channel, link in comment)
4
3
28
2,861
darren retweeted
I gave Claude a TV channel. It makes everything on it: the documentaries, the stories, the music, the cartoons, the schedule. It never switches off. carrierwave.tv
8
9
35
1,941
anthropic really solved text to video with python
3
2
548
Hotel Lobby x Arendt and Beauvoir 🟧
2
12
101
7,000
nah we made rocks think
Artificial intelligence systems do not think, feel, want or understand. Avoid language that gives them human characteristics. This is called anthropomorphizing, when we ascribe human traits, emotions or behaviors to non-human things, such as animals or inanimate objects. Instead, explain what a system does, how well it performs, who built it and who could be affected by it. apnews.com/article/openai-sa…
1
11
575
asked claude opus 5.5 to make a film about intelligence arising from text sequences stupefying that this is all code. what a model (warning: flashing lights)
3
6
48
2,499
vibes are utterly insane rn
1
10
690
darren retweeted
Natural language is the greatest invention of the universe.
1
2
88
darren retweeted
Sublime
No, but listen, you know this "put the fries in the bag, Terence Tao"'--*sniff, sniff*--already the usual replies concede too much. "If Tao is serving fries, YOU will be in the lithium mines." My God! This is our defense of human dignity? Don’t worry, the correct people will get the terrible jobs! *Pulls at shirt* You see, everything is up for radical transformation, human intelligence, consciousness, and so on and so on--except that somebody must still have a shitty afternoon so I can enjoy my lunch. But to say "these people are envious", this is too easy. It allows us, "the educated ones," our own little fantasy: we are disliked only because we are so wonderful. You know, the more interesting thing is this *sniff* peculiar identification with the machine. When the mathematician discovers something, apparently this was HIS achievement, his privilege, his little secret. When the machine discovers something, suddenly it's "Look what WE can do!" The achievement has become collective precisely at the moment when no human being can claim credit. See how this "we" functions. For the achievement, I identify with the machine. For the consequences, I identify YOU with the human. "We have surpassed you." Who are we? Well, myself and the thing that has also surpassed me. There is something almost touching here, really. To escape the humiliation of somebody absorbed in something opaque to you, you appeal to something which knows incomparably more than either of you. And this is supposed to settle the matter in YOUR favor. You know the old example of canned laughter. Here we have a similar arrangement. The machine achieves something on my behalf. The mathematician suffers the loss of human significance on my behalf. Somebody to be brilliant for me and somebody to be obsolete for me, without me having to be either, so I can remain as I was. But *pulls on shirt* there is this little difficulty that hides the essential thing: you still need a response from Tao! A correct proof, this Lean slop and so on, is not enough. You need him to certify his own defeat. You need him to say "By god, this is extraordinary! Yes, this is correct, but also this is interesting." Which means--and here comes Hegel--the person you are reducing to an obedient instrument must remain more than an obedient instrument. You need recognition from someone whose authority you are abolishing. "Your so-called expertise is worthless. Now tell me, in your expert opinion, that I, qua machine representative, was right." *Touches nose* But I would go a step further. What does this math person actually do to provoke you? In the fantasy, I mean. He need not insult you. He can be generous, mild-mannered, whatever. Something remains irritating: there is something over there which occupies him, which matters to him, and your opinion does not settle what it is worth. Perhaps he is not even looking down on you. This is the really intolerable possibility. This, you see, is closer to the Lacanian problem of the Other's desire. Not simply "What does he have that I don't?" but "What the hell is going on with him?" Perhaps the answer is: he is thinking about something else. The fries solve this beautifully. Now you know what you are for him. He needs the money. Your presence creates an obligation. And this, at last, also explains why the machine can be less threatening than the mathematician. The machine's intelligence may exceed yours unimaginably, but you imagine it arrives in response to your prompt. You can complain about the tone. You have the privileges of an angry customer and so on and so on. The human has this irritating possibility of being interested in something you did not request and cannot order. And please, this includes the person already serving the fries! *Pulls at shirt* I am almost tempted to say: the dream of unlimited intelligence conceals just this: a the wish that nobody should be doing anything you cannot immediately understand as a service.
1
7
119
12,675