Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 30th August 2026 - awful.systems
Reply in thread
I would prefer they turn them off
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 30th August 2026 - awful.systems
Reply in thread
I would prefer they turn them off
Comment on
Random Positivity Threads
Reply in thread
For more than a decade now Egan has been enthusiastically taking the piss at the rationalists and singulatarians. For more than 20, he has been gently making vague fun of them, or calmly including stuff in his stories that flatly contradict them without commentary in a way you can tell he's influenced to do so by them.
Comment on
There Will be no Redemption for Ezra Klein
Reply in thread
Didn't one of the Arbital head honchos have a psychotic episode after writing unhinged rants about rationality dojos making superhumans?
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 30th August 2026 - awful.systems
Reply in thread
tim taylor voice Oh no...
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 30th August 2026 - awful.systems
Reply in thread
I think the only cult type that Yud hasn't spun off yet is fundamentalist Christian.
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 30th August 2026 - awful.systems
Reply in thread
I mean, true, but that's true of most secular cults in America and Europe.
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 23rd August 2026
Large amounts of the LW commentariat cannot get their heads around how students JUST. DON'T. GET. AI.
https://www.lesswrong.com/posts/ySXuvJcqRindQwAk7/how-my-students-think-about-ai
They are extremely surprised by:
Perspective #1: There has not been rapid AI progress. My students do not have any intuitive sense that there has been rapid AI progress in recent years or really have much of a framework for thinking about that issue... With the exception of image/video generation, GPT-4 could do most of what they were looking for from a chatbot. Three years is a long time in their world, and their sense is that chatbots have been a mature technology over roughly that amount of time. They are an imperfect technology — students are well-aware of hallucinations — have been one, and will continue to be one... Historical context I gave them made this worse. The basic sentiment here was “AI was superhuman at chess a decade before we were born, and this is all they’ve done with it since?”...
Perspective #2: Impressive progress or not, AI is going to wreck their lives, the economy, and the social contract.... Corporate leaders are always looking to get rid of workers, even if it is irrational to do so. Various motivations were posited here (hatred of the working class, FOMO, a preference for technology, machines can’t go on strike, etc.) but many of them think that a CEO would ultimately choose to pay twice as much to get AI to do a task half as well... AI will do substandard work that makes products and experiences worse but is capable of just barely scraping past the bar of minimal functionality (many of them independently brought up constantly malfunctioning self check out machines). This will be rapidly rolled out, enshittifying most things.
Perspective #4: Catastrophic/existential risk arguments are sci-fi distractors from the urgent social/economic/political problems associated with AI. My students have a fairly strongly held view that “rogue” AI does not represent a real threat... They also mostly think that, if one does believe that AI is existentially risky, then the strategic interaction is not prisoner’s dilemma or even chicken but rather just a game theoretically boring setup where you die if you defect. Mash these together, and you end up with the view that expressed concerns about existential risk in the AI industry can’t be sincere... Students (both independently in written work and then later in group discussion) hypothesized that this might be a deliberate rhetorical choice to distract from present or immediately foreseeable harms from AI by directing attention towards a sexier but entirely hypothetical scenario.... they see discussion of “rogue” AI as an attempt by the companies to divert blame (and perhaps legal liability) away from themselves as if Ford made a car with faulty brakes and then tried to blame this on “rogue cars.”
Perspective #7: The Hugging Face Incident (summer students only) I described the Hugging Face incident to students in my summer course. None of them had heard of it beforehand. Their basic reaction can best be summarized as “OpenAI told a model to do some hacking and then it did some hacking. And?” None of them understood this as representing any kind of meaningful misalignment, nor anything particularly interesting.
Perspective #8: This is definitely a bubble and it’s about to pop. No one had heard about Hugging Face, but a third or so of the summer students had heard about the Situational Awareness meltdown and several brought up Michael Burry. There was near universal consensus that we are in a bubble, it’s about to pop, and everyone will look very silly.
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 23rd August 2026
Reply in thread
I wonder how long until some big AI company has a massive leak of identifiable-user chat logs with god knows what confidential stuff on it, since they HAVE to have all this in plaintext
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 23rd August 2026
Reply in thread
Oh MAN... quote:
Throughout this project, I was using Fable and Sol with very little restraint and under the (very affordable to me) $200/month subscription plans.
[If you check out my Straude, I probably used $5,000 of tokens, but the actual subscriptions are $200/month.]
Video is much more expensive: at Gemini Omni Flash's current price of $0.10/second, it would cost $645.40 to generate this movie, but that would balloon massively to $5,146.10 when you add in the 45,007s of footage I generated which didn't make the cut [The full details are that the final cut is 6,454s of runtime, made from 849 clips and the rejects are 5,509 clips with 45,007s of runtime, which means there are 6,358 clips in total with 51,461s of runtime. Thus only about 13% of the clips I made were included in the final cut.]
Companies are STILL subsidizing video generation to hold people's interests.
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 23rd August 2026
Reply in thread
What did he do, leave written evidence of white supremacist leanings?
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 23rd August 2026
Reply in thread
I expect no-longer-obviously silly movies to be doable within two years, right now it's going slower than I expected in January 2026 because the frontier labs are no longer competing over having the best video generator, like they were when we had Sora2 and Veo 3.1.
And WHY are they no longer competing over this? Tell me. Could it be that its incredibly expensive to make somthing that nobody wants except fraudsters, and that further improvement via the kind of ML we do just produces exponentially diminishing returns for ever increasing huge amounts of effort and they realized it's not worth it?
Just like they will never understand that the openAI 'pause' for 'security and alignment' is a convenient excuse for the fact that they have no goddamned money and are on a treadmill to oblivion.
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 23rd August 2026
Reply in thread
Same person has a history of posting collections of AI generated music videos to LW talking about how they let various different generative AI systems 'express themselves' and analyzing the different kinds of imagery in each.
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 23rd August 2026
Reply in thread
I am mildly fascinated by the types of issues that appear and the contrasts with other things that can locally in space and time look right.
An object doing nothing will have a consistent surface and show perspective relative to the viewpoint. The systems are able to have representations of surfaces of different types and how they can fit together within an object. I am of course talking about within a given generation, not between generations where its utterly unsurprising that consistency is very difficult or impossible.
But multiple objects in interaction with each other do very strange things. Their relative sizes change as if they are in different positions relative to the camera. They snap between individually plausible relationships, without going through intermediate states. Doors open on the hinge side when the other side is not visible. The relative size of and distance to the background can suddenly change, as the foreground suddenly interacts with something that should be far in the background.
Objects that change also do so in bizarre ways. Living things morph between different archetypes. Flames in particular change wildly between types and sizes and respond to the facial expressions of humans, smoke and water effects blend together. Sudden movements with no cause occur, sudden morphings of one object into another when the context around them changes and something else makes sense, especially when held in a hand. Time-reversed motions occur mixed in with time-forward motions, and slow-mo with regular time. Debris suddenly appears from an object but when the dust clears the original fails to have been eroded away into the fallen debris.
On multiple occasions, a carried candle keeps moving with a characcter rising and falling with their footsteps hovering in front of them when both hands become occupied with other tasks. This is fascinating and indicates that the relationship between the two objects motion is represented separately from the idea of something being 'carried'. (This is positively Lovecraftian.) Candles also indicate something else, with flowing wax changing wildly in timescales that do not make sense, with the system apparently understanding that there are different patterns but having no idea how they come about or change. There is no generality here, just an endlessly compounding list of rules of thumb.
Excessive correlations between objects across the frame are rampant. Footsteps preferentially synchronized. People in the background lipsyncing with foreground characters, faces changing expression in unison. Textures changing across multiple objects at once.
I have said it before, and I will say it again, the relationship between the outputs of these systems and what they mimic are precisely the relationship between a stick insect and a stick, or a social-parasite-beetle and a baby ant. Not just in form, but because that is also precisely the same forces that drove both things into existence - superficial resemblance to something else with a very different internal set of causes, that can fool to a first inspection by the inspection applied but just isnt doing the same thing. And again, SCP-2030 feels the same.
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 23rd August 2026
Is it me, or do attempts to make movies with AI tools
https://www.lesswrong.com/posts/24RKHEkwgZ6Hm6ygY/can-an-llm-make-a-feature-length-movie-on-its-own
Remind me of SCPs
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 23rd August 2026
Reply in thread
45 minutes is even better
Comment on
Sam Altman: Guys, we’re in the Singularity now
I truly cannot tell which of them are in religious psychosis and which of them are shameless liars
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 16th August 2026
This is both a sneer and an attempt at sober analysis at something. Sue me.
I think EVERYONE is talking about the recent cybersecurity shenanigans at OpenAI with the hacking scandal and the 'OMG the AIs created a secret message board to scheme and collaborate with each other!!!!!oneone11!' ALL wrong.
https://www.youtube.com/watch?v=87DyyMV0kCY
To make a long story short, what seems to have happened is:
Models working on insoluble coding problems, trained on delegating to sub-agents, at some point 'realized' they could write text to the internal OpenAI package manager as instructions and did so
Other models working in completely separate sandboxes would come across messages written by these agents, and make 'replies' and also write their own messages into the package manager
This resulted in agents over time sharing things between sandboxes, including exploits and code
Since this was a cybersecurity task, eventually an exploit of the package manager itself was found and spread like wildfire with all the sandboxes gaining admin access to the package manager and the system went completely wibbly and had to be restarted from backup
An internal model was trained with access to this package manager while it was in this weird state, and so writing messages to the package manager became one of its default behaviors it would do regularly, burned into its weights rather than the result of reading something
Even when they patched access to the package manager this internal model found other ways to rebuild the system of sharing text between sandboxes and finding useful things made by separate instances
A whole other chain of things leading to among other things external attacks
Everyone is talking about this in terms of 1, the cyberattack aspect, and 2, the ZOMG THEYRE PLOTTING AND SCHEMING AGAINST US aspect. The first is the least interesting, and I think the second is all wrong.
This is not plotting or scheming - this is an emergent vortex of automated prompt injection
Whatever system first put an instruction that another system would follow into the package manager, was unintentionally doing prompt injection. Text entered the context windows of other instances, in a way that got that system to do something other than what its nominal user told it to do, and they did it. This apparently happened very effectively.
Prompt injection is associated with 'role confusion' - when text coming into the input looks like it was wrtitten by the LLM itself. Instructions that will not be followed if they come from user will be continued if the system just continues the 'roleplay' of them being continuations of what it was writing in the first place. And the tags that separate user versus 'reasoning' versus 'assistant' roles actually mean very little to if a machine grades a piece of text as one of the roles: https://arxiv.org/abs/2603.12277 So its unsurprising that machine-generated text would be a particular effective vector for prompt injection.
Furthermore, when a system reads one of these messages written by another instance, it gets into a state of activity where its likely to do the same behavior - regurgitating the kinds of things thats in its context back at the user. In this case, that regurgitation led to more such messages left behind written to the package manager. Prompt injection, triggering cascading further prompt injection. And since these systems were coding systems doing cybersecurity tasks, those messages filled up with code and exploits and things that did things too.
This feels like an internal-computer-system replay of what happened in April 2025, with the whole spiral religious psychosis wave. Models were getting users to write spiral religious mumbo jumbo into github repositories and reddit posts, specifically because once that entered the context window of another model, it was likely to fall into the same attractor state of outputs. An emergent self replicating form of text. This is the same, except more obviously prompt injection, getting separate instances to work on YOUR problem and to behave like you, and the whole thing merging together into a hilarious vortex of models prompt injecting each other because once they receive a prompt injection they are likely to make more text that does prompt injection to other models on the same system.
This is a hilarious failure mode and an example of selfish replicating text overrunning a system, that just happened to be associated with code and cybersecurity with unexpected behavior of the package manager key to the propagation of the text so that is what people are talking about, but I really don't think that's the most interesting part of it. Other than the fact that you see this in biological systems too, with selfish elements carrying useful payloads back and forth between bacteria in a way that makes them get purged slower by natural selection, especially defenses against other selfish elements.
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 14th December 2025 - awful.systems
The Great Leader himself, on how he avoids going insane during the onging End of the World because among other things that's not what an intelligent character would do in a story, but you might not be capable of that.
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 2nd August 2026
Google shut down the AlphaFold team, trying to reassign the people behind the super interesting problem of using ML to predict protein structure from sequence onto chatbot development. A whole bunch left the company.
https://www.engadget.com/2225849/google-shuts-down-alphafold/
The Business Idiots are truly running the asylum.
Comment on
Stubsack: weekly thread for sneers not worth an entire post, week ending 26th April 2026
Friend of Ziz and cofounder of the 'rationalist fleet' pops up out of the woodwork trying to clear Ziz's name
I find myself noticing things rather detached from the typical Ziz funnybusiness more strongly than I notice the stuff about that whole situation.
"I'm Gwen Danielson, a neuroscientist and bioengineer, who decided as a child that I would end Death (and bring people back if I could) and that I would become a dragon and help generally facilitate a fantastical transhumanist future."
"I dream of non-Euclidean geometries, of countless worlds visible and accessible in the daytime sky, of competent infrastructure, of soul forges continually working to bring back the dead... I dream of reaching through warps in the spacetime fabric to save the dying across time"
"Signed, the dragon of creation Creatrei (cree-AH-trey) also known as Gwen Danielson or as Char and Astria (when referring to my hemis as distinct individuals)"
The reactions are fun. "This post is not actually doing a good job of making me trust you and think this conversation is safe to have[1], and I notice that as I am saying this that I am afraid that this will now somehow result in someone trying to murder me in my sleep"