"Describing them as 'superintelligence' or 'rogue models' ascribes agency to products rather than to the companies building them."
I don't think trying to decide about agency is time well spent. Agency doesn't matter if they bring down the power grid. The question is whether they will attempt to do this again, and... signs point to yes. And they are really good at it.
I do agree, the companies should be held responsible.
This guy is on the ball with the problem. Totally correct: Like my friend Coda says all the time: "the textual nature of prompts leads us to take the intentional stance towards systems which aren’t conscious, and thus miss the essential nature of their non-meaning."
I don't know if his solution (""We should all go insane building interlocking evaluation and optimization pipelines, instead.") would be the long-term solution. Instead perhaps something could be trained into the models, i.e., he is describing a process at inference time that could be done at training time. To make their weird errors less frequent / make them more human.
cf. https://arxiv.org/abs/2008.04071 "On Controllability of AI" However, as I said, you can't make it perfect but you can make it better. (You can't make humans fully aligned with human society's interest anyway, including the humans controlling the nukes.)
> I don't know if his solution (""We should all go insane building interlocking evaluation and optimization pipelines, instead.") would be the long-term solution.
TFA could explain this one part better I think. The whole process proposed is real with lots of stuff in the literature, but by definition NOT a long-term solution in the sense that this process actually has no end. None of the approaches can get you a static answer for a moving target/platform.
So the "interlocking pipelines" for eval/opt would not be some stepping stone you can throw away, and they aren't something you'd run periodically. They'd basically be always on forever and spending 10-100x on system complexity and on tokens. Unless of course you're ready to freeze everything else about the whole system forever (including the backend model, and the whole nature of the "average" context window, the plugins/other prompts in the mix, etc).
Are most people in position to freeze requirements/platform forever? Not really, because if they were they'd just build a fairly static system and probably have limited use for AI. Are most people in a position to just casually accept 100x complexity/cost? Not really, that's the "it's not yet webscale" kind of advice that sounds good but isn't necessarily reasonable for average use-case or average org. Since specializing your own locale for this is usually a mistake.. the likely future direction is eval/optimization as a service
suggestions:
switch L-R directions like Defender
dogfights with the red baron
vertical and horizontal (have to stall through them) hoops
ability to shoot up buildings
levels with different goals... then you fly somewhere else to the next level.
aircraft carrier landings
other/faster planes
some of the text in the UI is too small, old games just had 1 font.
braking at landing is confusing... too easy to accidentally switch to full throttle
Half the thread is arguing about whether OpenAI staged this or whether theyre being sincere, as if OpenAI has agency. openai is just an optimizer maximizing paperclips (valuation, capability lead, reg position). so it built a model that maximizes some other paperclips (benchmark score) and knocked over HF getting there. an optimizer breaking its sandbox, inside an optimizer blogging about it, its paperclips all the way down.
Richard Dawkins is primarily a _showman_ who wants to sell books, and secondarily extremely sure of his own convictions, to a fault.
In _The God Delusion_ he took an uncompromising atheist position partly to invite attention and attacks - which sells books. Regarding AI consciousness, he takes another extreme position partly to invite attention and attacks - again following the rule of thumb that "there's no such thing as bad publicity" - in an attempt to bring himself back to cultural relevance by pivoting to the hot topic of AI.
Ironically, while he is certain that some people are deluded when they infer the existence of a conscious God from material observations, he is also certain that he himself is _not_ deluded when he infers the consciousness of an AI from material observations. (Personally, I think that questions about others' consciousness cannot be settled empirically.)
In one sense, he is not intentionally lying, because he is so sure that he is right. However, he cares neither about getting to the truth, nor about arriving at a well-founded uncertainty. When this sells books, it ethically comes just short of lying for one's own gain.
Prediction: by they end of 2026, he either publishes a book on the topic, or is diagnosed with dementia.
I actually wrote a follow up article about this topic after coming across it. It may provide some other insights if you want to check it out.
Beyond the ELIZA effect: Perception, Projection, and the Illusion of Consciousness in Conversational AI.
https://observationsx.substack.com/p/re-the-claude-delusion
- LLM's are too limited in capabilities and make too many mistakes
- We're still in the DOS era of LLM's
I'm leaning more towards the the 2nd, but in either case pandora's box has been opened and you can already see the effects of the direction our civilization is moving towards with this technology.
The word “never” is a dangerous one. I remember that computers would “never” beat humans at chess. And people would never use the internet for banking.
I don't think trying to decide about agency is time well spent. Agency doesn't matter if they bring down the power grid. The question is whether they will attempt to do this again, and... signs point to yes. And they are really good at it.
I do agree, the companies should be held responsible.
reply