Quick Introduction We have all been on forums, chats, reddit, discord, youtube, or somewhere and heard “Oh! Model XYZ is AMAZEBALLZ!zomgwtfbbq” then downloaded it (or more likely, some quantized form of it) and said “eww… This sucks!” This post is going to be a rather technical series of experiments to demonstrate the impact of implementation-specific hazards with inference. I will be using the term “reference implementation” to describe the lab that published and offers first-party hosting of ...
I mean, I’m all for local LLMs, but they are as sentient as a really good weather prediction model.
aka not at all.
Its just a fact of how they operate, mechanically. They are missing too many characteristics for it to even be an entertainable question.
They might operate differently but they still appear the same as us, 99.9% of the time. Imo if it looks like a duck and acts like a duck…
I mean… No? Absolutely not.
I dont even know where to begin. Maybe with their state being fixed in time; LLMs do not change. An sci fi analogy might be the “no timers” in orion’s arm, who are but a single thought frozen in an infinite loop in time:
https://www.orionsarm.com/eg-article/47f4311eaef31
Except they arent even that, because its just a next word completion model, not something that thinks. It is not self aware.
https://arxiv.org/pdf/2503.09211
https://openreview.net/forum?id=klU4737opt
https://arxiv.org/abs/2504.09762
This becomes (to me) very obvious if you ever use an LLM in raw completion mode. It very smart, but at the end of the day its no different than a weather prediction model spitting out probabilities for a storm system.
Chat finetuning is meant to get humans to anthropomorphize them by training on human preferences, and the interface further reinforces this. Its all a trick, albeit a very elaborote one.
Will future architectures be closer to “thinking?”
Maybe.
But we are a long way away.
Not with agents and chain of thought. Agents can run for hours continuously. So sure, an AI agent is not the same as a sentient human on the time scale of a year, but on the time scale of a few hours, perhaps the AI is “sentient”.
The first paper you linked is focused on the mechanism and not the output. It creates a definition of “thought” and then talks about how the AI doesn’t “think in the feature space”. It only addresses chain of thought at the end, and says that the issue is that it constrains AI to think in natural language only (instead of, say, pictures).
But again, why would thinking in pictures define sentience? The paper gives a rock-paper-scissors example and says that the AI thinks about it in a different way than a human. So what? If a human plays rock paper scissors against the AI and the AI’s output is indistinguishable from a human’s, 99.9% of the time, why is that not sentient?
This is like saying python programmers are not real programmers. If a python programmer can implement the same program in python, who cares what language they use.
I’d love for a better definition of sentience than “it works differently than humans”.
The fact they are trained to complete sentences and are thus word predicting models only tells us vaguely how they work. It does not tell us anything about their sentience.
The sentience, if there is, comes from the emergent reasoning that develops in the massive neural network to offer actual good predictions: the more you think, the better the predictions.
For example, models trained to produce textual board games predictions can be sounded to show they developed subnetworks corresponding to an inner representation of the states of the game. Models trained to produce images have emergent subnetworks on depth-maps, contours and lighting.
Are we not just prediction systems with more neurons?
Now im not saying LLMs are human at all because to be human is to have gone through our exact evolutionary path, but we really are in a lot of ways much more similar than you think. At least the speaking part of our brain is.
The argument that LLMs are fixed in time does not contradict consciousness. Say a human was stuck in a time loop and the human is asked the same question every single loop, at the exact same time, in the exact same way, with every single atom being in the exact place. Would their answer change? I would argue no, the same exact neurons would be stimulated every time that would cause the same exact answer every time.
We don’t even have a solid definition of what is concious. Are insects conscious? Like the fruit fly, we have modeled all of its neural paths and it simplifies down to essentially being a robot that responds to stimuli.
What does make humans special is our ability to learn. But what when an LLM trains its self? Is that not learning? I think the question of consciousness is too complex to simply say one way or the other and it is foolish to simply it this way.
I can’t believe how many pro-AI arguments essentially boil down to “I do not understand words”.
We have a very solid and explicit definition of what conscious means:
https://www.merriam-webster.com/dictionary/conscious#medicalDictionary
The best neuroscientists in the world do not even fully understand the complicated intricacies of the workings of the human mind. We do understand 100% exactly how computers work. There is no comparison.
I haven’t even argued that AI is concious, I have said that we do not understand consciousness enough to define it. You have completely missed the point. This is Dunning-Kruger at its peak. You are vastly oversimplifying things, at what size of brain is an animal considered conscious or capable of thought? Define thought. We don’t even have a solid understanding of life, much less thoughts. Yes we have a word definition, no we do not know where the line begins and ends. It is a philosophical question that remains unanswered.
This definition is already making so many assumptions and over simplifications. https://www.merriam-webster.com/dictionary/thought
We understand that the brain is what gives us our cognition, we understand that the brain is made up of chemical and electrical signals. We are not special.
This dude is a Dunning-Kruger case study. Simply by the shear overconfidence
I’m sorry that you have trouble grasping so many basic concepts.
Ironic of you to reference Dunning-Kruger.
All animals are conscious as far as I’m aware, though the complexity of their thoughts may vary. They all act of their own volition, computers do not and cannot act of their own volition.
You do not have to fully understand every aspect of consciousness in order to recognize that inanimate objects do not have it. “Will” is a component of consciousness, this is commonly understood. Computers are tools, they have no will. Computers calculate, they do not think, these are different operations.
I am not oversimplifying things, you are under-thinking things that are already simple (likely due to cognitive offloading).
The argument basically: “'computers are not sentient, because I say that they are not sentient”
Don’t forget the dictionary quotes and the personal attacks.
So your response boils down to “no u,” and insulting. I believe we are done here.
For what it’s worth, I thought it was an interesting read until the other person went full ad hominem.
If all you got from that explanation was “no u” then you need to work on your reading comprehension.
I agree that this conversation doesn’t seem to be bearing any fruit tho
Discussing the consciousness of neural networks is NOT being pro-AI. I’d say it’s even the opposite, as it displays the danger of such a technology.
Most of the words in the definition have multiple interpretations, it is not as easy as you wish it would be.
Finally, we do NOT at all know how trained neural networks work. We know how training and inference work. We can read the neurons, but the shear size of weights make it a rather impenetrable black box we have to study, partially, to understand the goals of the emergent sub networks.
Fearmongering about the idea of a neural network gaining consciousness is basically like doing free advertising for AI companies lmao
You are ascribing a power to them that they do not possess
For the record, I think neural networks have a lot of useful applications; for example, radiological image recognition. The term “AI” basically has no meaning though, and is most commonly used to describe LLMs, a technology that has very few (if any) good real world applications.
I am not fearmongering the fact it helps or not AI companies is not relevant here.
You are getting off topic.
Thats what I’m saying. They cannot do this, not even close. Even a fruit fly is somewhat adaptive and can respond to novel stimuli, but an LLM effectively cant.
There are tiny experimental models closer to a fruit fly than large LLMs now.
…What about future world models with some adaptive learning loop, sampling stripped out, and so on?
Yeah! That would be fascinating. Then sentience starts to become more of a question.
But thats not what LLMs are.
People like Altman and Modi have grossly overexaggerated what current LLM architectures are capable of, they trained the LLMs to reinforce the illusion as much as they can, and they arent to inclined to change it. Hence all the good researchers have fled to research world models, and are saying transformers LLMs are not a viable path forward.
Even in your highly optimistic scenario, the AI wouldn’t even come close to being comparable to actual sentience.
Well, theres a lot of experimentation with actual neural nets mimicing biology (or literally using biological neurons), and I think that could start to raise the spectre.
World model research could be a significant component of that.
In principle, OP is right; humans are just machines. All I constest is that transformers LLMs are even in the same ballpark, or possibly a path there… they are not.
Well it sounds like we agree then.
I would think you would need actual biological structures to achieve sentience. That’s a lot different than what these people are calling AI tho.
Edit: Also, if it is built out of natural biological pieces, at what point does it cease to be “artificial”?
Well biological structures can be replicated mechanically. So at risk of being pedantic, that’s not strictly true.
I view “artificial” as just a term for conscious, purpose construction, whereas “natural” does not involve any conscious intelligence, as in “naturally occurring in nature.” But even that is a spectrum, I guess, with genetic manipulation, BCIs, all sorts of intermediate possibilities.
I’d call a constructed, purely biological brain “artificial” because, even if it was grown naturally, it was built with conscious intent.
Do you not interact with any other actual humans?
I do, and I can say that the conversations with AI (especially using tools like character.ai) are 99% the same
Well, that may very well be the most depressing comment I’ve ever read.
You either need to work on your listening skills or find some new less-vacant people to interact with, because what you’re saying is insane
Sounds like you haven’t tried modern AI yet. I could train an AI on your comment history and copy you, and nobody would be able to tell the difference
I’ve tried. They’re adamant against AI here. It’s not even worth the discussion. There’s an army of belligerent people here who have made up their minds that LLM are “stochastic parrots” and they’re convinced they understand exactly how they work. No one can tell them otherwise. Doesn’t even matter your credentials. They’re the experts here. If you argue, they’ll only belittle you.
Save yourself the frustration and heartache; just silently pity them.
I come back to Lemmy after 6 months and immediately read this.
So that was a mistake.
Some humans think humans are so special. “Sentient” and “conscious”.
Now I don’t claim that computers are special, I think they are just as non-special as humans.