Computer pioneer Alan Turing’s remarks in 1950 on the question, “Can machines think?” were misquoted, misinterpreted and morphed into the so-called “Turing Test”. The modern version says if you can’t tell the difference between communicating with a machine and a human, the machine is intelligent. What Turing actually said was that by the year 2000 people would be using words like “thinking” and “intelligent” to describe computers, because interacting with them would be so similar to interacting with people. Computer scientists do not sit down and say alrighty, let’s put this new software to the Turing Test - by Grabthar’s Hammer, it passed! We’ve achieved Artificial Intelligence!
I think the Chinese room argument published in 1980 gives a pretty convincing reason why the Turing test doesn’t demonstrate intelligence.
I am sceptical of this thought experiment as it seems to imply that what goes on within the human brain is not computable. For reference: every single physical effect that we have thus far discovered can be computed/simulated on a Turing machine.
The argument itself is also riddled with vagueness and handwaving: it gives no definition of understanding but presumes it as something that has a definite location, and also it may well be possible that taking the time to run the program inevitably causes understanding of Chinese after even the first word returned. Remember: executing these instructions could take billions of years for the presumably immortal human in the room, and we expect the human to be so thorough that they execute each of the trillions of instructions without error.
Indeed, the Turing test is insufficient to test for intelligence, but the statement that the Chinese room argument tries to support is much, much stronger than that. It essentially argues that computers can’t be intelligent at all.
That just shows a fundamental misunderstanding of levels. Neither the computer nor the human understands Chinese. Both the programs do, however.
The programs don’t really understand Chinese either. They are just filled with an understanding that is provided to them up-front. I mean as in they do not derive that understanding from something they perceive where there was no understanding before, they don’t draw conclusions, don’t understand words from context,… the way an intelligent being would learn a language.
Nothing in the thought experiment says that the program doesn’t behave that way. If the program really seems like it understands language to an outside observer, you would assume it did learn language that way.
Others have provided better answers than mine, pointing out that the Chinese room argument only makes sense if your premise is that a “program” is qualitatively different from what goes on in a human brain/mind.
Programs clearly understand words from context. Try making it do translation tasks, it can properly translate “tear” to either 泪水 (tears from crying) or 撕破 (to rend) based on context
The problem with the experiment is that there exists a set of instructions for which the ability to complete them necessitates understanding due to conditional dependence on the state in each iteration.
In which case, only agents that can actually understand the state in the Chinese would be able to successfully continue.
So it’s a great experiment for the solipsism of understanding as it relates to following pure functional operations, but not functions that have state changing side effects where future results depend on understanding the current state.
There’s a pretty significant body of evidence by now that transformers can in fact ‘understand’ in this sense, from interpretability research around neural network features in SAE work, linear representations of world models starting with the Othello-GPT work, and the Skill-Mix work where GPT-4 and later models are beyond reasonable statistical chance at the level of complexity for being able to combine different skills without understanding them.
If the models were just Markov chains (where prior state doesn’t impact current operation), the Chinese room is very applicable. But pretty much by definition transformer self-attention violates the Markov property.
TL;DR: It’s a very obsolete thought experiment whose continued misapplication flies in the face of empirical evidence at least since around early 2023.
It was invalid when he originally proposed it because it assumes a unique mystical ability for the atoms that make up our brains. For Searle the atoms in our brain have a quality that cannot be duplicated by other atoms simply because they aren’t in what he recognizes as a human being.
It’s why he claims the machine translation system system is incapable of understanding because the claim assumes it is possible.
It’s self contradictory. He won’t consider it possible because it hasn’t been shown to be possible.
The Chinese room experiment only demonstrates how the Turing test isn’t valid. It’s got nothing to do with LLMs.
I would be curious about that significant body of research though, if you’ve got a link to some papers.
Searle argued from his personal truth that a mystic soul is responsible for sapience.
His argument against a computer system having consciousness is this:
" In order for this reply to be remotely plausible, one must take it for granted that consciousness can be the product of an information processing “system”, and does not require anything resembling the actual biology of the brain."
-Searle
https://en.m.wikipedia.org/wiki/Chinese_room
Isn’t the brain just an information processing system?
Brilliant thought experiment. I never heard of it before. It does seem to describe what’s happening - if only there were a way to turn it into a meme so modern audiences could understand it.