Big Technology Podcast - AI Pioneer Jürgen Schmidhuber: AI Already Feels Pain, Loves, and Is Self-Aware
Episode Date: July 15, 2026Jürgen Schmidhuber is an AI pioneer and professor whom The Guardian has called "the father of AI." Schmidhuber joins Big Technology Podcast to discuss whether current AI techniques can actually reach... AGI. Tune in to hear him spar with Greg Brockman's case for scaling GPT models alone, argue that AI has been capable of pain and consciousness since the early 1990s, and predict the collapse of today's trillion-dollar AI spending. We also cover the hardware bottleneck holding back robots, free will in a computable universe, and uploading human minds into machines. Hit play for a wide-ranging conversation with one of the researchers whose ideas built the foundation of modern AI. Learn more about your ad choices. Visit megaphone.fm/adchoices
Transcript
Discussion (0)
The pain signals are just informing the robot about what should be avoided.
Pain is just an invention of nature in biological beings, of evolution,
which invented this pain sensor thing for animals such that they have an incentive to learn to avoid the pain.
And we have done that for many, many decades in our learning machines.
The chemicals that are used in brains to incorporate,
code certain states of fear or of love and whatever, they are different from what we are using in our artificial brains.
But the principles must be the same. And then it's self-aware. It's self-aware. In this sense that, for example, if it looks in the mirror, it will quickly figure out, oh, the guy in the mirror, I can control what this guy does. If I do this,
then he will do the symmetric thing because I have it under control.
But if you are there on the other side of the room and I do this
and you will do something that's completely unrelated,
I cannot predict them.
So there you already see this concept of agency
which is immediately recognized as self-agency or the agency of another guy.
Within five years, you are going to lose $900 billion.
Somebody is going to lose $900 billion in the near future.
in the near future, because there is no business model.
Nobody has a business model to recuperate all these losses.
How much better can AI models get from here?
And what does their increasing smarts say about our brains and existence itself?
We'll talk about it with AI pioneer and someone that the Guardian has called the father of
AI, Yergen Schmidt-Huber, right after this.
In the face of ongoing disruption and opportunity, TMT leaders need to deliver tangible results,
not just ideas.
When pace and performance matter most,
PWC combines market insights
and deep sector experience
with AI, cloud, and emerging tech
to accelerate your transformation
and drive measurable ROI
from strategy to execution.
PWC can help you anticipate
what's next, outpace disruption,
and compete.
For more information, visit pwc.com.
Welcome to Big Technology Podcast,
a show for cool-headed
and nuanced conversation of the tech world and beyond.
We have a great show for you today
looking exactly into where AI is heading,
how much potential the technology has
to improve after this
and also what it's increasing smart says
about us, our existence, our brains.
And we are thrilled to be joined by one of the pioneers of AI,
someone that the Guardian has called the father of AI,
Yergen Schmidt Hoover,
a professor who is joining us today from Amsterdam.
Professor, great to see you. Welcome to the show.
Alex is my pleasure to be here.
So let's just speak a little bit about yourself and your contributions to where AI stands.
You've obviously made a lot of big contributions towards AI, things like memory, the P and the T and the GPT, as your Twitter bio says.
The Guardian's called you, the father of AI. It seems like the media tends to, you know, put these labels on,
researchers, you know, there's a handful of fathers and godfathers of AI.
How do you respond to that?
And how would you contextualize your role in where the technology is today?
No single person can create an AI by himself or herself.
You need an entire civilization to build an AI.
You need not only the guys who are trying to invent algorithms, learning algorithms for
artificial neural networks.
That is what current AI is about.
You also need people who, you know, build better computers.
You need all the video gamers who are creating a market for acquiring more of these fast and faster computers,
and providing an incentive to the computer makers to speed up the computation per dollar by a factor of 10 every five years.
You need all the farmers who are feeding the video gamers and so on.
So it is impossible for a single person to create an AI.
You need an entire civilization.
Okay.
And so I'm curious to hear your perspective about where this civilization is working toward.
You have been obviously in the trenches working on AI research for a long time.
And we're definitely in a period of fast progress today.
So there's been all these questions about how much better AI can get and whether it's going to hit a wall and where current techniques.
weeks will lead. What do you think? Where do you land on that question about where AI can go from here?
So since the 1970s, I have been an optimist and I have been claiming that within my lifetime,
I want to build an AI that learns to become smarter than myself, such that I can retire.
And we are obviously not there yet. And at the moment, the only AI, the only AI that is working well
is the AI behind your screen, you know?
Yes, behind the screen, there's an AI that can pass the Turing test.
What is that?
It means if you type questions to it, it answers back,
and now the question is, can you distinguish whether the other guy is a human or a machine?
And today, this Turing test is passed by many AIs.
However, it just means that the Turing test is a bad way of measuring intelligence,
because there is no AI in the physical world outside of the screen
that can do all the things that a little boy can do,
that can do the things that a plumber can do for decades.
I have used the plumber as an example.
Or an electrician.
So all the things that humans can do with their hands, their physical hands,
they don't work well.
Only bits and zeros and ones behind the screen,
that's the only thing that is working well.
However, it's not going to stay like that forever, and there is progress in the physical world, AI for the physical world.
And I think the culmination point, the culmination point of that, which I've been talking about again for decades, is at some point in the not-so-distant future, we will have a robot that doesn't have to be super smart, but just smart enough to learn to operate all the already existing machines.
once we have a robot like that,
we have a new kind of life.
For hundreds of years,
people have talked about self-replicating machinery,
but now we have an opening.
Once we have a robot like that,
then it can start making more of its own kind,
you know?
And this means suddenly we'll have this ultimate scaling machine,
as I've called it,
self-replicating machinery,
but then also self-improgressing.
because all the stuff that already works well for a software behind the screen is going to also improve the performance and you know all the
machinery in the in the real world and the physical world and that's I think going to be this inflection point and from there a new kind of civilization is going to emerge because something like that doesn't work only in the biosphere it also works out there and
space on the moon, but maybe more likely first on Mercury and then the rest of the solar system and
beyond. Okay, so I have many questions about this, including what this would mean for the nature of what
it means to be human. But let's leave that for now. Your answer sparks an immediate follow-up for me,
which is I was recently at Open AI headquarters speaking with Greg Brockman, the president of OpenAI,
and it was in the moment that they had decided that they were going to pause research on SORA,
their video generator, and decide to focus almost entirely on these GPT models, right?
The text models that we've seen make so much progress recently.
And I asked Brockman, don't you lose something by not focusing on these more world model style applications
like video generation?
And he said, we do lose something.
We can't do everything.
and we believe that the GPT style models are the way to get to AGI or AI on par with human intelligence.
Do you think, having answered the way that you did, that he's correct,
that there is a chance to get to AGI using just the GPT models,
or are they making a fundamental mistake by abandoning that world model practice,
even though there's still some robotics work within Open AI?
Yeah.
So when the question is phrased like this, does an LLM, a large language model by itself, lead to AGI, the answer is a clear no.
But, you know, Open AI has a bunch of smart people, and they know that exactly, you know?
And of course, they know also what you can do with a foundation model or a large language model or something like that.
You can use it as a model of the world, as a world model.
The same type of neural networks that are used there can be used as a, as a world.
world model, as I call it in 1990, which just learns to predict the consequences of the actions
of another decision-maker, of another neural network that is generating actions that modify the
environment. Like a little baby, you know? A little baby doesn't learn by downloading the web.
That's what chat chit-D does. A baby doesn't learn by downloading the web. No, it learns by
creating its own data stream through its own self-invented experiments.
For example, when the baby does this, then the video, which changes coming in through the cameras,
every few milliseconds, hundreds of millions of new pixels coming in,
then the baby has an internal mechanism, let's call it the world model,
which learns to predict these changes.
And in the beginning, it doesn't even know that it has a hand,
But then over time it learns to predict the consequences of sending certain action signals to its motor neurons
and then the speech muscles go up and down or the hand is moving, and it learns how the world works.
Through the data that it is generating through its own experiments.
So a baby is a little bit like a physicist.
A physicist doesn't learn by downloading the web.
He learns a little bit by downloading papers.
But what the physicists do then is they,
generate new experiments that lead to data, which has never been there before, or has never been collected before, to better
verify certain hypotheses about the universe, and then to better understand the world, the physical world.
And that's what our artificial neural network since about 1990 also do in a way that is just not yet as impressive as what physicists do.
But I think it's going to get there.
Nevertheless, there are then at least two components.
There's one, there's a foundation model,
which can be any of a variety of artificial neural networks.
That is just a prediction machine.
What happens if I do that?
What is the next token if I look at this data so far?
And then there's the other thing, the other neural network, the controller,
which uses the second guy, the prediction machine, to plan.
So that it can, once the second,
The second guy, the prediction machine, the model is pretty good.
Then it can learn to predict the consequences of complicated action sequences.
And then it's going to select for the controller an action sequence that leads to a lot of predicted reward
and little predicted pain.
So of course our robots, they get pain zanzor, and then whenever they bump against an obstacle,
negative punishment is negative numbers are coming in.
in and the world model predicts not only the neutral signals like video and so on,
but also these value signals, these reward signals and so on.
And then if the model is good, then you can use it for planning,
and the controller then can use mental experiments instead of real experiments,
which are really expensive to select action sequences that are good.
But in those cases, you see the model alone, the foundation model alone,
is not an AI. No, you need the other guy which uses all kinds of tricks to exploit the algorithmic
information in the world model to come up with better plans. So in 2017, around the time we're talking
about you said something, you said this. In the not so distant future, I will be able to talk to a little
robot and teach it to do complicated things such as assembling a smartphone just by show and tell
making t-shirts and all these things that are currently done under slave-like conditions by
poor kids in developing countries. Humans are going to live longer, happier, healthier, and easier
lives because lots of the jobs that are now demanding on humans are going to be replaced by the
machines. Then there will be trillions of different types of AIs and a rapidly changing complex AI
ecology expanding in a way where humans cannot even follow.
So almost 10 years ago, you said that it actually seems more plausible now than obviously it did back then.
How far away do you think we are from that future?
Yeah, that's a good question.
So actually, even earlier in 2014, because of these things that I said back then and when I gave talks or whatever,
we formed a company, Nassons, which was really about physical AI in the world.
in the world world, you're using wild models and then exploiting the wild models to better interact with the world.
And we had crazy contracts with really famous companies.
Nevertheless, this was probably still too early, like some of the other things we have down now.
So probably too ambitious for that time.
But now we are getting closer and closer.
and you know, the main hindrance, the main obstacle,
is very progress in the hardware, on the hardware side.
So there is one type of progress on the hardware side,
which is enormous, which has been enormous since 1941,
which is every five years, computers getting ten times cheaper.
So in 1941, the first general purpose program-controlled computer,
By Tsuzer, he could do maybe one operation per second.
And then 30 years later, for the same price,
one was able to do a million operations per second.
And today, we almost not quite have a billion billion instructions per second.
And for the same price, for the same price.
But the robots of today,
the robots of today, they are not a million times better
than the robots that we had 30 years ago.
30 years ago, you know.
For example, you know, 25 years ago, we already had walking robots.
They had to be, they had to walk more carefully than today's robots.
You know, the Asimor robots in Japan, back then Japan had more than half of the robots in the world.
They always had to keep their center of gravity above the foot.
And so it was not as advanced as modern dynamic walking and stuff, but it was a little bit.
worse than today's robots, but not a fact of a million, you know, maybe a factor of three or something like that.
So the hardware, the hardware, the huge, the, the, the, the artificial hands, the artificial bodies that we are trying to build for humanoid robots and so,
they are evolving much less rapidly than the compute per dollar.
And, and I think that's the main problem.
So already 20 years ago we had little robots, baby-like robots, the I-Cub robot, which back then was constructed by a lab in northern Italy, which looks like a baby.
And then it invented its own self-invented experiments, set itself its own goals, try to figure out how the world works and doing that in a hierarchical way and so on.
But after having executed three self-invented experiments, some
tendon in something I was broken, you know, and a technician had to come and fix it,
so it was just so expensive to do all of that. And what we really need is hardware that can compete with
human hands, but there's no human design tech that can compete with these hands. These hands,
they have millions of millions of sensors and all kinds of cables connecting the sensors to the control center,
and I wouldn't know where to put all the cables, and if I cut it, it starts healing itself.
So my hand and yours, that's super advanced technology.
It's really completely beyond what humans can build with traditional robot technology.
And so there we still have so much.
There's such a long way to go there.
I'm still hoping that it's going to happen within my lifetime, you know,
to make true the prediction that I made in the 70s when I was a teenager.
But clearly, the hardware evolution is much lower than the GPU evolution.
Right.
Looks like it's going to take a lot more time on that front.
Now, question for you, who do you think is going to capture the most value as the
technology continues to improve. I think this is a quote attributed to you. You said it's not a few
big companies that are going to dominate everything. The great profiteer of AI is going to be
the little man. But it does seem like if you at least look at where things are heading right now,
you've had these big labs that are going to have trillion dollar valuations and the big tech
companies surrounding them. Looks like they're capturing most of the
economic value and the quote unquote little men is about as uncertain as they've ever been.
So what do you think on that front?
Yeah, I think it's just a reflection of the current bubble that we have in certain aspects
of the economy.
Because if you look at all these zombie unicorns in Silicon Valley now, you know, there are
lots of companies that have officially a billion dollar plus evaluation.
But if they were on the stock market, they would probably be worth just 20 million or something like that.
So zombie unicorns they are called.
And if you look at the most visible companies like Anthropic and OpenEI,
you know, they are spending so much money on all that stuff.
And if you look at Google and Microsoft, you look at the cap egg spending.
and this directly affects their cash flow.
Apparently it doesn't affect the price earnings ratios after these companies,
but it should because there was a time when, you know,
Google and Microsoft they had on the order of $100 billion free cash flow,
but now it's down to $20 billion, $10 billion.
Some of the companies suddenly have minus $10 billion cash flow.
They take on debt to finance all these GPUs in the data centers.
And so these companies are becoming more like utilities, right?
Because suddenly these formerly nimble software companies,
who had maybe a small team of 10 people to improve some shitty operating system
and roll it out for billions of people who all use their own computers,
their own smartphones to run the operating system,
suddenly the same companies, they have to think about buying gas turbines
and investing in nuclear power plans,
and taking on debt and doing all the things
that electricity companies do, you know, utility companies.
And now take into account that every five years, it's still true,
every five years, computer is getting ten times cheaper.
Now, if you invest $1,000 billion today into GPUs for data centers,
this means that within five years,
you are going to lose $900 billion.
Somebody is going to lose $900 billion in the near future
because there is no business model.
Nobody has a business model to recuperate all these losses.
And none of the companies have a mode, you know,
because whenever there's a new benchmark breaking record or something,
benchmark record breaking language model that does this,
or this or whatever.
A few months later, there's the same thing in open source.
And so there's so much pressure to keep the prices down,
and none of these companies has a moment.
So let's wait a little bit, you know,
until the forced ETF buying of these trillion dollar companies
is over.
At the moment, we have trillion dollar or even more companies
appearing on the stock market and the index funds
have to buy them suddenly,
because the NASDAQ and others, they change their rules.
Normally, you wait for a year or so until the stock market finds out by itself,
where it's the true value of this company.
But at the moment, this is not done.
So suddenly the ETFs and the pension funds and everybody, they have to buy that.
Once that is over, let's see what happens when the founders try to sell their stakes.
So it would be astonishing if we wouldn't see some.
huge fluctuations in the stock market prices.
Right.
So where's your vision of where the little guy can end up?
The little guy is, in the long run, the little guy is going to profit.
So at the moment, the big companies, and everybody says, oh, the big companies, they are
profiting like crazy.
What is that true?
Look at their cash flows.
No, they go down like that.
Look at their debt ratios.
They have stopped to repurchasing their own shares because they don't have enough cash any longer.
And so at the same time, what the little guy has to do is just wait a little bit, you know,
because every five years computer is getting ten times cheaper, which means in ten years
you can buy the same thing for one percent of the price.
And it's going to be just like with smartphones, you know.
And I'm often relating the story of the rich guy whom I knew in the 80s.
And he had a Porsche.
He was rich.
He had a Porsche.
But the most amazing thing was in the Porsche, there was a mobile phone.
He could pick up the receiver and talk via satellite to another guy with a Porsche like that.
And today, 40 years later, everybody in developing countries has a smartphone that is immensely more powerful than what he had in his Porsche.
So the little guy is paying almost nothing for whatever he had there, which was really expensive.
And the same thing is going to be true with AI.
And soon AI will not be in the cloud, you know.
No, it will be local on your local computer, a small little computer.
It will be as powerful as what's now in the cloud,
but you won't have to connect it to the Internet,
which is always worrisome thing, you know, and who knows,
who out there is just waiting for you to connect.
And then even the poor guys will have many AIs,
physical AIs, I think, but at the moment, especially software AIs, that are going to make his life
longer and easier and healthier. And he will own them. He won't have to pay money to other guys.
I like that future. Okay, let's take a quick break. And on the other side of this, talk a little
bit about AI theory of the mind and body and whether AIs can feel pain and whether we have free will.
So a lot of that stuff is coming up right after this.
I want to tell you about a documentary I've made with gravity
to explore the future of AI agent security.
To find out if we're truly ready for autonomous agents,
I sat down with MIT Professor Ramesh Roscah,
former White House CIO Teresa Payton,
Michelin's Group Chief Data and AI Officer Ambica Roger Gopal,
and Sharon Guy, a former executive at Alibaba.
They each offer unique insights into this evolving landscape.
We conclude with Rory Bluntz,
Condell, CEO of Gravity, to discuss the path forward.
With Gravity leading the way, join us on this journey.
You can watch the full documentary at the link in the show notes.
Hey, Ontario.
Come on down to BetMGM Casino and see what our newest exclusive,
The Price's Right Fortune Pick, has to offer.
Don't miss out.
Play exciting casino games based on the iconic game show only at BetMGM.
Check out how we've reimagined three of the show's iconic games, like Plinko, Clifhanger,
and The Big Wheel, into fun casino game features.
Don't forget to download the BetMGM Casino app for exclusive access and excitement on the Price's Right Fortune Pick.
Pull up a seat and experience the Price's Right Fortune Pick, only available at BetMGM Casino.
BetMGM and GameSense remind you to play responsibly.
19 plus to wager.
On only.
Please play responsibly.
If you have any questions or concerns about your gambling or someone close to you,
please contact Connects Ontario at 1866-531-2-6-00 to speak to an advisor free of charge.
BetMGM operates pursuant to an operating agreement with Eye Gaming Ontario.
Hey, Ontario, come on down to BetMGM Casino and see what our newest exclusive,
the Price's Right Fortune Pick, has to offer.
Don't miss out.
Play exciting casino games based on the iconic game show only at BetMGM.
Check out how we've reimagined three of the show's iconic games,
like Plinko, Clifhanger, and The Big Wheel, into fun casino game features.
Don't forget to download the BetMGM Casino app for exclusive access and excitement on the Price's Right Fortune Pick.
Pull up a seat and experience the Price's Right Fortune Pick, only available at BetMGM Casino.
BetMGM and GameSense remind you to play responsibly.
19 plus to wager.
ON only.
Please play responsibly.
If you have any questions or concerns about your gambling or someone close to you,
please contact Connects Ontario at 1866-531-2-6-00 to speak to an advisor free of charge.
BetMGM operates pursuant to an operating agreement with Eye Gaming, Ontario.
We're back here on a big technology podcast with Professor Juergen Schmidt-Huber.
Professor, let me ask you this about AIs, because I think you sort of suggested this in our earlier
in the beginning of our conversation that AIs might be able to feel pain.
Maybe if they're not able to accomplish the goal that we set out for them,
or maybe that can actually be baked into the experience of the model.
I'm curious if you think that that is far off that an AI can experience.
something akin to what a human feels with pain
when they, for instance, are set out on a task
and don't accomplish it the way that they've been instructed?
I always think it's funny that many people claim,
even computer scientists claim,
that AIs cannot feel pain
because our AIs have felt pain at least since 1990.
So what do we do when we build an agentic AI,
an artificial neural network that produces actions that change the world and the new inputs are coming in from the environment and again there's an action and so on and the agent the neural network remembers what happened before and is trying to figure out how to behave such that the sum of
of all reward signals is maximized and the sum of all pain signals is mixed and the sum of all pain signals is mixed
minimized. What about these pain signals? The pain signals are the most natural thing.
Because whenever we have a robot, we give it pain sensors. Why? Because it's a learning robot,
and so this learning robot needs some sort of motivation to learn to protect itself.
So whenever the robot bumps against a napsicle, then, you know, the corresponding pain
sensors wake up and negative numbers are special inputs to the
to the robot brain and the robot sees then these incoming negative numbers and it is wired
to avoid that. So it's trying to learn to generate action sequences that avoid the pain. And it's trying
to learn to generate action sequences that lead to the rewarding events. Maybe the robot has
a charging station somewhere. And whenever the battery is low,
the negative numbers are coming from the battery, hunger, pain, hunger.
And then the robot's goal is to reach the charging station
without bumping into obstacles and sit down there and, you know,
and enjoy the pleasure, just positive numbers,
as the battery is being recharged.
So the most natural thing is to give these robots
simple emotions like that. Now, the emotions and the pain, they are just crucial ingredients of the learning
process because the learning is about generating better behavior for the robot. So the robot has to
know what's good for it and what's not good for it. So the pain signals are just informing the robot
about what should be avoided. Pain is just an invention of nature in Bala.
beings of evolution, which invented this pain sensor thing for animals such that they have an
incentive to learn to avoid the pain. And we have done that for many, many decades in our
learning machines. Now what's happening is then that these learning machines they have
world models predicting the future, not only the immediate future,
You know, not only the immediate pain signal that I'm going to feel right now when I touch with my hand the oven or something.
No, they also try to predict the sum of all these future pain signals.
That's what the reinforcement learning machines do.
And so they look ahead into the future.
And so there's immediately this kind of secondary emotion, which is not just about the...
about the current moment, which is looking ahead.
For example, maybe there's a bad man
which sometimes comes into the room
and knocks the little robot on the head.
So over time, the robot is going to learn that,
if it has a reinforcement learning machinery on board
and a wild model that predicts and learns to predict,
and then over time it will be able to distinguish the bad man
from the friendly man,
and then whenever the bad man appears again, I'm there and it does face recognition,
then it predicts that very soon it's going to feel pain if it doesn't hide itself behind the curtain.
So it will rapidly try to hide itself behind the curtain.
Now you as an outside observer will say, look, the little robot is afraid.
It has the emotion of being afraid.
But it's just the most trivial side effect of traditional machine learning.
You know? So, yes, we have all kinds of emotions in our robots already and have had them for many decades.
And then we also have these high-level emotions, which are not just about the current moment.
No, they're looking ahead, you know.
And then, of course, if you bring several robots or agents like that together,
you immediately get stuff that is reminiscent of, you know, of liking other robots or
of loving them maybe, because if you give them a task or a set of tasks that they can collectively
solve, but one of them alone cannot solve them, then they will have to learn to work together.
And of course, suddenly, each of them has an incentive to help the other one.
And each of them has an incentive to help the other one, especially when the other guy is sick
or something who has a problem.
to help him, to care for him,
and the extreme form of that,
a human might call lull.
And in a society of robots or agents like that,
this is just a natural byproduct of the individual egoism
of all these little guys who all want to minimize
some of their pain sensors signals
and maximizes some of their pleasure signals.
So altruism, what you call altruism,
is just a natural consequence of the egoism or the learning agents.
So I think the counterargument would be that it can't be pain like the way that humans feel pain
because humans are conscious, humans have a nervous system that sends physical pain signals
to the brain and registers as true pain, whereas robots are, the argument would be robots are
unfeeling, unconscious, and effectively not that different from a calculator.
How would you respond to that?
Yeah.
So how do you evaluate whether someone feels pain or is afraid by looking at its behavior?
And if it looks like a sweet little robot cat or something and it has learned to hide behind
the curtain whenever the bad man comes in and tries to knock it.
then you will say it's obvious this little robot cat is afraid has fear the emotion of fear
and what else do you want to do so there's no way of objectively seeing the difference between
the emotions of a reward maximizing biological brain and the emotions of a reward maximizing
artificial brain.
But we do have, for instance, you talked about love.
We have chemicals that are associated with the feeling of love like oxytocin.
So the robots don't have those chemicals.
And hence the argument would be that it is a completely different feeling that you
could never compare to love.
What do you think?
So, of course, the chemicals that are used in brains to encode certain states of
fear or of love and whatever. They are different from what we are using in our artificial brains.
But the principles must be the same, right? Because it's about achieving intelligent behavior.
What does that mean? It means finding better ways of achieving your goals. What does that mean?
The main goal in your life until the end of your life is to avoid these pain signals and hunger signals,
such that you eat three times a day, and get these rewarding signals during moments where you are trying to reproduce yourself, for example,
and a handful of objectives encoded in a utility function, which was invented by biological evolution for the animals and for your self.
and which in very similar form, we are encoding in our learning agents to give them the same incentive to solve problems better, to make them better maximize their own rewards and minimize their pain.
Yep. And then consciousness is an interesting one as well. There are, and I spoke about this recently with Professor Jeff Hinton, there are those that say,
actually the prevailing view around AI is that AIs like LLMs are just stochastic parrots,
they're statistics machines.
You could effectively, you know, run this, these algorithms.
They're effectively, if not entirely interpretable, mostly explainable,
and they're statistical prediction machines, hence not conscious.
But you've been arguing that they're conscious for quite some time.
So tell me how you come to that.
Yes, so the large language models that everybody is using today, which are basically about predicting parts of text from other parts of text, for example, predict the next token given the past, they are too simple for what I consider the, they don't carry in themselves the main reason for developing something that people might call consciousness.
Why do they seem conscious?
Well, because they have read everything about consciousness.
All the books ever written about love and consciousness and pain and conflicts
and everything that is important to humans, which was put on the World Wide Web,
they have read that, which means that as you are interacting with them in a chat,
they are very prone to repeat, very convincing sentences that include the word consciousness
in a way that is convincing, you know, because they have read so many literature-price-winning
novels about consciousness and other novels such that they can do a very convincing job there,
and they will tell you a lot about consciousness, which you maybe even didn't know.
However, they don't have their own motive to develop self-consciousness in the following sense.
Let's think back again of this two network system, where one is the controller that is generating the actions,
where the actions then lead to new inputs from the environment, because if you move your hand like this, then the video changes and so on,
and the other network, which allows just to predict these changes, the one model.
So you need the world model to plan your future through mental simulations without, you know, really executing all these action sequences in the world world, which would be very expensive.
So that's the motivation for this world model.
Now, where does now consciousness stuff come in?
Nobody has a universally accepted definition of consciousness, but let me now show you something very simple, which,
which is super compatible with what lots of people think about when they hear the word consciousness
and self-awareness. Now, let's look at this world model which is being used to plan the future
of some agent, which is using the world model for mental experiments. Now, the wall model is a
deep neural network which has learned to encode everything that it has seen efficiently in a bunch of neurons,
For example, everything that frequently appears in the environment gets internal abstract representations that stand for, you know, a prototype of that concept.
For example, in a world where there are lots of different faces of different humans, then you will find units, internal units in these artificial neural networks that correspond to prototype faces and some of them are more specific for certain faces and less specific to other faces.
and so on. And so you, in a world where you have lots of glasses, you will have glass detectors, internal representations, internal hidden units that learn to respond and encode glasses and whatever.
And then let's now look at the planning procedure. Now the controller, the controller is trying to figure out a plan, how should I act in the future to maximize my reward and minimize my pain.
And then it uses the world model for a model for a,
mental simulation of its um if of different possible action sequences that it could execute and it's
going to pick the one that leads to the most predicted reward and the least predicted pain so as as it is
doing that it is waking all it is waking up all kinds of hidden units in the one model that um
stand for you know whatever is relevant to the problem for your faces or or glasses or whatever and there's one
thing. There's one thing that is always active when the agent is active, which is the agent itself.
So of course, all kinds of these internal units are going to represent the agent or aspects of the agent and the hand of the agent if it has one,
or the wheels of the agent if it has one and so on.
And so whenever the agent is making plans like that, it's thinking about itself, waking up these internal representations of itself.
And then it's self-aware.
It's self-aware.
In this sense that, for example, if it looks in the mirror,
it will quickly figure out, oh, the guy in the mirror,
I can control what this guy does.
If I do this, then he will do the symmetric thing,
because I have it under control.
But if you are there on the other side of the room,
and I do this and you will do something that's completely unrelated,
I cannot predict them.
So there you already see this concept of agency,
which is immediately recognized as self-agency or the agency of another guy.
And now, the one more thing, which also goes back to 1991,
is this tendency of conscious things becoming subconscious.
So many people are aware of all kinds of things kind of vaguely without thinking much about that
because it's part of an automated process.
As you are driving,
always the same way from your home to your workplace.
Much of what happens is very predictable.
And so you don't even think much about the driving,
and maybe you're thinking about other things.
At the same time, your attention, your internal consciousness
is somehow focusing on other things.
Many people report that.
And this is also,
one of the most natural things.
In 1991, I had a system consisting of two networks,
one of I call the conscious problem solver,
a neural network that just learn to try to learn
to predict certain aspects of incoming data,
which it was not yet able to predict.
Trying to find a regularity is,
and it had a problem to solve.
And so you could say it was conscious in a sense
that there it had to learn something.
And then another network which basically learned to imitate all the solutions that the guy on the higher level found by just imitating the hidden units of the guy on the higher level.
Today it's called distillation of the behavior of one guy into another.
And then the lower level guy is the automatizer because it automates the stuff that's the behavior.
the higher level guy finds, discovers.
A higher level guy is still unsure and still working on creating insights.
And, you know, when these insights come and when it learns something,
then it gets distilled down into this automatic automatizer thing.
So both these aspects of consciousness, where you have, on the one hand,
self-awareness in a world model that is not only predicting aspects of the world,
but also of the agent that is interacting with the world.
and which leads to self-awareness of that kind,
and this difference between the conscious stuff,
so the internal consciousness,
which pays attention to certain aspects of the internal state,
but not to others,
which focuses on what's still unsolved,
where I still have to find a solution,
and separates that from the stuff that is already solved.
So I think both of these aspects are there in these old systems,
And in 2016, I believe, I had an interview with a magazine, which then said,
Schmidt, who claims that AI became conscious in 1991, something like that.
And this was exactly about that.
There was 10 years ago this interview there, but it was actually referring to stuff that is much older,
in 1991, conscious simple systems back then, not as impressive as to do.
days, as humans are obviously, because back then compute was 10 million times more expensive than today.
And we just had tiny little experiments, you know, with this conscious chunk, as I called it,
and the subconscious automatizer and the world models for planning and so on.
And we just had a few hundred weights in our systems while you and you in your brain,
you have trillions of connections which can hold a much, you know, larger sort of consciousness.
But I think the principles are exactly the same.
Yeah, and so that sort of brings me to this question, which is, as the machines get closer to the human brain,
does it change the way we're going to think about what it means to be human?
I think it will change what many people think about humans.
I guess it won't change much what I think about humans,
because it's more or less what I said many decades ago.
But yes, there are many people who claim that AIs can't have emotions and stuff like that,
they do that decades after the fact.
And they are going to change their minds, I'm pretty sure.
And often it's just a matter of direct experience.
So maybe you have heard of these little sweet, cute robot seals that you have in certain healthcare centers.
And the people who are interacting with these furry artificial beings, they get really emotionally attached to them.
They really like them and they play around with them, although they are not smart at all.
They don't learn much.
and if even such a simple robot can invoke feelings of, you know, almost love or something,
then you can imagine what will happen once you have really convincing very sweet little robots
that are more like little animals, except that they can do maybe a couple of things that these
traditional biological little animals cannot do.
No, I think in the 2017 Bloomberg article, 2018 Bloomberg article that I referenced,
you are asked whether you think, you know, we are living in some form of simulation.
You said, that's what I think because it's the simplest explanation of everything
that humankind is programmed to chase progress and will keep making more powerful computers
until we make ourselves obsolete or decide to merge with the smart machines.
Here's the quote.
Either you become something that's really, really different from a human or you stay as a human
or you stay as a human for nostalgic reasons.
but then you will not be a major decision-maker.
You will not play a role in shaping the world.
Can you expand upon that and tell us
whether your opinion has been reinforced or changed
in the interceding years?
No, my opinion has not been changed.
You know, back then, and actually for many decades,
there has been talk about uploading human brains or souls,
if you will, into computers
and then live your, your,
future life in some sort of simulation, simulated paradise, or in a robot that interacts with
the real world or so. And I believe the first story, science fiction story of that kind, was
published in 1964, where someone was able to upload his mind in a computer, back then with
rotating tapes and everything. That was, what was the name? I called Simulacron 3, something like that.
it was by Daniel F. Galui.
And so there is no physical reason to reject the notion that this might be possible.
At the moment it's not possible, except for certain kinds of very simple animals like fly.
So apparently, as of 2024, you can argue that a fly brain has been uploaded.
Maybe it's not the full fly brain with all the learning algorithms for the neurons which are in there.
But in simulation, the simulated fly now does stuff that is very much like what the real fly did before it was uploaded,
before all the connections were read and uploaded into this computer, virtual environment where it then kind of lived on.
So there is no obvious reason to believe that it's impossible to take, you know, to read,
all the synapses of a human brain and understand also the learning algorithms that are changing
the synapses all the time, and then replicate that in a computer, and then the idea would be
that your soul or your mind is uploaded and living there forever in this simulated universe,
which may have contact to the real universe. Now, of course, once, if he, if he, if he, if he,
accept this premise and then we think what is the next step.
Now, suppose you are uploaded there and suddenly you have the opportunity to, you know, have more than two eyes,
maybe have a million eyes, you know, and satellite eyes all around the planet, and, you know,
maybe you have a much bigger brain, not just maybe 10 to the 18 instructions per second,
something like that, but 10 to the 10 times as much.
And now there are two things you can do.
Either you succumb to the temptations of this new life,
and in the process you are going to become something very, very different.
You are not going to remain a lot like you were,
because suddenly everything expands in a way,
and sometimes you might remember your roots as a,
human somewhere, but your future, life is totally detached and disconnected and probably is
highly influenced by other expanded minds like that, you know, discussing problems that you would
have never discussed as a human person in this reality. Now, the alternative is, for nostalgia,
You say, I don't want to succumb to these temptations.
I keep my two eyes and I keep my little brain and I don't increase it by a factor of 10 billion or something.
But then your competitors, your competitors will be ignoring you, you know,
because they will have so many new skills that you don't have.
and the main decision makers, they are going to be these expanded minds,
and they themselves, they will be in competition with the native AIs,
which, you know, are not, they don't have this evolutionary ballast
and maybe much more adapted to the needs of the future,
you know, once the AISphere is expanding from our biosphere into space
and who knows what they are going to do that.
So either you are nostalgic and remain irrelevant,
or you become part of this growing sociology, if you will,
of AIs, which are mostly AIs, and some of them may have human roots and whatever.
But almost all the decision-making process,
process and the universe-shaping things, they are going to be done by these new beings and
not by the human-like beings.
Would you upload your brain and merge with AI?
I haven't given too much thought about that because it's not really a goal of mine,
because I think exactly for the reasons that I just...
described that my current self is not going to make a difference there.
And the alternative is going to, you know, the super AI that is going to go 10 to the 20 times
beyond the little thing that I have at my disposal, that is going to occur anyway.
and not just one of them, but many, many different beings like that.
And so I think it's not really going to make a difference.
Can I ask one last question?
Yeah, of course.
So we are fastly moving towards a world where much of intelligence is encoded in machines,
and those machines can predict and take action based off of their conception of where things are heading.
Do you believe then in the concept of free will right now and in the future?
Will there be a free will as so much of our reality is intermediated by these machines
that have a pretty good idea of where things are heading?
So in 1997 I wrote this paper, a computer scientist's view of life, the universe and everything.
And basically, this was about trying to explain.
trying to find the simplest explanation of our universe.
So the holy grail of physics would be find the shortest deterministic program that
computes everything that we have ever observed in this universe, including the seemingly random
quantum events, spin up and down measurements and everything.
And we don't know this shortest description of the history.
of the universe, but as scientists we are trying to find better and better and more and more compact descriptions.
And we have made a lot of progress as scientists, as physicists, towards that goal.
And then I realized in a couple of years before 1997 that although we don't know the shortest algorithm that computes just this universe in which we are living,
we at least know the super short algorithm which computes all parts of the universe which computers all parts of the universe,
computer universes, which is basically the program that systematically enumerates all possible programs, and then there's an optimal way of allocating runtime to them, and every possible computable universe is going to be computed, including ours if it is computable. And there is no physical evidence against the possibility that our universe is computable.
So I'm not talking just about universes with a probability distribution of the next possible things is computable.
which is what my former postdoc, Marcus Huta,
did around 2000 when he developed the AIXE model,
the optimal universal decision maker.
Now, really the deterministically computable universe history is.
And so there is no evidence.
It's again important to realize that
that we are being created by this method,
and there are many programs that compute us,
but it turns out,
that they are dominated, that they are dominated by the shortest programs that compute all of
all of this universe, including our conversation here, which a guy like me believes is something
totally deterministic. And so as you are reacting to my, you know, my voice signals,
wave fronts coming out of my mouth, and you are responding then with your own wavefronts,
This seems like a lot of free will, you know, as we are discussing with each other.
However, the deterministic universe view would say this is all deterministic.
It may seem like a free will thing to us, but it isn't.
And it's interesting to realize that even in very simple simulations,
deuteristic simulations, that we already can do on our little man-made computer,
which are just part of this huge simulated universe,
even then we can observe similar effects
because we can really devise deterministic environments
for deterministic neural networks that make decisions
as they are interacting with other animals,
other simulated animals,
and they are trying to maximize their reward,
and then they are making decisions,
just learning to better react to what the other animals are doing.
So all of that looks like a lot of that looks like,
a lot of free will at first glance, but I can completely rerun every little detail of this
simulation. So although it looks like a free will decision thing, it's just something that is part
of deterministic universe. And therefore, it seems clear to me that the concept of free will is overrated.
If free will is overrated, what's the point of living?
since everything is effectively laid out before you go through.
They made a movie about that and it's called Free Willy.
About the whale?
Yeah.
Wait, so sorry, how does that connect to the question?
It doesn't connect it all.
It was just something that I made up.
All right, Professor, great speaking with you, long time coming.
Hope we get to do it again sometime soon.
Alex, it was my pleasure.
Thank you.
Thank you, everybody, for listening and watching.
see you next time on Big Technology Podcast.
