Judea Pearl: Human-Level AI and the Test of Free Will | AI Podcast Clips
Watch on YouTubeVideo summary
Judea Pearl, a renowned figure in artificial intelligence research, expresses his lifelong drive to create human-level intelligence through an incremental approach that moves from established knowledge to new steps without necessarily visualizing the final destination immediately. However, he clarifies that this end goal is not merely about machines answering sophisticated questions or handling counterfactuals; rather, it involves creating systems capable of great compassion and possessing a sense of responsibility and free will. Pearl emphasizes that current AI lacks these qualities because robots cannot yet communicate with one another using reward and punishment mechanisms in the way humans do. He illustrates this concept through an analogy of soccer players who can critique each other's performance—such as saying, "You didn't pass the ball at the right time"—and face consequences like sitting on the bench for two games. This ability to engage in honest natural language communication about actions and their outcomes is crucial because it allows a coach to convey complex knowledge that directly tweaks an agent's software modules, leading to improved performance over time. The transcript highlights a significant distinction between current AI capabilities and true intelligence regarding value alignment and ethics. Pearl argues that for machines to align with human values, morals, and ethics, they must possess cause-and-effect reasoning rather than just statistical correlation or simple communication skills. He posits that empathy is essential for ethical decision-making, which requires an agent to build a model of the recipient (the user) as if suffering their pain were its own. Pearl suggests that this process is not overly difficult because humans naturally map others onto themselves; we recognize similarities and say, "You are like me," thereby avoiding harm. For machines to achieve similar ethical standards, they must essentially learn to fake being human by constructing a model of the user based on self-knowledge, effectively defending their own existence within that framework. A critical component in achieving this level of sophistication is the development of consciousness and free will through self-modeling. Pearl explains that an entity gains a sense of itself when it builds a blueprint of its own abilities and desires by looking at itself as part of the environment rather than just being embedded within it. This internal representation allows for modification; much like looking in a mirror to see how tweaking specific parts changes performance, having this self-model enables an agent to understand that changing certain variables will lead to different outcomes. Pearl acknowledges limitations regarding the halting problem but asserts that possessing even a partial blueprint of oneself is sufficient to define free will and consciousness at a functional level. Ultimately, the discussion concludes that while current AI can communicate effectively in natural language, true human-level intelligence requires more than just data processing; it demands an internal architecture capable of self-reflection and ethical reasoning. The ability for robots to reason about reward and punishment among themselves is presented as a foundational step toward playing "better soccer," which serves as a metaphor for complex social interaction and cooperation. By integrating cause-and-effect thinking with the capacity to model others based on oneself, machines could eventually bridge the gap between computational logic and human-like compassion. Pearl's vision suggests that without these specific cognitive structures—specifically the ability to simulate selfhood and empathize through shared models of pain and desire—the creation of truly responsible AI remains out of reach despite rapid advancements in communication technologies.
Read the full video transcript
I know you're not a futurist but are you
excited have you when you look back in
your life long for the idea of creating
a human level intelligence yeah I'm
driven by that all my life I'm doing
just by one thing but I go slowly
I go from what I know to the next step
incrementally so without imagining what
the end goal looks like do you imagine
what the end goal is gonna be a machine
that can answer sophisticated questions
counterfactuals of a great compassion
[Music]
responsibility and free will so what is
a good test there's a touring test a
reasonably free will doesn't exist yet
how would you test free well and that's
so far we know only one thing meaning if
robots can communicate with reward and
Punishment among themselves hitting each
other on the wrist and say you shouldn't
have done it
okay playing better soccer because you
can do that what do you mean because
they can do that because it can
communicate among this because of the
communication they can do because of the
communicate like us reward and
punishment yes you didn't pass the ball
the right the right time and so for
therefore you're gonna sit on the bench
for the next two if they start
communicating like that the question is
will they play a better soccer it
supposed to work as if what what they do
now without this ability to reason about
reward and Punishment responsibility and
it's fine I can only think about
communication communications and in an
honest a natural language but just
communication just communication and
that's important to have a quick and
effective means of communicating
knowledge if the coach tells you should
have passed the ball pink he conveys so
much knowledge to
supposed to would go down and change
your software that's the alternative but
the coach doesn't know your software so
how can it coach tell you you should
have passed the ball but that our
language is very effective if you just
pass the ball you know your software you
tweak the right module and next time you
don't do it now that's for playing
soccer and the rules are well-defined
well not well defined when you should
pass the ball is not what the fuck'd
know it's very soft there is no Z yes
it's art but in terms of aligning values
between computers and humans do you
think this cause and effect type of
thinking is important to align the
values values morals ethics under which
the machines make decisions is is the
cause-effect where the two can come
together qualification is necessary
component to build a ethical machine
because the machine has to empathize to
understand what's good for you to build
a model of use of you as a recipient we
should be very much but what is
compassion they imagine it you suffer
pain as much as me as much as I do have
already a model of myself right so it's
very easy for me to map you to mine I
don't have to rebuild the model it's
much easier to say oh you are like me
okay therefore I would not hate you and
the machine has to imagine it has to try
to fake to be human essentially so you
can imagine that you're they you're like
me right
moreover with me let's defend it that's
consciousness they have a model of
yourself where do you get this model you
look at yourself as if you are a part of
the environment if you build a model of
yourself versus the environment then you
can say I need to have a model of myself
I have abilities I have desires and so
forth okay I have a blue
print of myself though not the full
detail because I cannot get the halting
problem right but I have a blueprint so
that level of a blueprint I can modify
things I can look at myself in the
mirror and say hmm if I change this much
tweak this one I'm gonna perform
differently that is what we mean by free
will and consciousness gorgeous
you