Essentials: The Science of Learning & Speaking Languages | Dr. Eddie Chang
33m 4s
Dr. Eddie Chang discusses the neurobiology of speech and language, distinguishing speech—controlled by motor actions in the vocal tract—from language, which involves understanding meaning, grammar, and context. He explains how the larynx generates voice through vocal fold vibrations, and how speech production relies on precise coordination of multiple structures. Chang highlights that non-linguistic vocalizations stem from distinct neural pathways, separate from language centers. His work on brain-machine interfaces (BCIs) enables paralyzed individuals with locked-in syndrome to communicate by decoding neural signals associated with speech intentions. A clinical trial involving a patient paralyzed for 15 years demonstrated successful translation of brain activity into words using machine learning and a small vocabulary, which is now expanding. The technology is being enhanced with features like autocorrect and animated avatars that mimic facial expressions and speech, improving communication realism and user feedback. While such neurotechnologies are advancing rapidly, Chang cautions that broader applications for human enhancement—like super-memory or faster communication—raise significant ethical questions about access, equity, and societal implications. He emphasizes that while augmentation is not new (e.g., through caffeine or nicotine), the invasive nature of current BCIs demands thoughtful evaluation. Ultimately, the focus remains on restoring communication for those with severe disabilities, with future innovations aiming for more natural, embodied interactions that bridge the gap between thought and expression.
Welcome to Huberman Lab Essentials, where we revisit past episodes for the most potent and actionable science-based tools for mental health, physical health, and performance.
I'm Andrew Huberman, and I'm a professor of neurobiology and ophthalmology at Stanford School of Medicine.
And now for my discussion with Dr. Eddie Chang.
Eddie, welcome.
Hi. Hi, Andrew.
Great to be here with you.
Your main focus these days is the neurobiology of speech and language.
So for those that aren't familiar, could you please distinguish for us speech versus language in terms of whether or not different brain areas control them?
When I think about language, I think about words and just talking.
If I sit down to do a long podcast or I think about asking you a question, I don't even think about the words I want to say very much.
I mean, I have to think about them a little bit, one would hope.
But I don't think about individual syllables unless I'm trying to, you know, accent something or it's a word that I have a particular difficulty saying.
is represented to me is perhaps one of the most interesting questions. And I know this lands
square in your wheelhouse. Sure. Let's get into this, Andrew, because this is one of the most
exciting stuff that's happening right now is understanding how the brain processes these
exact questions. And speech corresponds to the communication signal. It corresponds to me moving
my mouth and my vocal tract to generate words. And you're hearing these as an auditory signal.
Language is something much broader. So it refers to what you're extracting from the words that I'm
saying. We call that pragmatics and sort of are you getting the gist of what I'm saying?
There's another aspect of it that we call semantics. Do you understand the meaning
of these words and the sentences? There's another part that we call syntax, which refers to how the
words are assembled in a grammatical form.
So those are all really critical parts of language. And speech is just one form
of language. There's many other forms like sign language, reading. Those are all important
modalities for reading. Our research really focuses on this area that we're calling speech.
Again, the production of this audio signal, which you can't see, but your microphones are picking up.
There are these vibrations in the air that are created,
by my vocal tract that are picked up by the microphone in the case of this recording,
but also picked up by the sensors in your ear. The very tiny vibrations in your ear are picking
that up and translating that into electrical activity. It's such a complex feat. Some people
would say it's the most complex motor thing that we do as a species is speaking, not the extreme
feats of acrobatics or athleticism, but speaking. Especially when one observes,
opera or people who freestyle rappers. And of course, it's not just the lips. It's the tongue.
And you've mentioned two other structures. Pharynx and larynx are the main ones. Can you tell us,
just educate us at a superficial level, what the pharynx and larynx do differentially? Because I
think most people aren't going to be familiar with that.
Okay, sure. I'll talk primarily about the larynx here for a second, which is that
if you think about when we're speaking, really what we're doing is we're shaping the brain.
Okay. Sure. So even before you get to the larynx, you got to start with the expiration. We fill up our lungs
and then we push the air out. That's a normal part of breathing. What is really amazing about speech
and language is that we evolved to take advantage of that normal physiologic thing at a larynx. And
what the larynx does is that when you're exhaling, it brings the vocal folds together. Some people
call them vocal cords. They're not really cords. They're really vocal folds. They're two pieces of
tissue that come together. And a muscle brings them together. And then what happens is when the air comes through the vocal
folds, when they're together, they vibrate at really high frequencies, like a hundred to 200
Hertz. And the reason why men and women generally have different voice qualities is it has to do with
the size of the larynx and the shape of it. Okay. So in general, men have a larger voice box or
larynx and the vibrating frequency, the resonance frequency of the vocal folds when the air comes
through them.
It's about a hundred Hertz for men and about 200 for women. So you take a breath in as the air is
coming out, the vocal folds come together and the air goes through. That creates the sound of the
voice that we call voicing. It's not just your voice characteristic. It's the energy of your
voice. It's coming from the larynx there. It's a noise. And then it's the source of the voice.
And then what happens is that energy, that sound goes up.
Up through the parts of the vocal track, like the pharynx into the oral cavity, which is your mouth
and your tongue and your lips. And what those things are doing is that they're shaping this,
the air in particular ways that create consonants and vowels. That's what I mean by shaping the
breath. It just starts with this exhalation. You generate the voice in the larynx, and then
everything above the larynx is moving around. Just like the way my mouth is doing right now,
to shape that air into particular patterns that you can hear is words.
Immediately makes me wonder about more primitive or non-learned vocalizations like crying or
laughter. Are those produced by the language areas or do they have their own unique neural
structures? We call those vocalizations. A vocalization is basically where someone
can create a sound, like a cry,
or a moan, that kind of sound. And it also involves the exhalation of air. It also involves
some phonation at the level of the larynx where the vocal folds come together to create that
audible sound. But it turns out that those are actually different areas. So people who have
injuries in the speech and language areas oftentimes can still moan. They can still
vocalize. And it is a different part of the brain. I would say an area that even non-human
primates have that can be specialized for vocalization. It's a different form of
communication than words, for example. I'd like to take a quick break and
acknowledge our sponsor, Function. Function provides over 160 advanced lab tests to give
you a clear snapshot of your bodily health. This snapshot gives you insights into your
heart health, your hormone health, autoimmune function, nutrient levels, and much more.
They've also recently added access to advanced MRI and CT scans,
function not only provides testing of over 160 biomarkers key to your physical and mental health.
It also analyzes these results and provides recommendations for improving your health
from top doctors. For example, in a recent test with Function, I learned that some of my blood
lipids were slightly out of range. As a result, I decided to start supplementing with natokinase,
which can naturally help reduce LDL cholesterol. And it did. In a follow-up test, I could confirm
that this strategy worked. My blood lipids are now back exactly where I want them.
Comprehensive lab testing of the sort that Function offers is just so important for health.
I mean, how else are you going to know what's going on under the hood? And while I've been
doing blood work for years, it used to be time consuming, complicated, and expensive. In fact,
I used to spend thousands of dollars per year trying to get this kind of data. And the data,
frankly, were not all that good. But now with Function, it's extremely easy and affordable.
A Function membership is only a dollar a day, $365 a year.
And if you think about the information it provides and the health challenges it helps you
avoid and the proactive things that it can do for you to enhance your health, I truly look at it as
a savings. To learn more, visit functionhealth.com slash Huberman and use the code Huberman for a
$50 credit towards your membership. Again, that's functionhealth.com slash Huberman.
Speaking of storage of and ability to speak, you are doing some amazing work and have achieved
some pretty incredible, well-deserved recognition for your work in bringing language out to the
world of paralyzed people, essentially allowing people who are locked in to a paralyzed state or
otherwise unable to articulate speech using brain machine interface, essentially translating the
neural activity of areas of the brain that would produce speech into hardware, artificial,
non-biological tools in order to allow paralyzed people to communicate.
So there are a series of conditions. They include things like brainstem,
stroke. The brainstem is the part of the brain that connects the cerebrum, which is the top part,
does our thinking and a lot of the motor control, speech, language, everything. And the brainstem is
what connects that to the spinal cord and the nerves that go out to the face and vocal tract.
So if you have a stroke there, you could be thinking all the wild, creative, intelligent
thoughts you have in the mind and the cerebrum, but you can't get them out into words or you can't
get them out to your hand to write them down. So that's a very severe form of paralysis called
brainstem stroke. There's another
kind of conditions that we call neurodegenerative, where the nerve cells die, basically, or atrophy
in a condition called ALS. That's a very severe form of paralysis. In its extreme form, people
essentially lose all voluntary movement. The muscles to their diaphragm and their lungs
essentially give out as well. They get weakness there and then they can't breathe anymore.
In our field, these are kind of like the most devastating things that can happen.
This condition of what we call being locked in refers to,
this idea that you can have completely
intact cognition and awareness but have no way to express that no voluntary movement
no ability to speak and that is devastating because psychologically and socially you know
you're completely isolated that's what we call locked-in syndrome and it's devastating so we've
been studying this patterning of electrical activity for consonants and vowels and essentially
once we figured out a lot of these codes for the individual phonetic elements part of the lab
started to focus on this very specific question for people who have these kind of paralysis could
we intercept those signals from the brain the cerebral cortex as someone is trying to say those
words and then can we intercept them and then have them taken out of the brain through wires
to a computer that are going to interpret those signals
you
to a computer that are going to interpret those signals and translate them into words so we started
a clinical trial it's called the bravo trial it's still underway and the first participant in the
bravo trial was a man who had been paralyzed for 15 years he was in a car accident he actually
walked out of the hospital day after that car accident but the next day had a complication
related to it where he had a very large stroke in the brainstem and that turned out to be devastating
he didn't wake up from that stroke
for about a week he was in a coma for about a week and when he woke up from that coma he realized that
he couldn't speak or move his arms or legs as he told me or communicated to us that was absolutely
devastating he wanted really to die at that time could he blink his eyes or move his mouth in any
way he could blink his eyes he had some limited mouth movements but couldn't produce any intelligible
speech it was like completely slurred and incomprehensible he survived this injury a lot of
people who have that kind of stroke just don't survive the way he actually communicates because
he has a little bit of residual neck movements is that he improvised and had his friends basically
put a stick attached to his baseball cap because he could move his neck he would essentially type out
letters on a keyboard screen to get out words in fact this is how he communicated
was through a device that he would essentially peck out letters one by one by moving
his neck to control this stick attached to his baseball cap he hadn't really spoken for about
15 years oh goodness yeah so it was part of a clinical trial it was you know something that
our hospital and also the fda you know had to approve and looked at very carefully but given
a lot of the work that we had done there were some basis for for why this might work
and so we did a surgery where we implanted electrodes onto these areas that control the vocal tract
the areas that control the larynx the areas that control the lips and tongue and jaw movements
when we normally speak these are areas that presumably may be active that was our hope
and he underwent a surgery a brain surgery we put an electrode array and we connected it to a port
that was sculled to screw to his skull and the port actually goes through
his scalp and he's lived with this now for the last three years
so he has an electrode array that's implanted over the part of this brain that's important for
for speech it's connected to a port and then we connect a wire to that port that translates those
what we call analog you know brain waves and converts them into digital signals
we put them through machine learning or artificial intelligence algorithm
that can pick up these very very subtle patterns you can't actually see them with your eye
in in the brain activity and translate those into words and this is something that took
weeks to train the algorithm to interpret it correctly but what was incredible about it was
to see how he reacted he would be prompted to say a given word like you know outside for example
and then he would think about it try to say it and finally those words would appear
on the screen and what was really amazing about it was you could really tell that he like got a kick
out of that because you know his body was shaken away and his head would shake in a way that he
would start to giggle that was cool to see but then i also realized that when he was
giggling it kind of screwed up the next words decoding is that a bug you've since fixed no
we haven't fixed that it's easier just to tell him to stop giggling the way this worked was we trained
this computer to recognize 50 words we started with a very small vocabulary that's expanding
as we speak i think that this is just a matter of time before these vocabularies become much much
much larger but we started with a 50 set of words we created essentially all the possible sentences
that you could generate from those 50 words why that was important was you can use those
all those possible sentences to create a computational model computer model
of all the different word combinations to give different sentences given those 50 words and then
you can essentially do what we call autocorrect it's the same kind of thing that we do when you're
texting for example you get the wrong letter in there your phone actually knows you know because
it's context what to correct it so because the decoding is not 100 correct all the time in fact
it's far from that it's really helpful to have these other features like autocorrect the stuff
that we use routinely now with texting that makes it correct and then updates it so it's a combination
of a lot of things it's the ai that is translating those brain activity patterns but it's also things
that we've learned from speech and speech technologies that you know you put all together
and then all of a sudden it starts to work that was the first time that someone was paralyzed and
could create words and sentences uh that was just decoded from the brain activity
i'd like to take a quick break and acknowledge our sponsor better help
better help offers professional therapy with a licensed therapist carried out entirely online
i've been doing therapy for a long time and i can tell you that it's a lot like physical workouts
there are times when i want to do it and there are times when i don't want to do it but whenever i
finish a therapy session i come away feeling better and i have specific actionable items
that hadn't occurred to me that when i implement improve various aspects of my life the data tell
us that there are just so many benefits that come through effective therapy and effective
therapy involves having a good rapport with your therapist getting insights from them and with them
and coming away with specific actionable tools that you can apply outside of therapy with better
help they make it very easy to find an expert therapist who can help provide those benefits
that come through effective therapy and the data say it works better help has an average rating of
4.9 out of 5 for its live sessions based on over 1.7 million client reviews also because better
help is done entirely online it's extremely efficient there's no driving to a therapist's
office looking for parking or anything like that you simply log on and do your session
wherever you happen to be if you would like to try better help go to betterhelp.com
huberman to get 10 off your first month again that's betterhelp.com
these days we hear a lot about neural link elon musk's company while brain machine interface
of the sort that you do and that other laboratories do has been going on for a
long time there's been some press around neural link about the promise of what brain machine
interface could do what are your thoughts about manipulating neural circuitry to achieve
supra human or super human or super physiological functions and here we don't even have to think
about neural link in particular it's just but one example of companies and people in laboratories
that are quite understandably considering all this it's a really interesting time right now
the science has been going on for decades the work that we've done in this field that you call brain
machine interface it's been going on for a while and a lot of the early work was just trying to
restore things like arm movement or having people or monkeys control a computer cursor for example
on the screen that's been going on for decades what's been really new is that industry is now
involved and some some of this now becoming commercialized and we're starting to see
us now cross over to this field where it's no longer just research that we're talking about
medical products we're starting to see that we're now crossing over to this field where it's no longer just research that we're talking about medical products
um that are designed to be you know surgically implanted in some cases you know there's people
doing this kind of work non-invasively as well they don't require surgery the specific question
that you're asking about is an area that we call augmentation so can you build a device
that essentially enhances someone's ability beyond super normal super memory
super communication speeds beyond speech for example
superior precision athletic abilities i think that these are very serious kind
of questions to be asking now because as you mentioned the pathway so far is really to
focus on these medical applications i personally don't think that we've thought enough actually
about what these kind of scenarios are going to look like and i don't think we've thought
through all the ethical implications of what this means for augmentation in particular there's part
that is not new at all humans throughout history have been doing things to augment our function
coffee nicotine all kinds of medications that cross over from medical to consumer that is
everywhere so the the pursuit of augmentation or performance
performance, or enhancement is really not a new thing. The questions really, as they relate to
neurotechnologies, for example, have to do with the invasive nature. For example, if these
technologies require surgery, for example, to do something that is not for a medical application.
Again, there, that is not exactly new territory either. People do that routinely for cosmetic
kind of procedures for physical appearance, not necessarily cognitive. So I do think that
provided the technology continues to emerge the way that it does, that it's going to be around
the corner. And it probably is not going to be in ways that are super obvious. I don't think it's
going to be like, can we easily memorize every fact in the world, but in forms that are going
to be much more incremental and maybe more subtle. In many ways, we already have that now.
For example,
you don't have to have a neural interface embedded in your brain to get information,
essentially access to all information in the world. You just have to have your iPhone.
Whether you could do it faster through a brain interface, I definitely wouldn't rule that out.
But think about this, that the systems that we have already to speak and to communicate have
evolved over thousands and millions of years. And they're supported by neural structures
that have bandwidth of millions.
There's no technology that exists right now that people are thinking about that are in commercial
form, certainly not even in research labs that come anywhere close to what has been evolved for
those natural purposes. So I'm essentially saying two sides of this, which is we're already getting
into this now. This is not new territory. This topic of augmentation, both physical and cognitive,
we've already surpassed that.
It's part of what humans do in general, but we are entering this area of enhanced cognition,
these areas that I think the technology is going to be the rate limiting step and how
far we can go. And we have not had the full conversations about number one, is this what
we actually want? Is this going to be good for society? Who gets access to this technology?
These are all things that are going to become real world problems.
Could you tell us what you're doing in terms of merging the brain machine interface
with extraction of speech signals from people who are locked in like poncho with facial expressions?
Sure. Yeah. I'm here with you in person. We could have done this virtually, probably. It's pretty
easy to do that. We could have recorded this really separate, but there is something about
being able to actually see your expressions and to understand other forms of communication.
So another really important one is nonverbal expressions that you're making. For example,
if you have a quizzical look on your face, if I'm looking at you, if I'm looking at your face, if I'm
saying something not clear, that's a sign to me that I need to rephrase it or to say it in a
different way or to slow down. Facial expressions actually are a really important part of the way
we speak. And there's two things. It's not just the expressions of how you're feeling and perceiving
what I'm saying, but it's also seeing my mouth move. In your eyes, I actually see my mouth move
and my jaw move in a particular way that actually allows you to hear those sounds better. So having
both the visual information, but also the sounds go into your brain is going to improve intelligibly,
also make it more natural. And the reason why we're also very interested in this idea of not
just having text on a screen, but essentially a fully computer-animated face, like an avatar of
the person's speech movements and their facial expressions is going to be a more complete
form of expression. Now, you're going to want to make sure that when you're speaking, you're
going to be able to hear what the person is saying. So you're going to want to be able to see what the
person is saying. So you're going to want to make sure that when you're speaking, you're going to
be able to hear what the person is saying. So you're going to want to
be able to hear what the person is saying. But I think the way things are going in the next couple of years, a lot more of our social
interactions, more than even now, are going to move into this digital virtual space. Of course,
most people are thinking about what that means for most consumers, but it also has really important
implications for people who are disabled, right? And how are they going to participate in that?
And so we're thinking really about for people like Pancho and other people who are paralyzed,
what other forms of BCI can we do in order to help improve their ability to communicate?
So one is essentially building out more holistic avatars, you know, things that can
essentially decode, you know, essentially their expressions or the movements associated with their
mouth and jaw when they actually speak to improve that communication.
So do you envision a time not too long from now where instead of tweeting out something in text,
my avatar will, I'll type it out, but my avatar will just say it. It'll be an image of my avatar
saying whatever it is I happen to be tweeting at that moment.
That's what we're working on. That is going to happen and it's going to happen soon. And
there's a lot of progress in that. And again, we're just trying to enrich the field of communication
expression to make it more normal. And we actually think that having that kind of avatar is a way of
getting feedback to people learning how to speak through a speech neuroprosthetic. That's the device
that we call it. It's a speech neuroprosthetic. That is going to be the way that can help people
learn how to do it the quickest, not necessarily like trying to say words and having it come on a
screen, but actually have people embody, feel like it's part of themselves or that they are directly
controlling that, that speech.
That illustration or animation.
As many of you know, I've been taking AG1 for nearly 15 years now. I discovered it way back in
2012, long before I had a podcast and I've been taking it every day since. AG1 is to my knowledge,
the highest quality and most comprehensive of the foundational nutritional supplements on the
market. It combines vitamins, minerals, prebiotics, probiotics, and adaptogens into a single scoop
that's easy to drink and tastes great. It's designed to support things like gut health,
immune health, and over
all energy. And it does so by helping to fill any gaps that you might have in your daily nutrition.
And of course, we should all eat high quality whole foods, but most of us are probably not
getting enough prebiotics, vitamins, and minerals, and AG1 ensures that those gaps are filled.
If you'd like to try AG1, you can go to drinkag1.com/huberman to get a special offer.
For a limited time, AG1 is giving away a week's supply of AGZ, which is their sleep supplement,
and a free bottle of vitamin D3K2 with your subscription. AGZ is something that I helped
design. It tastes great, and it's the only sleep supplement I take. It has a collection of different
things in it that has dramatically improved my sleep, both my slow wave deep sleep and my rapid
eye movement sleep. And I absolutely love it. Again, that's drinkag1.com/huberman to get a
week's supply of AGZ and a bottle of D3K2 with your subscription. I get a lot of questions about
stutter. What can people with stutter do if they'd like to relieve their stutter?
Stutter is a condition where
the words can't come out fluently. So you have all the ideas, you've got the language intact.
Remember we talked about this distinction between language and speech. Stuttering is a problem of
speech, right? So the ideas, the meanings, the grammar, it's all there in people's stutter,
but they can't get the words out fluently. So that's a speech condition. And in particular,
it's a condition that affects articulation, specifically about controlling the production
of words.
Stuttering is a condition where people have a predisposition to it. So there's
an aspect of stuttering. You are a stutterer or you're not a stutterer, but people who stutter
don't stutter all the time either. So you could be a stutterer who stutters sometimes but not others.
The main link between stuttering and anxiety is that anxiety can provoke it and make it worse.
That's certainly true, but it's not necessarily caused by anxiety. It can essentially trigger it
or make it worse, but it's not the cause of it per se. So the cause of it is still really not clear,
but it does have to do with these kind of brain functions that we've been talking about earlier,
which is that in order to produce normal fluent speech, we're not even conscious of what is going
on in our mouths, in our larynx. We're not conscious. And if we were, we would not be able
to speak because it's too complex. It's too precise. It's something that we have really
developed the abilities to do, and we do it naturally, right? It's part of our programming
and part of what we learn inherently and, you know, it's just through exposure. So
stuttering is essentially a breakdown at certain times in that machinery being able to work in a
really coordinated way. You can think about, you know, the operations of these areas that
are controlling the vocal tract. Let's say speech is like a symphony. In order for it to come out,
normally you've got to have not just one part, the larynx, but the lips, the jaw. They can't be
doing their own thing. They have to be very, very precisely
activated and very, very precisely controlled in a way to actually create words. And so in stuttering,
there's a breakdown of that coordination. If somebody has a stutter, is it better
to address that early in life when there's still neuroplasticity?
is very robust. And if so, what's the typical route for treatment? I have to imagine it's not
brain surgery typically. I'm guessing there are speech therapists that people can talk to and
they can help them work out where they're getting stuck in the relationship to anxiety.
Yeah, exactly. I mean, part of it is about that anxiety, but a lot of it really has to do
with therapy to sort of like work through and think of tricks basically sometimes to create
conditions where you can actually get the words to come out. A lot of, some forms of stuttering
are really initiation problems. Just getting started itself is very hard. You want to start
with initial vowel or consonant, but it won't emit. So a lot of that therapy is really just
focusing on like, how do you create the conditions, you know, for that to happen?
And there's another aspect to it that I find very interesting is that the feedback, essentially
what we hear ourselves say, for example, every time that I say a word, I'm also hearing what
I'm saying. So that's what we call auditory feedback. That turns out to be very important.
And sometimes when you change that, it can actually change the amount someone stutters
for better or for worse. And it's giving us a clue that the brain is not just focused on
sending the commands.
But it's also possibly interacting with the part that is hearing the sounds. And there's something
might be going on in that connection that breaks down when stuttering occurs. So there are
individuals that are stutterers, but they don't stutter all the time. In those instances, there's
something happening in those particular moments where this very, very precise coordination needs
to happen in the brain in order to get the words out fluently. Eddie, I have to say from the first
time we became friends,
38 years ago, something like that. To be sitting here with you today for me is an absolute thrill,
not just because we've been friends for that long or that we got reacquainted through literally the
halls of medicine and science, but because I really do see what you're doing as really
representing that front, absolute cutting edge of exploration and application. I mean,
the story of Poncho is about one of your many patients that has derived tremendous benefit from
your work. And now as a chair of a department, you, of course, work alongside individuals who are
also doing incredible work in the spinal cord, et cetera. So on behalf of myself and everyone
listening, I just really want to thank you for joining us today to share this information,
but also just for the work you do. It's truly spectacular. So thank you ever so much.
Thanks.
Fresh groceries delivered right to your door. Our refrigerated trucks keep your items fresh
while our associates are committed to getting every detail right. Carefully selecting and
packing your order with care because bringing you fresh quality groceries is what we do best.
And right now, enjoy $20 off your first online order of $75 or more.
Restrictions apply. See site for details. Kroger, fresh for everyone.
Podcast Summary
Key Points:
Speech refers to the motor production of sound through the vocal tract, while language involves the comprehension of meaning, grammar, and context through syntax, semantics, and pragmatics.
The larynx generates vocalization through vibrating vocal folds during exhalation, with voice pitch differing between men and women due to anatomical size.
Speech production requires coordinated activity across the larynx, pharynx, tongue, lips, and jaw, forming consonants and vowels through shaping of air flow.
Non-linguistic vocalizations like crying or laughter originate in distinct brain areas separate from speech and language networks, indicating evolutionarily ancient vocal control.
Brain-machine interfaces (BCIs) allow paralyzed individuals with locked-in syndrome to communicate by decoding neural signals related to speech intentions and translating them into words.
Current BCIs use implanted electrodes and machine learning to interpret brain activity, with ongoing improvements in vocabulary size and accuracy, including features like autocorrect for better output.
Future developments include fully animated avatars that combine speech and facial expressions to create more natural, intuitive, and emotionally resonant communication.
While augmentation of cognitive or physical abilities via neurotechnology is feasible, ethical concerns around access, equity, and societal impact remain significant and underexplored.
Summary:
Dr. Eddie Chang discusses the neurobiology of speech and language, distinguishing speech—controlled by motor actions in the vocal tract—from language, which involves understanding meaning, grammar, and context. He explains how the larynx generates voice through vocal fold vibrations, and how speech production relies on precise coordination of multiple structures.
Chang highlights that non-linguistic vocalizations stem from distinct neural pathways, separate from language centers. His work on brain-machine interfaces (BCIs) enables paralyzed individuals with locked-in syndrome to communicate by decoding neural signals associated with speech intentions. A clinical trial involving a patient paralyzed for 15 years demonstrated successful translation of brain activity into words using machine learning and a small vocabulary, which is now expanding.
The technology is being enhanced with features like autocorrect and animated avatars that mimic facial expressions and speech, improving communication realism and user feedback. While such neurotechnologies are advancing rapidly, Chang cautions that broader applications for human enhancement—like super-memory or faster communication—raise significant ethical questions about access, equity, and societal implications. , through caffeine or nicotine), the invasive nature of current BCIs demands thoughtful evaluation.
Ultimately, the focus remains on restoring communication for those with severe disabilities, with future innovations aiming for more natural, embodied interactions that bridge the gap between thought and expression.
FAQs
Speech refers to the physical production of sound, like moving the lips and vocal tract to create words. Language encompasses the meaning, grammar, and understanding of those words—such as semantics, syntax, and pragmatics.
The larynx produces vocal sound through vibrations of the vocal folds when air passes through, creating the fundamental frequency of the voice. The pharynx and oral cavity shape the airflow to form consonants and vowels, shaping the speech signal.
Yes, researchers have developed brain-machine interfaces that decode neural activity associated with speech intentions. These systems can convert brain signals into words, allowing paralyzed individuals to communicate, as demonstrated in clinical trials like the BRAVO trial.
No, primitive vocalizations are controlled by different brain regions than speech. These vocalizations are linked to ancient brain structures and can occur even when speech areas are damaged, showing they are distinct from language-based communication.
These interfaces detect neural signals from brain areas involved in speech and translate them into words. This enables individuals with severe paralysis from ALS or brainstem stroke to communicate, even when they cannot move or speak voluntarily.
While augmentation is possible, it's likely to be subtle and incremental rather than dramatic. Current technology is not near the complexity of natural human speech systems, and ethical questions about access and societal impact remain significant.
Chat with AI
Loading...
Pro features
Go deeper with this episode
Unlock creator-grade tools that turn any transcript into show notes and subtitle files.