New Issue: Science’s Impossible Questions. Read Now

Brain Activity Decoded to Produce Intelligible, Synthesized Speech

New device is a step toward translating thoughts into machine-spoken words

Join Our Community of Science Lovers!

Neurological conditions that can cause paralysis, such as amyotrophic lateral sclerosis (ALS) and strokes in the brain stem, also rob many patients of their ability to speak. Assistive technologies enable keyboard control for some of these individuals (like the famed late physicist Stephen Hawking), and brain-computer interfaces make it possible for others to control machines directly with their thoughts. But both types of devices are slow and impractical for people with locked-in syndrome and other communication impairments.

Now researchers are developing tools to eavesdrop on speech-related brain activity, decode it and convert it into words spoken by a machine. A recent study used state-of-the-art machine learning and speech-synthesis technology to yield some of the most impressive results to date.

Electrical engineer Nima Mesgarani of Columbia University’s Zuckerman Institute and his colleagues studied five epilepsy patients who had electrodes implanted in or on their brain as part of their treatment. The electrodes covered regions involved in processing speech sounds. The patients listened to stories being read aloud as their brain activity was recorded. The team trained a “deep learning” neural network to match this activity with the corresponding audio. The test was then whether, given neural data it had not seen before, the system could reproduce the original speech.


On supporting science journalism

If you're enjoying this article, consider supporting our award-winning journalism by subscribing. By purchasing a subscription you are helping to ensure the future of impactful stories about the discoveries and ideas shaping our world today.


When the patients heard the digits zero through nine spoken four times each, the system transformed the neural data into values needed to drive a vocoder, a special kind of speech synthesizer. A separate group of participants heard the synthesized words and identified them correctly 75 percent of the time, according to the study, published in January in Scientific Reports. Most previous efforts have not measured how well such reconstructed speech can be understood. “We show that it’s intelligible,” Mesgarani says.

Researchers already knew it was possible to reconstruct speech from brain activity, but the new work is a step toward higher performance. “There’s a lot of room for improvement, but we know the information is there,” says neurosurgeon Edward Chang of the University of California, San Francisco, who was not involved in the study. “Over the next few years it’s going to get even better—this is a field that’s evolving quickly.”

There are some limitations. Mesgarani’s team recorded brain activity from speech-perception regions, not speech-production ones; the researchers also evaluated their system on only a small set of words instead of complete sentences drawing on a large vocabulary. (Other researchers, including Chang, are already working on these problems.) Perhaps most important, the study was designed to decode activity related to speech that was actually heard rather than merely imagined—the latter feat will be required to develop a practical device. “The challenge for all of us is actual versus imagined” speech, Mesgarani says.

Simon Makin is a freelance science journalist based in the U.K. His work has appeared in New Scientist, the Economist, Scientific American and Nature, among others. He covers the life sciences and specializes in neuroscience, psychology and mental health. Follow Makin on X (formerly Twitter) @SimonMakin

More by Simon Makin
Scientific American Magazine Vol 320 Issue 5This article was published with the title “Decoding Speech” in Scientific American Magazine Vol. 320 No. 5 (), p. 18
doi:10.1038/scientificamerican0519-18a

Subscribe to Support Independent Journalism

Great science journalism requires human expertise, time, effort and creativity. And it costs money. That’s why I and the journalists here at Scientific American hope you’ll join our community.

When you subscribe, you are supporting staff and freelance journalists who are passionate about telling science stories that are true, important and compelling. Our editors and reporters are often experts in their fields, which means they understand the nuances of big discoveries and can untangle the breakthroughs from the hype. With a subscription, you are also supporting rigorous fact-checking to ensure the words we publish are precise and accurate. And you’re supporting original illustrations, graphics and photos that bring you closer to an advanced laboratory, an ice sheet in Antarctica or a space mission in orbit. You’re helping us craft other types of high-quality journalism as well: Our newsletters are carefully written, edited and curated by staffers you have or will come to know and love. Our Science Quickly podcast is based on original reporting, collaboration with editors and scientists and exacting production.

Subscriptions keep this engine running so we can continue to deliver thoughtful, rigorous and independent science journalism to you. In an era of viral misinformation, this work is crucial. If you value what we do, I hope you’ll consider joining us as a subscriber

Thank you,

Jeanna Bryner, Editor in Chief, Scientific American

Subscribe