What to know about How do computers talk with voices that sound like people?
The article explains the technical evolution of text-to-speech technology, from early mechanical bellows and synthesizers to modern AI-driven machine learning. It describes how computers simulate human speech patterns and discusses both the beneficial applications and the potential for misuse through audio deepfakes.
Propaganda risk0%
Claims checked11
Techniques found0
Topics0
Coverage spectrum
Coverage gap: Low Left coverage
Left0%
Center100%
Right0%
4 sources compared across this story cluster. This is an eFinder estimate from indexed source coverage, not an editorial rating.
What happened
Curious Kids is a series for children of all ages.
Why it matters
If you have a question you’d like an expert to answer, send it to CuriousKidsUS@theconversation.com.
Common ground
– Sarah G٫ age 11٫ Seguin٫ Texas When you talk to computerized assistants like Siri or Alexa, they reply in voices that sound very human.
Perspective signals
No major persuasion pattern has been attached yet, so the source, headline, and evidence should carry most of the weight for readers.
Follow-up questions
What concrete event or decision sits underneath the headline: How do computers talk with voices that sound like people??
What evidence would most clearly confirm or weaken the claim that Today’s computers use machine learning, a kind of artificial intelligence, to sound like a person?
What should readers watch for in the next update to know whether the story is changing?
The article explains the technical evolution of text-to-speech technology, from early mechanical bellows and synthesizers to modern AI-driven machine learning. It describes how computers simulate human speech patterns and discusses both the beneficial applications and the potential for misuse through audio deepfakes.
Low risk. This article shows minimal use of propaganda techniques.
fact_checkClaims Checked
eFinder analyzed this article and checked 11 claims against available evidence, cross-references, web search, and Wikipedia. Here is what the fact-checking layer found.
helpInsufficient Evidence3
check_circleCorroborated3
infoSingle Source2
verifiedVerified By Reference2
schedulePending1
help
Claim 1: “Today’s computers use machine learning, a kind of artificial intelligence, to sound like a person.”
INSUFFICIENT EVIDENCE
No evidence was provided in the search results to verify the use of machine learning in modern TTS systems.
info
Claim 2: “If a computer wants to say “Hello, how are you?” it breaks each word into bits of sound called phonemes.”
SINGLE SOURCE
The provided evidence for this claim consists of search results for 'text messaging' and 'texting apps', which are irrelevant to the technical process of text-to-speech phoneme breakdown.
travel_explore
web search
NEUTRAL
— Text messaging, or texting, is the act of composing and sending electronic messages, typically consisting of alphabetic and numeric characters, between two or more users of mobile phones, tablet compu…
https://en.wikipedia.org/wiki/Text_messaging
travel_explore
web search
NEUTRAL
— Ditch your monthly phone bills with free calling and texting. Get the TextNow SIM card or use the TextNow app over wifi. Start saving!
https://www.textnow.com/login
travel_explore
web search
NEUTRAL
— Make TextFree your everyday line. Call and text friends, family, and anyone else—it’s free, unlimited, and available on all of your devices. Need a spare line for a side hustle, marketplace listings, …
https://textfree.com/
verified
Claim 3: “Way back in the 1700s, inventors tried to make machines work like your lungs and throat do. They used bellows... to push air from inside the bag through pipes, whistles and leather tubes.”
VERIFIED BY REFERENCE
Web search results confirm that 18th-century inventors like C.G. Kratzenstein and Wolfgang von Kempelen created 'talking heads' using organ pipes and other mechanical means to simulate human speech.
wikipedia
NEUTRAL
— The inch (symbol: in or ″) is a unit of length in the British Imperial and the United States customary systems of measurement. It is equal to 1/36 yard or 1/12 of a foot. Derived from the Roman un…
https://en.wikipedia.org/wiki/Inch
menu_book
wikipedia
NEUTRAL
— Killed in action (KIA) is a casualty classification generally used by militaries to describe the deaths of their personnel at the hands of enemy or hostile forces at the moment of action. The United S…
https://en.wikipedia.org/wiki/Killed_in_action
+ 3 more evidence sources
help
Claim 4: “This advanced technology even allows a sophisticated AI computer to listen to a recording of your voice for just a few seconds, learn your exact speech patterns and copy it.”
INSUFFICIENT EVIDENCE
No evidence was provided in the search results to verify the specific claim regarding voice cloning from a few seconds of audio.
verified
Claim 5: “One famous machine was called Voder, which made its debut at the 1939 World’s Fair in New York City.”
VERIFIED BY REFERENCE
Wikipedia explicitly confirms that the Voder (voice operating demonstrator) was invented by Homer Dudley and debuted at the 1939 New York World's Fair.
menu_book
wikipedia
NEUTRAL
— The 1939 New York World's Fair (also known as the 1939–1940 New York World's Fair) was an international exposition held at Flushing Meadows–Corona Park in Queens, New York City, United States. The fai…
https://en.wikipedia.org/wiki/1939_New_York_World's_Fair
menu_book
wikipedia
NEUTRAL
— The 1939 New York World's Fair took place at Flushing Meadows–Corona Park in Queens, New York, United States, during 1939 and 1940. The fair included pavilions with exhibits by 62 nations, 34 U.S. sta…
https://en.wikipedia.org/wiki/1939_New_York_World's_Fair_pav…
menu_book
wikipedia
NEUTRAL
— The Bell Telephone Laboratory's voder (abbreviation of voice operating demonstrator) was the first attempt to electronically synthesize human speech by breaking it down into its acoustic components. I…
https://en.wikipedia.org/wiki/Voder
+ 3 more evidence sources
help
Claim 6: “When you speak, your lungs push air up your windpipe and through the vocal cords in your throat. That makes the vocal cords vibrate, which creates sound.”
INSUFFICIENT EVIDENCE
No evidence was provided in the search results to verify the biological process of speech production described in the claim.
check_circle
Claim 7: “older software programs had to stitch together small sounds that had been mapped out from recorded voices. The maps, called spectrograms, look like graphs with peaks and valleys representing how strong each tone was”
CORROBORATED
Web search results confirm that spectrograms are used to visualize the pitch, volume, and timbre of speech and have a history in speech acoustics dating back to the late 1800s.
travel_explore
web search
NEUTRAL
— An introduction to how spectrograms help us "see" the pitch, volume and timbre of a sound. The spectrogram used in this video is called Signal Spy for iPadNMC Learning at Home: Spectrograms & Your Voi…
https://www.youtube.com/watch?v=_FatxGN3vAM
web search
NEUTRAL
— If you capture a voice in your recording then a spectrogram will tell you one of two things, either the voice falls within the normal frequency range of human speech, or the voice does not match the t…
https://www.higgypop.com/news/spectrograms-and-evp-analysis/
schedule
Claim 8: “Advanced software can create highly realistic fake voices, sometimes called audio deepfakes.”
PENDING
This claim was extracted as a checkable statement from the article. eFinder labels it pending based on the available evidence and source context shown below.
info
Claim 9: “Computers began speaking by putting together phonemes in the 1960s”
SINGLE SOURCE
While web results discuss phoneme concatenation in the context of TTS, there is no specific evidence provided that confirms this practice began specifically in the 1960s.
menu_book
wikipedia
NEUTRAL
— Acorn Computers Ltd. was a British computer company established in Cambridge, England in 1978 by Hermann Hauser, Chris Curry and Andy Hopper. The company produced a number of computers during the 1980…
https://en.wikipedia.org/wiki/Acorn_Computers
menu_book
wikipedia
NEUTRAL
— A computer is a machine that can be programmed to automatically carry out sequences of arithmetic or logical operations (computation). Modern digital electronic computers can perform generic sets of o…
https://en.wikipedia.org/wiki/Computer
menu_book
wikipedia
NEUTRAL
— A personal computer (PC), or simply computer, is a computer designed for personal use. It is typically used for tasks such as word processing, web browsing, email, file management, spreadsheets, and v…
https://en.wikipedia.org/wiki/Personal_computer
+ 3 more evidence sources
check_circle
Claim 10: “The first electronic speech machines, called synthesizers, were built in the 1930s.”
CORROBORATED
Evidence from YouTube and other web sources mentions the development of speech synthesizers starting around 1939 (The Voder), supporting the timeline of the 1930s.
menu_book
wikipedia
NEUTRAL
— Electronic Arts Inc. (EA) is an American video game company headquartered in Redwood City, California. Founded in May 1982 by former Apple employee Trip Hawkins, the company was a pioneer of the early…
https://en.wikipedia.org/wiki/Electronic_Arts
menu_book
wikipedia
NEUTRAL
— Electronic music broadly is a group of music genres that employ electronic musical instruments, circuitry-based music technology and software, or general-purpose electronics (such as personal computer…
https://en.wikipedia.org/wiki/Electronic_music
menu_book
wikipedia
NEUTRAL
— Electronic may refer to:
Electronics, a scientific and technical field
Electronics (magazine), a defunct American trade journal
Electronic storage, the storage of data using an electronic device
Elec…
https://en.wikipedia.org/wiki/Electronic
+ 3 more evidence sources
check_circle
Claim 11: “They use a technology called text-to-speech.”
CORROBORATED
Multiple sources, including Fast Company and general web search results, confirm that virtual assistants like Siri and Alexa utilize text-to-speech technology to communicate.
menu_book
wikipedia
NEUTRAL
— In machine learning, deep learning (DL) focuses on utilizing multilayered neural networks to perform tasks such as classification, regression, and representation learning. The field takes inspiration …
https://en.wikipedia.org/wiki/Deep_learning
menu_book
wikipedia
NEUTRAL
— Electronic discovery (also ediscovery or e-discovery) refers to discovery in legal proceedings such as litigation, government investigations, or Freedom of Information Act requests, where the informat…
https://en.wikipedia.org/wiki/Electronic_discovery
menu_book
wikipedia
NEUTRAL
— A virtual assistant (VA) is a software agent that can perform a range of tasks or services for a user based on user input, such as commands or questions, including verbal ones. Such technologies often…
https://en.wikipedia.org/wiki/Virtual_assistant
+ 3 more evidence sources
infoDisclaimer: This analysis is generated by AI and should be used as a starting point for critical thinking, not as definitive truth. Claims are verified against publicly available sources. Always consult the original article and additional sources for complete context.