Speech Recognition & Synthesis

Speech Recognition & Synthesis

Developer(s)	Google
Initial release	November 13, 2013; 10 years ago (2013-11-13)

Stable release	20231225.02_p0.593665078(Android 8-14) / December 25, 2023; 35 days ago (2023-12-25)^[1]

Operating system	Android
Type	Screen reader

Speech Recognition & Synthesis, formerly known as Speech Services,^[2] is a screen reader application developed by Google for its Android operating system. It powers applications to read aloud (speak) the text on the screen, with support for many languages. Text-to-Speech may be used by apps such as Google Play Books for reading books aloud, Google Translate for reading aloud translations for the pronunciation of words, Google TalkBack, and other spoken feedback accessibility-based applications, as well as by third-party apps. Users must install voice data for each language.

Supported languages

Albanian (Albania)
Arabic
Bengali (Bangladesh)
Bengali (India)
Bosnian (Bosnia and Herzegovina)
Bulgarian (Bulgaria)
Cantonese (Hong Kong)
Catalan (Spain)
Chinese (China)
Chinese (Taiwan)
Croatian (Croatia)
Czech (Czech Republic)
Danish (Denmark)
Dutch (Belgium)
Dutch (Netherlands)
English (Australia)
English (Nigeria)
English (India)
English (United Kingdom)
English (United States)
Estonian (Estonia)
Filipino (Philippines)
Finnish (Finland)
French (Canadian)
French (France)
German (Germany)
Greek (Greece)
Gujarati (India)
Hebrew (Israel)
Hindi (India)
Hungarian (Hungary)
Icelandic (Iceland)
Indonesian (Indonesia)
Italian (Italy)
Japanese (Japan)
Javanese (Indonesia)
Kannada (India)
Khmer (Cambodia)
Korean (South Korea)
Latvian (Latvia)
Lithuanian (Lithuania)
Malay (Malaysia)
Malayalam (India)
Marathi (India)
Nepali (Nepal)
Norwegian Bokmål (Norway)
Polish (Poland)
Portuguese (Brazil)
Portuguese (Portugal)
Punjabi (India)
Romanian (Romania)
Russian (Russia)
Sinhala (Sri Lanka)
Slovak (Slovakia)
Spanish (Spain)
Spanish (United States)
Sundanese (Indonesia)
Swahili (Kenya)
Swedish (Sweden)
Tamil (India)
Telugu (India)
Thai (Thailand)
Turkish (Turkey)
Ukrainian (Ukraine)
Urdu (Pakistan)
Vietnamese (Vietnam)
Welsh (United Kingdom)

History

Some app developers have started adapting and tweaking their Android Auto apps to include Text-to-Speech, such as Hyundai in 2015.^[3] Apps such as textPlus and WhatsApp use Text-to-Speech to read notifications aloud and provide voice-reply functionality.

Google Cloud Text-to-Speech is powered by WaveNet,^[4] software created by Google's UK-based AI subsidiary DeepMind, which was bought by Google in 2014.^[5] It tries to distinguish from its competitors, Amazon and Microsoft.^[6]

DeepMind's AI voice synthesis tech is notably advanced and realistic. Most voice synthesizers (including Apple's Siri) use concatenative synthesis,^[4] in which a program stores individual phonemes and then pieces them together to form words and sentences.

WaveNet generates speech that sounds more natural than other text-to-speech systems. It synthesizes speech with more human-like emphasis and inflection on syllables, phonemes, and words. On average, a WaveNet produces speech audio that people prefer over other text-to-speech technologies. Unlike most other text-to-speech systems, a WaveNet model creates raw audio waveforms from scratch. The model uses a neural network that has been trained using a large volume of speech samples. During training, the network extracts the underlying structure of the speech, such as which tones follow each other and what a realistic speech waveform looks like. When given a text input, the trained WaveNet model can generate the corresponding speech waveforms from scratch, one sample at a time, with up to 24,000 samples per second and smooth transitions between the individual sounds.^[4]

The service was renamed Speech Recognition & Synthesis in 2023.^{[citation needed]}

References

^ "Speech Services by Google APKs". APKMirror.
^ Wang, Jules (November 8, 2021). "You'll never guess the latest Google app to cross 10 billion installs (seriously)". Android Police. Archived from the original on November 8, 2021. Retrieved November 18, 2021.
^ "Google, Hyundai show off new third-party Android Auto apps". CNET. CBS Interactive. Retrieved 17 January 2015.
^ ^a ^b ^c "WaveNet". www.deepmind.com. Retrieved 2023-06-22.
^ Gibbs, Samuel (2014-01-27). "Google buys UK artificial intelligence startup Deepmind for £400m". The Guardian. ISSN 0261-3077. Retrieved 2023-06-22.
^ "Text-to-Speech AI: Lifelike Speech Synthesis". Google Cloud. Retrieved 2023-06-22.

External links

Speech Recognition & Synthesis on Google Play

Google

Company

Divisions

Ads
AI
- Brain
- DeepMind
Android
China
- Goojje
Chrome
Cloud
Glass
Google.org
Health
Maps
Pixel
Search
- Timeline
Sidewalk Labs
Sustainability
YouTube
- History
- "Me at the zoo"
- Social impact
- YouTuber

People

Current	Krishna Bharat Vint Cerf Jeff Dean John Doerr Sanjay Ghemawat Al Gore John L. Hennessy Urs Hölzle Salar Kamangar Ray Kurzweil Ann Mather Alan Mulally Rick Osterloh Sundar Pichai (CEO) Ruth Porat (CFO) Rajen Sheth Hal Varian Susan Wojcicki Neal Mohan
Former	Andy Bechtolsheim Sergey Brin (Founder) David Cheriton Matt Cutts David Drummond Alan Eustace Timnit Gebru Omid Kordestani Paul Otellini Larry Page (Founder) Patrick Pichette Eric Schmidt Ram Shriram Amit Singhal Shirley M. Tilghman Rachel Whetstone

Real estate

Design

Fonts
- Croscore
- Noto
- Product Sans
- Roboto
Logo
- Doodle
  - Doodle Champion Island Games
  - Magic Cat Academy
Material Design

Events

Android Developer Challenge Developer Day Developer Lab Code-in Code Jam Developer Day Developers Live Doodle4Google G-Day I/O Jigsaw Living Stories Lunar XPRIZE Mapathon Science Fair Summer of Code Talks at Google
YouTube	Awards CNN/YouTube presidential debates Comedy Week Live Music Awards Space Lab Symphony Orchestra

Projects and
initiatives

20% project
Area 120
- Reply
- Tables
ATAP
Business Groups
Computing University Initiative
Data Liberation Front
Data Transfer Project
Developer Expert
Digital Garage
Digital News Initiative
Digital Unlocked
Dragonfly
Founders' Award
Free Zone
Get Your Business Online
Google for Education
Google for Startups
Labs
Liquid Galaxy
Made with Code
Māori
ML FairnessNative Client
News Lab
Nightingale
OKR
PowerMeter
Privacy Sandbox
Quantum Artificial Intelligence Lab
RechargeIT
Shield
Silicon Initiative
Solve for X
Starline
Student Ambassador Program
Submarine communications cables
- Dunant
- Grace Hopper
Sunroof
Versus Debates
YouTube
- Creator Awards
- Next Lab and Audience Development Group
- Original Channel Initiative
Zero

Criticism

2018 data breach 2018 walkouts Alphabet Workers Union Censorship DeGoogle "Did Google Manipulate Search for Hillary?" Dragonfly FairSearch "Ideological Echo Chamber" memo Litigation Privacy concerns Street View San Francisco tech bus protests Services outages Smartphone patent wars Worker organization
YouTube	Back advertisement controversy Censorship Copyright issues Copyright strike Elsagate Fantastic Adventures scandal Headquarters shooting Kohistan video case Reactions to Innocence of Muslims Slovenian government incident

Development

Operating systems

Android
- Automotive
- Glass OS
- Go
- gLinux
- Goobuntu
- Things
- TV
- Wear OS
ChromeOS
- ChromiumOS
- Neverware
Fuchsia
TV

Libraries/
frameworks

Platforms

App Engine AppJet Apps Script Cloud Platform Anvato Firebase Cloud Messaging Crashlytics Global IP Solutions Internet Low Bitrate Codec Internet Speech Audio Codec Gridcentric, Inc. ITA Software Kubernetes LevelDB Neatx SageTV
Apigee	Bigtable Bitium Chronicle VirusTotal Compute Engine Connect Dataflow Datastore Kaggle Looker Mandiant Messaging Orbitera Shell Stackdriver Storage

Tools

Search algorithms

Others

BERT BigQuery Chrome Experiments Flutter Gemini Googlebot Keyhole Markup Language LaMDA Open Location Code PaLM Programming languages Caja Carbon Dart Go Sawzall Transformer Viewdle Webdriver Torso Web Server
File formats	AAB APK AV1 On2 Technologies VP3 VP6 VP8 libvpx VP9 WebM WebP WOFF2

Products

Entertainment

Currents (news app) Green Throttle Games Owlchemy Labs Oyster PaperofRecord.com Podcasts Quick, Draw! Santa Tracker Songza Stadia games Typhoon Studios TV Vevo Video
Play	Books Games most downloaded apps Music Newsstand Pass Services
YouTube	BandPage BrandConnect Content ID Instant Kids Music Official channel Preferred Premium original programming YouTube Rewind RightsFlow Shorts Studio TV

Communication

Aardvark
Alerts
Answers
Base
BeatThatQuote.com
Blog Search
Books
- Ngram Viewer
Code Search
Data Commons
Dataset Search
Dictionary
Directory
Fast Flip
Flu Trends
Finance
Goggles
Google.by
Images
- Image Labeler
- Image Swirl
Kaltix
Knowledge Graph
- Freebase
- Metaweb
Like.com
News
- Archive
- Weather
Patents
People Cards
Personalized Search
Public Data Explorer
Questions and Answers
SafeSearch
Scholar
Searchwiki
Shopping
Catalogs
- Express
Squared
Tenor
Travel
- Flights
Trends
- Insights for Search
Voice Search
WDYL

Navigation

Earth
Endoxon
ImageAmerica
Maps
- Latitude
- Map Maker
- Navigation
- Pin
- Street View
  - Coverage
  - Trusted
Waze

Business
and finance

Ad Manager
AdMob
Ads
Adscape
AdSense
Attribution
BebaPay
Checkout
Contributor
DoubleClick
- Affiliate Network
- Invite Media
Marketing Platform
- Analytics
- Looker Studio
- Urchin
Pay (mobile app)
- Wallet
- Pay (payment method)
- Send
- Tez
PostRank
Primer
Softcard
Wildfire Interactive
Widevine

Organization
and productivity

Bookmarks Browser Sync Calendar Cloud Search Desktop Drive Etherpad fflick Files iGoogle Jamboard Notebook One Photos Quickoffice Quick Search Box Surveys Sync Tasks Toolbar
Docs Editors	Docs Drawings Forms Fusion Tables Keep Sheets Slides Sites
Publishing	Apture Blogger Pyra Labs Domains FeedBurner One Pass Page Creator Sites Web Designer

Education

Others

Account Dashboard Takeout Android Auto Android Beam Arts & Culture Assistant Authenticator Bard Body BufferBox Building Maker BumpTop Cast List of supported apps Cloud Print Crowdsource Digital Wellbeing Expeditions Family Link Find My Device Fit Google Fonts Gboard Gesture Search Impermium Knol Lively Live Transcribe MyTracks Nearby Share Now Offers Opinion Rewards Person Finder Poly Question Hub Quick Share Reader Safe Browsing Sidewiki SlickLogin Sound Amplifier Speech Services Station Store TalkBack Tilt Brush URL Shortener Voice Access Wavii Web Light WiFi
Chrome	Apps Chromium Dinosaur Game GreenBorder Remote Desktop Web Store V8
Images and photography	Camera Lens Snapseed Nik Software Panoramio Photos Picasa Web Albums Picnik

Hardware

Smartphones	Android Dev Phone Android One Nexus Nexus One S Galaxy Nexus 4 5 6 5X 6P Comparison Pixel Pixel 2 3 3a 4 4a 5 5a 6 6a 7 7a Fold 8 Comparison Play Edition Project Ara
Laptops and tablets	Chromebook Nexus 7 (2012) 7 (2013) 10 9 Comparison Pixel Chromebook Pixel Pixelbook Pixelbook Go C Slate Tablet
Wearables	Fitbit List of products Pixel Buds Pixel Watch Pixel Watch 2 Project Iris (unreleased) Virtual reality Cardboard Contact Lens Daydream Glass
Others	Chromebit Chromebox Clips Digital media players Chromecast Nexus Player Nexus Q Dropcam Liquid Galaxy Nest Smart Speakers Thermostat Wifi OnHub Pixel Visual Core Search Appliance Sycamore processor Tensor Tensor Processing Unit Titan Security Key

v t e Litigation
Advertising	Feldman v. Google, Inc. (2007) Rescuecom Corp. v. Google Inc. (2009) Goddard v. Google, Inc. (2009) Rosetta Stone Ltd. v. Google, Inc. (2012) Google, Inc. v. American Blind & Wallpaper Factory, Inc. (2017) Jedi Blue
Antitrust	European Union (2010–present) United States v. Adobe Systems, Inc., Apple Inc., Google Inc., Intel Corporation, Intuit, Inc., and Pixar (2011) Umar Javeed, Sukarma Thapar, Aaqib Javeed vs. Google LLC and Ors. (2019) United States v. Google LLC (2020) United States v. Google LLC (2023)
Intellectual property	Perfect 10, Inc. v. Amazon.com, Inc. and A9.com Inc. and Google Inc. (2007) Viacom International Inc. v. YouTube, Inc. (2010) Lenz v. Universal Music Corp.(2015) Authors Guild, Inc. v. Google, Inc. (2015) Field v. Google, Inc. (2016) Google LLC v. Oracle America, Inc. (2021) Smartphone patent wars
Privacy	Rocky Mountain Bank v. Google, Inc. (2009) Hibnick v. Google, Inc. (2010) United States v. Google Inc. (2012) Judgement of the German Federal Court of Justice on Google's autocomplete function (2013) Joffe v. Google, Inc. (2013) Mosley v SARL Google (2013) Google Spain v AEPD and Mario Costeja González (2014) Frank v. Gaos (2019)
Other	Garcia v. Google, Inc. (2015) Google LLC v Defteros (2020) Epic Games v. Google (2021) Gonzalez v. Google LLC (2022)
Category

Terms and phrases	"Don't be evil" Gayglers Google (verb) Google bombing 2004 U.S. presidential election Google effect Googlefight Google hacking Googleshare Google tax Googlewhack Googlization "Illegal flower tribute" Rooting Search engine manipulation effect Sitelink Site reliability engineering YouTube poop
Documentaries	AlphaGo Google: Behind the Screen Google Maps Road Trip Google and the World Brain The Creepy Line
Books	Google Hacks The Google Story Google Volume One Googled: The End of the World as We Know It How Google Works I'm Feeling Lucky In the Plex The Google Book
Popular culture	Google Feud Google Me (film) "Google Me" (Kim Zolciak song) "Google Me" (Teyana Taylor song) Is Google Making Us Stupid? Proceratium google Matt Nathanson: Live at Google The Billion Dollar Code The Internship Where on Google Earth is Carmen Sandiego?
Others	"Attention Is All You Need" elgooG g.co .google Pimp My Search Predictions of the end Relationship with Wikipedia Sensorvault Stanford Digital Library Project

Italics indicate discontinued products or services.
Category
Commons
Outline
WikiProject

Android

Android Go
- Comparison of products

Software
development

Development tools

Official

Android Runtime (ART)
Software development kit (SDK)
- Android Debug Bridge (ADB)
- Fastboot
- Android App Bundle
- Android application package (APK)
Bionic
Dalvik
Firebase
- Google Cloud Messaging (GCM)
- Firebase Cloud Messaging (FCM)
Google Mobile Services (GMS)
Native development kit (NDK)
Open accessory development kit (OADK)
RenderScript
Skia
AdMob
Material Design
Fonts
- Droid
- Roboto
- Noto
Google Developers

Other

Integrated
development
environments (IDE)

Languages, databases

Virtual reality (VR)

Events, communities

Releases

Derivatives

Devices

Pixel	C Pixel & Pixel XL 2 & 2 XL 3 & 3 XL 3a & 3a XL 4 & 4 XL 4a & 4a (5G) 5 5a 6 & 6 Pro 6a 7 & 7 Pro 7a Fold
Nexus	One S Galaxy Nexus 4 10 Q 5 5X 6 6P 7 2012 2013 9 Player
Play edition	HTC One (M7) HTC One (M8) LG G Pad 8.3 Moto G Samsung Galaxy S4 Sony Xperia Z Ultra
Android One other smartphones

Custom
distributions

AliOS
Android-x86
- Remix OS
AOKP
Baidu Yi
Barnes & Noble Nook
CalyxOS
ColorOS
- realme UI
CopperheadOS
EMUI
- Magic UI
Fire OS
Flyme OS
GrapheneOS
LeWa OS
LineageOS
- /e/
- CrDroid
- CyanogenMod
- DivestOS
- iodéOS
- Kali NetHunter
LiteOS
MicroG
MIUI
- MIUI for POCO
Nokia X software platform
OmniROM
OPhone
OxygenOS
PixelExperience
Pixel UI
Replicant
Resurrection Remix OS
SlimRoms
TCL UI
Ubuntu for Android
XobotOS
ZUI

Booting and recovery

APIs

Alternative UIs

Rooting

Lists

From Wikipedia, the free encyclopedia

Supported languages

History

See also

References

External links