Speech Recognition & Synthesis

(Redirected from Google Text-to-Speech) Screen reader application by Google

Speech Recognition & Synthesis

Developer(s)	Google
Initial release	10 October 2013; 11 years ago (2013-10-10)

Stable release	20241030.02/p3 (Build 702043126) / 2 December 2024; 32 days ago (2024-12-02)

Operating system	Android
Type	Screen reader

Speech Recognition & Synthesis, formerly known as Speech Services, is a screen reader application developed by Google for its Android operating system. It powers applications to read aloud (speak) the text on the screen, with support for many languages. Text-to-Speech may be used by apps such as Google Play Books for reading books aloud, Google Translate for reading aloud translations for the pronunciation of words, Google TalkBack, and other spoken feedback accessibility-based applications, as well as by third-party apps. Users must install voice data for each language.

Supported languages

Afrikaans (South Africa)
Albanian (Albania)
Amharic (Ethiopia)
Arabic (Saudi Arabia)
Assamese (India)
Basque (Spain)
Bengali (Bangladesh)
Bengali (India)
Bodo (India)
Bosnian (Bosnia and Herzegovina)
Bulgarian (Bulgaria)
Burmese (Myanmar)
Cantonese (Hong Kong)
Catalan (Spain)
Chinese (China)
Chinese (Taiwan)
Croatian (Croatia)
Czech (Czech Republic)
Danish (Denmark)
Dogri (India)
Dutch (Belgium)
Dutch (Netherlands)
English (Australia)
English (Nigeria)
English (India)
English (United Kingdom)
English (United States)
Estonian (Estonia)
Filipino (Philippines)
Finnish (Finland)
French (Canada)
French (France)
Galician (Spain)
German (Germany)
Greek (Greece)
Gujarati (India)
Hausa (Nigeria)
Hebrew (Israel)
Hindi (India)
Hungarian (Hungary)
Icelandic (Iceland)
Indonesian (Indonesia)
Italian (Italy)
Japanese (Japan)
Javanese (Indonesia)
Kannada (India)
Kashmiri (India)
Khmer (Cambodia)
Konkani (India)
Korean (South Korea)
Latin (Vatican City)
Latvian (Latvia)
Lithuanian (Lithuania)
Maithili (India)
Malay (Malaysia)
Malayalam (India)
Manipuri (India)
Marathi (India)
Nepali (Nepal)
Norwegian (Norway)
Odia (India)
Polish (Poland)
Portuguese (Brazil)
Portuguese (Portugal)
Punjabi (India)
Romanian (Romania)
Russian (Russia)
Sanskrit (India)
Santali (India)
Serbian (Serbia)
Sindhi (India)
Sinhala (Sri Lanka)
Slovak (Slovakia)
Slovenian (Slovenia)
Spanish (Spain)
Spanish (United States)
Sundanese (Indonesia)
Swahili (Kenya)
Swedish (Sweden)
Tamil (India)
Telugu (India)
Thai (Thailand)
Turkish (Turkey)
Ukrainian (Ukraine)
Urdu (Pakistan)
Urdu (India)
Vietnamese (Vietnam)
Welsh (United Kingdom)

History

This article needs additional citations for verification. Please help improve this article by adding citations to reliable sources. Unsourced material may be challenged and removed.
Find sources: "Speech Recognition & Synthesis" – news · newspapers · books · scholar · JSTOR (November 2023) (Learn how and when to remove this message)

Some app developers have started adapting and tweaking their Android Auto apps to include Text-to-Speech, such as Hyundai in 2015. Apps such as textPlus and WhatsApp use Text-to-Speech to read notifications aloud and provide voice-reply functionality.

Google Cloud Text-to-Speech is powered by WaveNet, software created by Google's UK-based AI subsidiary DeepMind, which was bought by Google in 2014. It tries to distinguish from its competitors, Amazon and Microsoft.

Most voice synthesizers (including Apple's Siri) use concatenative synthesis, in which a program stores individual phonemes and then pieces them together to form words and sentences. WaveNet synthesizes speech with human-like emphasis and inflection on syllables, phonemes, and words. Unlike most other text-to-speech systems, a WaveNet model creates raw audio waveforms from scratch. The model uses a neural network that has been trained using a large volume of speech samples. During training, the network extracts the underlying structure of the speech, such as which tones follow each other and what a realistic speech waveform looks like. When given a text input, the trained WaveNet model can generate the corresponding speech waveforms from scratch, one sample at a time, with up to 24,000 samples per second and smooth transitions between the individual sounds.

The service was renamed Speech Recognition & Synthesis in 2023.

References

"Speech Recognition & Synthesis". Google Play. Retrieved 2024-12-11.
"Speech Recognition & Synthesis googletts.google-speech-apk_20241125.02_p2.702443970". APKMirror. 2024-12-11. Retrieved 2024-12-11.
Wang, Jules (November 8, 2021). "You'll never guess the latest Google app to cross 10 billion installs (seriously)". Android Police. Archived from the original on November 8, 2021. Retrieved November 18, 2021.
"Google, Hyundai show off new third-party Android Auto apps". CNET. CBS Interactive. Retrieved 17 January 2015.
^ "WaveNet". www.deepmind.com. Retrieved 2023-06-22.
Gibbs, Samuel (2014-01-27). "Google buys UK artificial intelligence startup Deepmind for £400m". The Guardian. ISSN 0261-3077. Retrieved 2023-06-22.
"Text-to-Speech AI: Lifelike Speech Synthesis". Google Cloud. Retrieved 2023-06-22.

External links

Speech Recognition & Synthesis on Google Play

Google

a subsidiary of Alphabet

Company

Divisions

Subsidiaries

Active

Defunct

Programs

Events

Infrastructure

111 Eighth Avenue
Android lawn statues
Androidland
Barges
Binoculars Building
Central Saint Giles
Chelsea Market
Chrome Zone
Data centers
GeoEye-1
Googleplex
Ivanpah Solar Power Facility
James R. Thompson Center
King's Cross
Mayfield Mall
Pier 57
Sidewalk Toronto
St. John's Terminal
Submarine cables
- Dunant
- Grace Hopper
- Unity
WiFi
YouTube Space
YouTube Theater

People

Current	Krishna Bharat Vint Cerf Jeff Dean John Doerr Sanjay Ghemawat Al Gore John L. Hennessy Urs Hölzle Salar Kamangar Ray Kurzweil Ann Mather Alan Mulally Rick Osterloh Sundar Pichai (CEO) Ruth Porat (CFO) Rajen Sheth Hal Varian Neal Mohan
Former	Andy Bechtolsheim Sergey Brin (co-founder) David Cheriton Matt Cutts David Drummond Alan Eustace Timnit Gebru Omid Kordestani Paul Otellini Larry Page (co-founder) Patrick Pichette Eric Schmidt Ram Shriram Amit Singhal Shirley M. Tilghman Rachel Whetstone Susan Wojcicki

Criticism

General	Censorship DeGoogle FairSearch "Google's Ideological Echo Chamber" No Tech for Apartheid Privacy concerns Street View YouTube Worker organization Alphabet Workers Union YouTube copyright issues
Incidents	Backdoor advertisement controversy Blocking of YouTube videos in Germany Data breach Elsagate Fantastic Adventures scandal Kohistan video case Reactions to Innocence of Muslims San Francisco tech bus protests Services outages Slovenian government incident Walkouts YouTube headquarters shooting

Other

Development

Software

A–C	Accelerated Linear Algebra AMP Actions on Google ALTS American Fuzzy Lop Android Cloud to Device Messaging Android Debug Bridge Android NDK Android Runtime Android SDK Android Studio Angular AngularJS Apache Beam APIs App Engine App Inventor App Maker App Runtime for Chrome AppJet Apps Script AppSheet ARCore Base Bazel Bigtable BigQuery Bionic Blockly Borg Caja Cameyo Chart API Charts Chrome Enterprise Premium Chrome Frame Chromium Blink Closure Tools Cloud Connect Cloud Dataflow Cloud Datastore Cloud Messaging Cloud Shell Cloud Storage Code Search Compute Engine Cpplint
D–N	Dalvik Data Protocol Dialogflow Exposure Notification Fast Pair Fastboot Federated Learning of Cohorts File System Firebase Firebase Cloud Messaging FlatBuffers Flutter Freebase Gadgets Ganeti Gears Gerrit GLOP gRPC Gson Guava Guetzli Guice gVisor GYP JAX Jetpack Compose Keyhole Markup Language Kubernetes Kythe LevelDB Lighthouse Looker Studio lmctfy MapReduce Mashup Editor Matter Mobile Services Namebench Native Client Neatx Neural Machine Translation Nomulus
O–Z	Open Location Code OpenRefine OpenSocial Optimize OR-Tools Pack PageSpeed Piper Plugin for Eclipse Polymer Programmable Search Engine Project IDX Project Shield Public DNS reCAPTCHA RenderScript SafetyNet SageTV Schema.org Search Console Shell Sitemaps Skia Graphics Engine Spanner Sputnik Stackdriver Swiffy Tango TensorFlow Tesseract Test Translator Toolkit Urchin UTM parameters V8 VirusTotal VisBug Wave Federation Protocol Weave Web Accelerator Web Designer Web Server Web Toolkit Webdriver Torso WebRTC

Operating systems

Android
- Cupcake
- Donut
- Eclair
- Froyo
- Gingerbread
- Honeycomb
- Ice Cream Sandwich
- Jelly Bean
- KitKat
- Lollipop
- Marshmallow
- Nougat
- Oreo
- Pie
- 10
- 11
- 12
- 13
- 14
- 15
- 16
- version history
- smartphones
Android Automotive
Android Go
- devices
Android Things
Android TV
- devices
Android XR
ChromeOS
ChromiumOS
Fuchsia
Glass OS
gLinux
Goobuntu
TV
Wear OS

Language models

Neural networks

Computer programs

Formats and codecs

Programming languages

Search algorithms

Domain names

Typefaces

Products (software and services)

Defunct or discontinued

Hardware

Pixel

Smartphones	Pixel (2016) Pixel 2 (2017) Pixel 3 (2018) Pixel 3a (2019) Pixel 4 (2019) Pixel 4a (2020) Pixel 5 (2020) Pixel 5a (2021) Pixel 6 (2021) Pixel 6a (2022) Pixel 7 (2022) Pixel 7a (2023) Pixel Fold (2023) Pixel 8 (2023) Pixel 8a (2024) Pixel 9 (2024) Pixel 9 Pro Fold (2024)
Smartwatches	Pixel Watch (2022) Pixel Watch 2 (2023) Pixel Watch 3 (2024)
Tablets	Pixel C (2015) Pixel Slate (2018) Pixel Tablet (2023)
Laptops	Chromebook Pixel (2013–2015) Pixelbook (2017) Pixelbook Go (2019)
Other	Pixel Buds (2017–present)

Nexus

Smartphones	Nexus One (2010) Nexus S (2010) Galaxy Nexus (2011) Nexus 4 (2012) Nexus 5 (2013) Nexus 6 (2014) Nexus 5X (2015) Nexus 6P (2015)
Tablets	Nexus 7 (2012) Nexus 10 (2012) Nexus 7 (2013) Nexus 9 (2014)
Other	Nexus Q (2012) Nexus Player (2014)

Other

Android Dev Phone
Android One
Cardboard
Chromebit
Chromebook
Chromebox
Chromecast
Clips
Daydream
Fitbit
Glass
Liftware
Liquid Galaxy
Nest
- smart speakers
- Thermostat
- Wifi
Play Edition
Project Ara
OnHub
Pixel Visual Core
Project Iris
Search Appliance
Sycamore processor
Tensor
Tensor Processing Unit
Titan Security Key

v t e Litigation
Advertising	Feldman v. Google, Inc. (2007) Rescuecom Corp. v. Google Inc. (2009) Goddard v. Google, Inc. (2009) Rosetta Stone Ltd. v. Google, Inc. (2012) Google, Inc. v. American Blind & Wallpaper Factory, Inc. (2017) Jedi Blue
Antitrust	European Union (2010–present) United States v. Adobe Systems, Inc., Apple Inc., Google Inc., Intel Corporation, Intuit, Inc., and Pixar (2011) Umar Javeed, Sukarma Thapar, Aaqib Javeed vs. Google LLC and Ors. (2019) United States v. Google LLC (2020) United States v. Google LLC (2023)
Intellectual property	Perfect 10, Inc. v. Amazon.com, Inc. (2007) Viacom International Inc. v. YouTube, Inc. (2010) Lenz v. Universal Music Corp.(2015) Authors Guild, Inc. v. Google, Inc. (2015) Field v. Google, Inc. (2016) Google LLC v. Oracle America, Inc. (2021) Smartphone patent wars
Privacy	Rocky Mountain Bank v. Google, Inc. (2009) Hibnick v. Google, Inc. (2010) United States v. Google Inc. (2012) Judgement of the German Federal Court of Justice on Google's autocomplete function (2013) Joffe v. Google, Inc. (2013) Mosley v SARL Google (2013) Google Spain v AEPD and Mario Costeja González (2014) Frank v. Gaos (2019)
Other	Garcia v. Google, Inc. (2015) Google LLC v Defteros (2020) Epic Games v. Google (2021) Gonzalez v. Google LLC (2022)

Concepts

Products

Android	Booting process Custom distributions Features Recovery mode Software development
Street View coverage	Africa Antarctica Asia Israel Europe North America Canada United States Oceania South America Argentina Chile Colombia
YouTube	Copyright strike Education Features Moderation Most-disliked videos Most-liked videos Most-subscribed channels Most-viewed channels Most-viewed videos Arabic music videos French music videos Indian videos Pakistani videos Official channel Social impact Suspensions YouTube Premium original programming
Other	Gmail interface Maps pin Most downloaded Google Play applications Stadia games

Documentaries

Books

Google Hacks
The Google Story
Google Volume One
Googled: The End of the World as We Know It
How Google Works
I'm Feeling Lucky
In the Plex
The Google Book
The MANIAC

Popular culture

Google Feud
Google Me (film)
"Google Me" (Kim Zolciak song)
"Google Me" (Teyana Taylor song)
Is Google Making Us Stupid?
Proceratium google
Matt Nathanson: Live at Google
The Billion Dollar Code
The Internship
Where on Google Earth is Carmen Sandiego?

Other

Italics denote discontinued products.

Android

Android Go
- Comparison of products

Software
development

Development tools

Official

Android Runtime (ART)
Software development kit (SDK)
- Android Debug Bridge (ADB)
- Fastboot
- Android App Bundle
- Android application package (APK)
Bionic
Dalvik
Firebase
- Google Cloud Messaging (GCM)
- Firebase Cloud Messaging (FCM)
Google Mobile Services (GMS)
Native development kit (NDK)
Open accessory development kit (OADK)
RenderScript
Skia
AdMob
Material Design
Fonts
- Droid
- Roboto
- Noto
Google Developers

Other

Integrated
development
environments (IDE)

Languages, databases

Virtual reality (VR)

Events, communities

Releases

Derivatives

Devices

Pixel	C Pixel & Pixel XL 2 & 2 XL 3 & 3 XL 3a & 3a XL 4 & 4 XL 4a & 4a (5G) 5 5a 6 & 6 Pro 6a 7 & 7 Pro 7a Fold Tablet 8 & 8 Pro 8a 9, 9 Pro & 9 Pro XL 9 Pro Fold
Nexus	One S Galaxy Nexus 4 10 Q 5 5X 6 6P 7 2012 2013 9 Player
Play edition	HTC One (M7) HTC One (M8) LG G Pad 8.3 Moto G Samsung Galaxy S4 Sony Xperia Z Ultra
Android One other smartphones

Custom
distributions

AliOS
Android-x86
- Remix OS
AOKP
Baidu Yi
Barnes & Noble Nook
CalyxOS
ColorOS
- realme UI
CopperheadOS
EMUI
- Magic UI
Fire OS
Flyme OS
GrapheneOS
Xiaomi HyperOS
- MIUI
- MIUI for POCO
LeWa OS
LineageOS
- /e/
- CrDroid
- CyanogenMod
- DivestOS
- iodéOS
- Kali NetHunter
LiteOS
Meta Horizon OS
MicroG
Nokia X software platform
OmniROM
OPhone
OxygenOS
PixelExperience
Pixel UI
Replicant
Resurrection Remix OS
SlimRoms
TCL UI
Ubuntu for Android
XobotOS
ZUI

Booting and
recovery

APIs

Alternative UIs

Rooting

Lists

Misplaced Pages