Skip to content

Latest commit

History

15 Commits

Folders and files

NameName
Last commit message
Last commit date

Repository files navigation

Google Cloud Speech + Web Audio API

Speech recognition and synthesis using the Google Cloud Speech APIs integrated with the Web Audio API for microphone input and playback directly in the browser.

For now authorization only works with an API Key, you can create one at https://console.cloud.google.com/apis/credentials. Make sure you restrict it to only Cloud Speech-to-Text and Cloud Text-to-Speech APIs.

Here's a demo page.

Usage

npm install google-cloud-speech-webaudio

Speech Recognition

import{GoogleSpeechRecognition}from'google-cloud-speech-webaudio';constGOOGLE_API_KEY='...';// Optionally define a regional endpoint: https://cloud.google.com/speech-to-text/docs/endpointsconstEU_ENDPOINT='https://eu-speech.googleapis.com'// If this parameter is not defined, the US endpoint will be used by default.// const US_ENDPOINT = 'https://us-speech.googleapis.com'constspeechRecognition=newGoogleSpeechRecognition(GOOGLE_API_KEY,EU_ENDPOINT);// start recording microphone audio to a buffer.// on first run it will request microphone permission.awaitspeechRecognition.startListening();// stop recording your audio, send the buffer to Google for transcriptionconstresult=awaitspeechRecognition.stopListening();

Speech Synthesis

import{GoogleSpeechSynthesis}from'google-cloud-speech-webaudio';constGOOGLE_API_KEY='...';constspeechSynthesis=newGoogleSpeechSynthesis(GOOGLE_API_KEY);// make the API call and play the produced speech audio bufferawaitspeechSynthesis.speak('hello world');

Author

Andrei Gheorghe

About

Speech recognition and synthesis using the Google Cloud Speech APIs integrated with the Web Audio API for microphone input and playback directly in the browser.

Resources

Stars

19 stars

Watchers

1 watching

Forks

Releases

Packages

Used by

Contributors

Languages