Speech Markdown grammar, parser, and formatters for use with JavaScript.
Supported platforms:
- amazon-alexa
- amazon-polly
- amazon-polly-neural
- apple-avspeechsynthesizer
- google-assistant
- ibm-watson
- microsoft-azure
- microsoft-sapi
- w3c
- samsung-bixby
- elevenlabs
Find the architecture here
Platform-specific SSML notes are tracked in docs/platforms. Use npm run docs:update-voices to refresh the auto-generated voice maps in src/formatters/data when vendor credentials are available.
Convert Speech Markdown to SSML for Amazon Alexa
constsmd=require('speechmarkdown-js');constmarkdown=`Sample [3s] speech [250ms] markdown`;constoptions={platform: 'amazon-alexa',};constspeech=newsmd.SpeechMarkdown();constssml=speech.toSSML(markdown,options);The resulting SSML is:
<speak>
Sample <breaktime="3s"/> speech <breaktime="250ms"/> markdown
</speak>Convert Speech Markdown to SSML for Google Assistant
constsmd=require('speechmarkdown-js');constmarkdown=`Sample [3s] speech [250ms] markdown`;constoptions={platform: 'google-assistant',};constspeech=newsmd.SpeechMarkdown();constssml=speech.toSSML(markdown,options);The resulting SSML is:
<speak>
Sample <breaktime="3s"/> speech <breaktime="250ms"/> markdown
</speak>Convert Speech Markdown to SSML for Microsoft Azure with automatic MSTTS namespace injection
constsmd=require('speechmarkdown-js');constmarkdown=`(This is exciting news!)[excited:"1.5"] The new features are here.`;constoptions={platform: 'microsoft-azure',};constspeech=newsmd.SpeechMarkdown();constssml=speech.toSSML(markdown,options);The resulting SSML is:
<speakxmlns:mstts="https://www.w3.org/2001/mstts">
<mstts:express-asstyle="excited"styledegree="1.5">This is exciting news!</mstts:express-as> The new features are here.
</speak>Azure supports 27 express-as styles including emotional styles (excited, disappointed, friendly, cheerful, sad, angry, etc.) and scenario-specific styles (newscaster, customerservice, chat, etc.). See Azure platform documentation for complete details.
Convert Speech Markdown to Plain Text
constsmd=require('speechmarkdown-js');constmarkdown=`Sample [3s] speech [250ms] markdown`;constoptions={};constspeech=newsmd.SpeechMarkdown();consttext=speech.toText(markdown,options);The resulting text is:
Sample speech markdown
You can pass options into the constructor:
constsmd=require('speechmarkdown-js');constmarkdown=`Sample [3s] speech [250ms] markdown`;constoptions={platform: 'amazon-alexa',};constspeech=newsmd.SpeechMarkdown(options);constssml=speech.toSSML(markdown);Or in the methods toSSML and toText:
constsmd=require('speechmarkdown-js');constmarkdown=`Sample [3s] speech [250ms] markdown`;constoptions={platform: 'amazon-alexa',};constspeech=newsmd.SpeechMarkdown();constssml=speech.toSSML(markdown,options);Available options are:
platform(string) - Determines the formatter to use to render SSML. Valid values are:- "amazon-alexa"
- "amazon-polly"
- "amazon-polly-neural"
- "apple-avspeechsynthesizer"
- "google-assistant"
- "ibm-watson"
- "microsoft-azure"
- "microsoft-sapi"
- "w3c"
- "samsung-bixby"
- "elevenlabs"
includeFormatterComment(boolean) - Adds an XML comment to the SSML output indicating the formatter used. Default isfalse.includeSpeakTag(boolean) - Determines if the<speak>tag will be rendered in the SSML output. Default istrue.includeParagraphTag(boolean) - Determines if the<p>tag will be rendered in the SSML output. Default isfalse.preserveEmptyLines(boolean) - keep empty lines in markdown in SSML. Default istrue.escapeXmlSymbols(boolean) - Currently only foramazon-alexaandmicrosoft-azure. Escape XML text. Default isfalse.voices(object) - give custom names to voices and use that in your markdown:{ "platform": "amazon-alexa", "voices": { "Scott": { "voice": { "name": "Brian" } }, "Sarah": { "voice": { "name": "Kendra" } } } }{ "platform": "google-assistant", "voices": { "Brian": { "voice": { "gender": "male", "variant": 1, "language": "en-US" } }, "Sarah": { "voice": { "gender": "female", "variant": 3, "language": "en-US" } } } }
The biggest place we need help right now is with the completion of the grammar and formatters.
- break
- emphasis - strong
- emphasis - moderate
- emphasis - none
- emphasis - reduced
- ipa
- sub
Short-form examples:
(pecan)/'pi.kæn/→<phoneme alphabet="ipa" ph="'pi.kæn">pecan</phoneme>(Al){aluminum}→<sub alias="aluminum">Al</sub>/ˈdeɪtə/→<phoneme alphabet="ipa" ph="ˈdeɪtə">ipa</phoneme>
- address
- audio
- break (time)
- break (strength)
- characters / chars
- date
- defaults (section)
- disappointed
- disappointed (section)
- dj (section)
- emphasis
- excited
- excited (section)
- expletive / bleep
- fraction
- interjection
- ipa
- lang
- lang (section)
- mark
- newscaster (section)
- number
- ordinal
- telephone / phone
- pitch
- rate
- sub
- time
- unit
- voice
- voice (section)
- volume / vol
- whisper
clean- remove coverage data, Jest cache and transpiled files,build- perform all build tasksbuild:ts- transpile TypeScript to ES5build:browser- creates single file./dist.browser/speechmarkdown.jsfile for use in browser,build:minify- creates single file./dist.browser/speechmarkdown.min.jsfile for use in browser,watch- interactive watch mode to automatically transpile source files,lint- lint source files and tests,test- run tests,test:watch- interactive watch mode to automatically re-run tests
Licensed under the MIT. See the LICENSE file for details.