A React-Native Bridge for the Google Dialogflow (API.AI) SDK
-
Updated
May 4, 2023 - JavaScript
A React-Native Bridge for the Google Dialogflow (API.AI) SDK
A speech-to-text framework and bot for Discord. Take control of your Discord server using speech and voice commands. Can also be useful for hearing impaired and deaf people.
Typing to Listen at the Cocktail Party: Text-Guided Target Speaker Extraction (LLM-TSE)
JS speech analyzer for fast speech analysis and labeling
Extract formant features such as frequency, power, energy, and bandwidth of formants at syllable or word level from audio sources in a web browser using WebAudio API.
A speech-to-text bot for discord with music commands and more using NodeJS. Ideally for controlling your Discord server using voice commands, can also be useful for hearing-impaired people.
a simple speech recognition app using the Web Speech API Interfaces
(1st place at HopHacks) A dynamic webVR memory palace for speech training, utilizing natural language processing and Google Streetview API
Speech Recognition and Voice Activity Detection using a Convolutional Neural Network Architecture built with Tensorflow.js
Online web based mel-spectrum, power spectrum, FFT analyzer for speech and music processing
Creating music with the browser. From Speech to Musical Instrument
DeepSpeech runtime transcript NodeJs native client
On-device wake-phrase detection in the browser — WebAssembly + SIMD128, ~275 KB total (runtime + model), no server. Same runtime as voxrt-wake-word-{android,ios,linux}. Free demo tier of the VoxRT wake-word family.
On-device 14-way keyword spotting in the browser — WebAssembly + SIMD128, ~1.5 MB total (runtime + model), no server. Same runtime as voxrt-kws-{linux,android,ios}.
The Duolingo for speech therapy–we're making better speech accessible to all, without the crazy costs.
Simplifying the Speech Synthesis and Speech Recognition engines for Javascript. Listen for commands and perform callback actions, make the browser speak and transcribe your speech!
A voice-controlled PDF viewer app
Speech To Code is Google Chrome Extension to convert Speech into Code.
VoiceBridge: Assistive technology designed to transform atypical speech patterns into clear, intelligible voice output, empowering individuals with speech impairments.
Privacy-first, 100% in-browser Voice Impression & Pitch Analysis web app powered by Transformers.js, WebGPU, and ONNX. 瀏覽器端純本地聲學與語音印象分析工具。
To associate your repository with the speech-processing topic, visit your repo's landing page and select "manage topics."