A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
-
Updated
May 24, 2024 - Python
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System, End-to-End Speech Translation and Keyword Spotting. Won NAACL2022 Best Demo Award.
A PyTorch-based Speech Toolkit
🤖 wukong-robot 是一个简单、灵活、优雅的中文语音对话机器人/智能音箱项目,支持ChatGPT多轮对话能力,还可能是首个支持脑机交互的开源智能音箱项目。
Production First and Production Ready End-to-End Speech Recognition Toolkit
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
pytorch-kaldi is a project for developing state-of-the-art DNN/RNN hybrid speech recognition systems. The DNN part is managed by pytorch, while feature extraction, label computation, and decoding are performed with the kaldi toolkit.
Lingvo
The official repository of the Eesen project
OpenAI Whisper ASR Webservice API
DELTA is a deep learning based natural language and speech processing platform.
Silero Models: pre-trained speech-to-text, text-to-speech and text-enhancement models made embarrassingly simple
This is a python API which allows you to get the transcript/subtitles for a given YouTube video. It also works for automatically generated subtitles and it does not require an API key nor a headless browser, like other selenium based solutions do!
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
SincNet is a neural architecture for efficiently processing raw audio samples.
A Python wrapper for Kaldi
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
Add a description, image, and links to the asr topic page so that developers can more easily learn about it.
To associate your repository with the asr topic, visit your repo's landing page and select "manage topics."