Fully local, private and cross platform Speech-to-Text with LLM Post-processing
-
Updated
Sep 6, 2026 - Rust
Fully local, private and cross platform Speech-to-Text with LLM Post-processing
Open source inference code for Rev's model
SOVA ASR (Automatic Speech Recognition)
Kroko ASR - Speech-to-text
Trained models for automatic speech recognition (ASR). A library to quickly build applications that require speech to text conversion.
End-to-End Vietnamese Speech Recognition using wav2vec 2.0
A multilingual ASR model that can recognize ten Turkic languages—Azerbaijani, Bashkir, Chuvash, Kazakh, Kyrgyz, Sakha, Tatar, Turkish, Uyghur, and Uzbek.
An implementation of RNN-Transducer loss in TF-2.0.
Whisper Studio is an independent desktop interface for OpenAI Whisper.
fine-tune Wav2vec2. an ASR model released by Facebook
Summarization, topic generation using GPT3
ComfyUI nodes for Qwen3-ASR (0.6B/1.7B) and ForcedAligner. Supports high-accuracy ASR and language identification for 52 languages/dialects, including 22 Chinese dialects and various English accents. Features word-level timestamps, long audio transcription, and VRAM-optimized inference.
Quartznet implementation on pytorch [https://arxiv.org/abs/1910.10261]
Collection and resources for Bulgarian Corpus, Datasets and Models used in ASR, TTS or NLP tasks together with the links of corresponding tools/apps.
Nemotron ASR rewrite to GGML
Automatic speech recognition (ASR) for Indonesian language built by using HTK and Julius. Web interface is built using Node.js.
Many ASRs under one roof. With Benchmarking... answering the question. What is the best ASR for my dataset?
Audio Preprocessing and finetuning of wav2vec2-large-xlsr model on AI4D Baamtu Datamation - Automatic Speech Recognition in WOLOF Data.
Deepspeech ASR Model for the Catalan Language
To associate your repository with the asr-model topic, visit your repo's landing page and select "manage topics."