stt

32 projects share this GitHub topic

stt — khoj ★36.1ksttvosk-api — ★15kVision-Agents — ★8kstt — ★4.7kwhishper — ★3.1kSTT — ★2.6ktensorflow-speech-recognition — ★2.2kElatoAI — ★1.9kava-whatsapp-agent-course — ★1.7kdsnote — ★1.6kvllm-mlx — ★1.5kSpeech-AI-Forge — ★1.4kopen-speech-corpora — ★1.4kSoniTranslate — ★1.4kgp.nvim — ★1.3ksokuji — ★1klobe-tts — ★798TTS-Voice-Wizard — ★795whisper.unity — ★749mlx-audio-swift — ★736mlx-omni-server — ★734cheetah — ★669sonus — ★638Starmoon — ★549vosk-browser — ★527leopard — ★482CrispASR — ★482skills — ★395onnx-asr — ★353LangHelper — ★349vakyansh-models — ★327flowflow — ★160vosk-api★ 15kVision-Agents★ 8kstt★ 4.7kwhishper★ 3.1kSTT★ 2.6ktensorflow-speech-recogn…★ 2.2kElatoAI★ 1.9kava-whatsapp-agent-cours…★ 1.7kdsnote★ 1.6kvllm-mlx★ 1.5kSpeech-AI-Forge★ 1.4kopen-speech-corpora★ 1.4kSoniTranslate★ 1.4kgp.nvim★ 1.3ksokuji★ 1klobe-tts★ 798TTS-Voice-Wizard★ 795whisper.unity★ 749mlx-audio-swift★ 736mlx-omni-server★ 734cheetah★ 669sonus★ 638Starmoon★ 549vosk-browser★ 527leopard★ 482CrispASR★ 482skills★ 395onnx-asr★ 353LangHelper★ 349vakyansh-models★ 327flowflow★ 160

Lines connect members that are measurably related to each other. Dot size reflects stars.

🧬 Members
khoj
Your AI second brain. Self-hostable. Get answers from the web or your docs. Build custom agents, schedule…
★ 36.1k
vosk-api
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
★ 15k
Vision-Agents
Open Vision Agents by Stream. Build voice and vision agents quickly with any model or video provider. Uses…
★ 8k
stt
★ 4.7k
whishper
Transcribe any audio to text, translate and edit subtitles 100% locally with a web UI. Powered by whisper…
★ 3.1k
STT
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so…
★ 2.6k
tensorflow-speech-recognition
🎙Speech recognition using the tensorflow deep learning framework, sequence-to-sequence neural networks
★ 2.2k
ElatoAI
Realtime Voice AI with 100+ Models on Arduino ESP32 with Secure Websockets and Edge Functions for AI Toys,…
★ 1.9k
ava-whatsapp-agent-course
Meet Ava, the WhatsApp Agent
★ 1.7k
dsnote
Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and…
★ 1.6k
vllm-mlx
OpenAI and Anthropic compatible server for Apple Silicon. Run LLMs and vision-language models (Llama,…
★ 1.5k
Speech-AI-Forge
🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a…
★ 1.4k
open-speech-corpora
💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies
★ 1.4k
SoniTranslate
Synchronized Translation for Videos. Video dubbing
★ 1.4k
gp.nvim
Gp.nvim (GPT prompt) Neovim AI plugin: ChatGPT sessions & Instructable text/code operations & Speech to text…
★ 1.3k
sokuji
Real-time two-way speech translation for bilingual meetings — auto-detects the spoken language and…
★ 1k
lobe-tts
🎤 Lobe TTS - A high-quality & reliable TTS/STT library for Server and Browser
★ 798
TTS-Voice-Wizard
Speech to Text to Speech. Song now playing. Sends text as OSC messages to VRChat to display on avatar.…
★ 795
whisper.unity
Running speech to text model (whisper.cpp) in Unity3d on your local machine.
★ 749
mlx-audio-swift
A modular Swift SDK for audio processing with MLX on Apple Silicon
★ 736
mlx-omni-server
MLX Omni Server is a local inference server powered by Apple's MLX framework, specifically designed for Apple…
★ 734
cheetah
On-device streaming speech-to-text engine powered by deep learning
★ 669
sonus
:speech_balloon: /so.nus/ STT (speech to text) for Node with offline hotword detection
★ 638
Starmoon
A conversational, AI device + software framework for companionship, entertainment, education, healthcare, IoT…
★ 549
vosk-browser
A speech recognition library running in the browser thanks to a WebAssembly build of Vosk
★ 527
leopard
On-device speech-to-text engine powered by deep learning
★ 482
CrispASR
C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B…
★ 482
skills
Collections of skills for building with ElevenLabs
★ 395
onnx-asr
A lightweight Python package for Automatic Speech Recognition using ONNX models
★ 353
LangHelper
Striving to create a great Application with full functions of learning languages by ChatGPT, TTS, STT and…
★ 349
vakyansh-models
Open source speech to text models for Indic Languages
★ 327
flowflow
Voice notes for iPhone and macOS - 100% Rust, Dioxus, local-first (SQLite + LanceDB + RIG)
★ 160
🔗 Related families

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.