Retour au classement

coqui-ai/STT

C++coqui.ai

🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.

sttspeech-to-texttensorflowdeep-learningautomatic-speech-recognitionasrvoice-recognitionspeech-recognitionspeech-recognizerspeech-recognition-api
Croissance des étoiles
Étoiles
2.6k
Forks
299
Croissance hebdomadaire
Issues
89
2.6k2.6k2.6k
18 juil.19 juil.21 juil.22 juil.
Dépôts similaires
khoj-ai/khoj

Your AI second brain. Self-hostable. Get answers from the web or your docs. Build custom agents, schedule automations, do deep research. Turn any online or local LLM into your personal, autonomous AI (gpt, claude, gemini, llama, qwen, mistral). Get started - free.

PythonPyPIGNU Affero General Public License v3.0semantic-searchemacs
khoj.dev
35.9k2.3k
alphacep/vosk-api

Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node

Jupyter NotebookApache License 2.0speech-recognitionasr
15k1.7k
moonshine-ai/moonshine

Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces

C++Otherintent-recognitionstt
moonshine.ai
10.2k539
GetStream/Vision-Agents

Open Vision Agents by Stream. Build voice and vision agents quickly with any model or video provider. Uses Stream's edge network for ultra-low latency.

PythonPyPIApache License 2.0aiai-agents
visionagents.ai
8k668
jianchang512/stt

Voice Recognition to Text Tool / 一个离线运行的本地音视频转字幕工具,输出json、srt字幕、纯文字格式

PythonPyPIGNU General Public License v3.0speechspeech-recognition
pyvideotrans.com
4.7k492
pluja/whishper

Transcribe any audio to text, translate and edit subtitles 100% locally with a web UI. Powered by whisper models!

SvelteGNU Affero General Public License v3.0aiaudio-to-text
whishper-docs.pages.dev
3k177
pannous/tensorflow-speech-recognition

🎙Speech recognition using the tensorflow deep learning framework, sequence-to-sequence neural networks

PythonPyPIOthertensorflowspeech-recognition
2.2k631
akdeb/ElatoAI

Realtime Voice AI with 100+ Models on Arduino ESP32 with Secure Websockets and Edge Functions for AI Toys, Companions, and Devices

TypeScriptnpmOtheraiarduino
elatoai.com
1.8k226
neural-maze/ava-whatsapp-agent-course

Meet Ava, the WhatsApp Agent

PythonPyPIMIT Licenseagentagent-based
1.7k419
mkiol/dsnote

Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.

C++Mozilla Public License 2.0asrsailfishos
1.5k69
waybarrios/vllm-mlx

OpenAI and Anthropic compatible server for Apple Silicon. Run LLMs and vision-language models (Llama, Qwen-VL, LLaVA) with continuous batching, MCP tool calling, and multimodal support. Native MLX backend, 400+ tok/s. Works with Claude Code.

PythonPyPIApache License 2.0apple-siliconaudio-processing
1.4k198
lenML/Speech-AI-Forge

🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.

PythonPyPIGNU Affero General Public License v3.0chatttsssml
1.4k186