๐ข
gitzw.comไธ็บฟไบ๏ผๅ่ฝ้็ปญๆดๆฐไธญ๏ผๅฆๆ้ฎ้ขๆๅ้ฆ่ฏทๅจไธๆนๅ้ฆ/ๅปบ่ฎฎไธญ็ปๆไปฌ็่จใ
โ
Understand open source, starting here
Home
Ranking
Tutorial Hub
Knowledge Base
News
Categories
Collections
๐
EN
โพ
็ฎไฝ
็น้ซ
English
Log in
User center
โพ
๐ค User center
๐ฌ My subscriptions
โ๏ธ My email
๐ Change password
โญ My bookmarks
๐ฌ My feedback
โฉ Log out
Subscribe
โฐ
Home
Ranking
Tutorial Hub
Knowledge Base
News
Categories
Collections
Search
Subscribe
RSS
็ฎไฝ
็น้ซ
English
๐
Search
๐ก
RSS
๐ฅ Hot searches
llm
vuejs
pytorch
langchain
sql
graphql
peg
leetcode
tokio
redis
unity
php8
"speech-processing"
espnet
/
espnet
espnet โ End-to-End Speech Processing Toolkit
Python
โ 9.9k
โ 2.4k
ai-ml
โ
coqui-ai
/
TTS
TTS โ ๐ธ๐ฌ - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
Python
โ 45.8k
โ 6.2k
ai-ml
โ
handy-computer
/
transcribe.cpp
transcribe.cpp โ ggml speech-to-text inference for 16+ model families
C++
โ 1.4k
โ 35
other
+395 โ
โ
mozilla
/
DeepSpeech
DeepSpeech โ DeepSpeech is an open source embedded (offline, on-device) speech-to-text engine which can run in real time on devices ranging from a Raspberry Pi 4 to high power GPU servers.
C++
โ 26.8k
โ 4.1k
ai-ml
โ
babysor
/
MockingBird
MockingBird โ ๐Clone a voice in 5 seconds to generate arbitrary speech in real-time
Python
โ 36.9k
โ 5.2k
ai-ml
โ
CorentinJ
/
Real-Time-Voice-Cloning
Real-Time-Voice-Cloning โ Clone a voice in 5 seconds to generate arbitrary speech in real-time
Python
โ 60k
โ 9.4k
ai-ml
โ
moonshine-ai
/
moonshine
moonshine โ Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces
C++
โ 10.1k
โ 526
ai-ml
+282 โ
โ
langchain-ai
/
open_deep_research
open_deep_research โ
Python
โ 12.2k
โ 1.7k
other
+14 โ
โ
andrewyng
/
context-hub
context-hub โ
JavaScript
โ 13.8k
โ 1.2k
web
โ
MiniMax-AI
/
skills
skills โ
C#
โ 13.1k
โ 1.1k
other
โ
RVC-Project
/
Retrieval-based-Voice-Conversion-WebUI
Retrieval-based-Voice-Conversion-WebUI โ Easily train a good VC model with voice data <= 10 mins!
Python
โ 36.5k
โ 5.1k
data
โ
hankcs
/
HanLP
HanLP โ Natural Language Processing for the next decade. Tokenization, Part-of-Speech Tagging, Named Entity Recognition, Syntactic & Semantic Dependency Parsing, Document Classification
Python
โ 36.5k
โ 10.9k
ai-ml
โ
myshell-ai
/
OpenVoice
OpenVoice โ Instant voice cloning by MIT and MyShell. Audio foundation model.
Python
โ 37k
โ 4.1k
other
โ
explosion
/
spaCy
spaCy โ ๐ซ Industrial-strength Natural Language Processing (NLP) in Python
Python
โ 33.8k
โ 4.7k
ai-ml
โ
FunAudioLLM
/
CosyVoice
CosyVoice โ Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
Python
โ 22.1k
โ 2.6k
ai-ml
โ
snakers4
/
silero-vad
silero-vad โ Silero VAD: pre-trained enterprise-grade Voice Activity Detector
Python
โ 9.6k
โ 804
ai-ml
โ
KoljaB
/
RealtimeSTT
RealtimeSTT โ A robust, efficient, low-latency speech-to-text library with advanced voice activity detection, wake word activation and instant transcription.
Python
โ 10k
โ 850
other
โ
huggingface
/
transformers
transformers โ ๐ค Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Python
โ 162.7k
โ 33.9k
ai-ml
โ
supertone-inc
/
supertonic
supertonic โ Lightning-Fast, On-Device, Multilingual TTS โ running natively via ONNX.
Swift
โ 13.4k
โ 1.4k
mobile
โ
sinaptik-ai
/
pandas-ai
pandas-ai โ Chat with your database or your datalake (SQL, CSV, parquet). PandasAI makes data analysis conversational using LLMs and RAG.
Python
โ 23.7k
โ 2.3k
data
โ
keras-team
/
keras
keras โ Deep Learning for humans
Python
โ 64.2k
โ 19.7k
ai-ml
โ
interviewstreet
/
hiring-agent
hiring-agent โ AI agent to evaluate and score resumes.
Python
โ 2.5k
โ 623
ai-ml
+203 โ
โ
open-mmlab
/
Amphion
Amphion โ Amphion (/รฆmหfaษชษn/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audio, music, and speech generation research and development.
Python
โ 10k
โ 829
other
โ
AudioKit
/
AudioKit
AudioKit โ Audio synthesis, processing, & analysis platform for iOS, macOS and tvOS
Swift
โ 11.4k
โ 1.6k
mobile
โ
2noise
/
ChatTTS
ChatTTS โ A generative speech model for daily dialogue.
Python
โ 39.7k
โ 4.2k
ai-ml
โ
jamiepine
/
voicebox
voicebox โ The open-source AI voice studio. Clone, dictate, create.
TypeScript
โ 44.6k
โ 5.4k
ai-ml
+821 โ
โ
deepset-ai
/
haystack
haystack โ Open-source AI orchestration framework for building context-engineered, production-ready LLM applications. Design modular pipelines and agent workflows with explicit control over retrieval, routing, memory, and generation. Built for scalable agents, RAG, multimodal applications, semantic search, and conversational systems.
MDX
โ 25.9k
โ 2.9k
ai-ml
โ
BVLC
/
caffe
caffe โ Caffe: a fast open framework for deep learning.
C++
โ 34.6k
โ 18.5k
ai-ml
โ
zesterer
/
chumsky
chumsky โ [Chumsky has moved to Codeberg!] Write expressive, high-performance parsers with ease.
Rust
โ 4.5k
โ 210
other
โ
karpathy
/
autoresearch
autoresearch โ AI agents running research on single-GPU nanochat training automatically
Python
โ 91.6k
โ 13.1k
ai-ml
โ
pjreddie
/
darknet
darknet โ Convolutional Neural Networks
C
โ 26.5k
โ 21k
ai-ml
โ
stanfordnlp
/
dspy
dspy โ DSPy: The framework for programmingโnot promptingโlanguage models
Python
โ 36.2k
โ 3.1k
backend
โ
harvard-edge
/
cs249r_book
cs249r_book โ Machine Learning Systems
Python
โ 27k
โ 3.2k
other
+329 โ
โ
google-ai-edge
/
mediapipe
mediapipe โ Cross-platform, customizable ML solutions for live and streaming media.
C++
โ 36.2k
โ 6.1k
ai-ml
โ