What Did I Just Say? Self-Listening for Full-Duplex Speech Models Paper • 2609.05592 • Published 10 days ago • 23
deepseek-ai/DeepSeek-R1-Distill-Qwen-32B Text Generation • 33B • Updated Feb 24, 2025 • 505k • • 1.61k
SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing Paper • 2110.07205 • Published Oct 14, 2021 • 6
SpeechT5 Collection The SpeechT5 framework consists of a shared seq2seq and six modal-specific (speech/text) pre/post-nets that can address a few audio-related tasks. • 8 items • Updated May 1, 2025 • 28
reazon-research/reazonspeech-nemo-v2 Automatic Speech Recognition • Updated Feb 13, 2024 • 1.46k • 39
Running on Zero Agents 1.22k ChatGPT Prompt Generator 👨 1.22k Generate ChatGPT prompts from a given persona
Running on CPU Upgrade 14.1k Open LLM Leaderboard 🏆 14.1k Track, rank and evaluate open LLMs and chatbots
Running on CPU Upgrade Agents Featured 1.46k Open ASR Leaderboard 🏆 1.46k Explore and compare speech recognition model performance