-
OWSM v3.1: Better and Faster Open Whisper-Style Speech Models based on E-Branchformer
Paper • 2401.16658 • Published • 14 -
Reproducing Whisper-Style Training Using an Open-Source Toolkit and Publicly Available Data
Paper • 2309.13876 • Published • 2 -
OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning
Paper • 2506.00338 • Published • 11
Quijada
RaulQF
·
AI & ML interests
Passionate about the intersection of Artificial Intelligence, Machine Learning, and Media Technology, I specialize in:
Computer Vision – Advanced image and video analysis for content understanding and automation.
Natural Language Processing – Extracting insights from text using cutting-edge NLP models.
Audio & Speech Processing – Enhancing accessibility through speech recognition, transcription, and dubbing solutions.
AI for Media & Broadcasting – Automating workflows for TV companies, from metadata extraction to content monitoring.
Optimization & Deployment – Scalable AI solutions leveraging GPU acceleration and cloud-based architectures.
Recent Activity
upvoted a paper 2 days ago
OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and
Cleaning updated a collection 2 days ago
ASR updated a collection 2 days ago
ASR