Courses from 1000+ universities
AI got cheap enough that Duolingo’s most expensive plan may not survive it. I read the earnings call transcript and opened the app to see what is actually changing for learners.
600 Free Google Certifications
Artificial Intelligence
Data Science
Cybersecurity
L'Italiano nel mondo
Introduction to HTML5
Umano Digitale
Organize and share your learning with Class Central Lists.
View our Lists Showcase
Learn Speech & Audio, earn certificates with free online courses from Stanford, Alexander Amini, IIT Kharagpur, Higher School of Economics and other top universities around the world. Read reviews to decide if a class is right for you.
Diagnose audio model failures in production: calculate Word Error Rate and F1-scores, run segment and qualitative error analysis, and correlate spectrograms with failure patterns.
Explore cutting-edge audio-visual analysis techniques, from enhanced speech recognition to object tracking, dynamic highlighting, and advanced video processing using Kalman Filters.
Explore the different system components involved in acoustic communication using a representative set of practical applications.
Introducción práctica a Amazon Transcribe para convertir voz en texto mediante transcripción en tiempo real, por lotes y vocabulario personalizado.
Master SpeechLMs, neural audio codecs, diffusion & flow matching to build real-time agentic voice AI systems
Generate AI voiceovers in ElevenLabs Creator: produce natural narrations, adjust voice settings, build short dubbing projects, choose export options, and handle consent and disclosure.
Create text, images, PPTs, audio, videos, blogs, quizzes, and more using ChatGPT, Gemini, and NotebookLM AI Tools.
Voice cloning, speech to text, text to speech, and AI agents — 5 real Python projects on Mistral's free plan
Learn multimodal AI by building applications that combine text, speech, images, and video using Whisper, DALL·E, Sora, Llama, Granite, Mixtral, Flask, and Gradio through hands-on labs and real project
This four-week course provides a hands-on deep dive into the full spectrum of modern AI capabilities. You will master Image-to-Text (Vision), Text-to-Speech (TTS), and Speech-to-Text (Whisper), before
Microsoft Azure AI Fundamentals: Explore natural language processing
Learn the fundamentals of deep learning with PyTorch! This beginner friendly learning path will introduce key concepts to building machine learning models in multiple domains include speech, vision, and natural language processing.
Build Python speech applications that recognize audio, analyze sentiment, summarize podcasts, and support a real-time voice assistant.
In this lab you take audio from the client's microphone and stream it to a Java servlet. The Java servlet passes the data to the Cloud Speech API, which then streams transcriptions back to the servlet.
In this lab you will create a series of audio files using the Text-to-Speech API, then listen to them to compare the differences.
Get personalized course recommendations, track subjects and courses with reminders, and more.