Courses from 1000+ universities
AI got cheap enough that Duolingo’s most expensive plan may not survive it. I read the earnings call transcript and opened the app to see what is actually changing for learners.
600 Free Google Certifications
Artificial Intelligence
Data Science
Cybersecurity
L'Italiano nel mondo
Introduction to HTML5
Umano Digitale
Organize and share your learning with Class Central Lists.
View our Lists Showcase
Learn Speech & Audio, earn certificates with free online courses from Stanford, Alexander Amini, IIT Kharagpur, Higher School of Economics and other top universities around the world. Read reviews to decide if a class is right for you.
Build a C# media-processing pipeline that downloads files from Google Drive, extracts audio with FFmpeg, transcribes with Whisper, and generates structured summaries.
Build TypeScript applications that transcribe local and remote videos with OpenAI Whisper, including large-file processing and video summarization.
Build scalable video transcription systems in Java with OpenAI GPT-4o mini, including large-file processing, remote video scraping, and optimization.
Build a C#/.NET application that transcribes audio and video files with OpenAI APIs, handles long-form media, downloads remote files, and summarizes results.
Build pseudo-realtime browser transcription by chunking microphone input and using Whisper’s segment timing, prompts, and language options.
Build a TypeScript app that turns audio and remote video into structured text with OpenAI Whisper, Howler.js, and real-time browser audio.
Build a TypeScript video transcription workflow with OpenAI’s Whisper API, including remote media handling and retry logic.
①滴滴语音简介 ②ICASSP论文解读 | 整合频谱和空间信息的鲁棒波束形成技术 ③ICASSP论文解读 | 基于图卷积编码器的摘要生成方法
An Online General Transcription Course for Both Beginners and Experts.
Master AI videos creation step-by-step, mastering top AI tools including Seedance, Google OMNI & VEO, Kling & Midjourney
Master the Skills Needed to Professionally Transcribe Any Audio File Accurately and in Less Time
Build AI apps with open source models from Hugging Face Hub: chatbots, translation, speech recognition, text-to-speech, zero-shot image segmentation, and deployment with Gradio and Spaces.
Build multimodal AI applications combining text, speech, and images, using DALL·E, Sora, Whisper, and Llama, with cross-modal retrieval and web apps in Flask and Gradio.
Learn to transcribe recorded and live audio and video with Amazon Transcribe, exploring its architecture, built-in features, and demonstrations in the AWS Management Console.
In this guided project, you will learn how to use Amazon Transcribe to generate your own audio transcripts quickly.
Get personalized course recommendations, track subjects and courses with reminders, and more.