✓ Link copied to clipboard!
Speech-to-text with open models | Sri AI
AI: Language & Voice AI

Speech-to-text with open models | Sri AI

(0 reviews)
Intermediate 3 views

Course overview

Turning speech into text offline, on a server, a laptop or a phone, with Whisper and its faster cousins.

Level: Intermediate  ·  Mode: Evenings, online

Who this course is for

Developers adding voice input to apps, assistants and transcription tools.

What you will learn

  • Run Whisper on CPU and GPU
  • Transcribe live audio in real time
  • Pick the right model size for a device
  • Measure accuracy with word error rate

Syllabus

  1. Module 1: How speech recognition works
  2. Module 2: Whisper and faster-whisper
  3. Module 3: whisper.cpp on CPU and phones
  4. Module 4: Small offline models with Vosk
  5. Module 5: Streaming and voice activity detection
  6. Module 6: Measuring accuracy

Final project

Every module ends in hands-on practice, and the course ends with a project you build and present. Your certificate names that project.

Before you start

Python for AI.

Open-source tools you will use

⭐ Rate This Course