AI: Language & Voice AI
Audio AI: classification, speaker ID & diarisation | Sri AI
Advanced
2 views
Course overview
Beyond words: recognising sounds, telling speakers apart and working out who spoke when in a meeting recording.
Level: Advanced · Mode: Evenings, online
Who this course is for
Engineers working with recordings, call centres and monitoring systems.
What you will learn
- Classify sounds and audio events
- Identify and verify speakers
- Split a recording by speaker
- Build meeting-minute pipelines
Syllabus
- Module 1: Audio features
- Module 2: Sound classification
- Module 3: Speaker identification
- Module 4: Diarisation
- Module 5: Meeting transcription pipeline
Final project
Every module ends in hands-on practice, and the course ends with a project you build and present. Your certificate names that project.
Before you start
Speech-to-text with open models.
Open-source tools you will use
Adapts the Hugging Face Audio Course (Apache-2.0).