WeAreDevelopers LIVE Dec 16, 2024

Building a Browser-Based Karaoke Game with Web Speech API

Ana Rodrigues

Building "useless" side projects is the ultimate antidote to developer burnout. Watch Ana build a browser-based karaoke game to explore the Web Speech API and rediscover her frontend passion.

Pause
Mute Enter Fullscreen
#1 about 3 min

Motivation for building a gamified karaoke web app

The desire to create a browser-based karaoke game solving the lack of available songs for specific bands.

#2 about 2 min

Exploring the features of speech recognition and synthesis

How speech recognition and synthesis functionality can be utilized for continuous detection and forms input.

#3 about 3 min

Browser support and privacy concerns for speech recognition

Why features like native speech recognition often require a connected web server and lack consistent cross-browser support.

#4 about 5 min

Demonstrating voice navigation and synthesis bookmarklets

Live experiments showing how browser bookmarklets can turn text to speech or enable voice-activated page navigation.

#5 about 4 min

Implementing karaoke game logic and lyrics synchronization

The technical steps to initialize speech recognition, load timed lyric subtitles, and evaluate interim results.

#6 about 7 min

Testing the karaoke application with live singing

A live demonstration verifying text matching accuracy by comparing spoken word inputs against expected mapped lyrics.

#7 about 3 min

Current limitations of browser-based voice integrations

An overview of alternative polyfills, community projects, and the common voice dataset to improve voice technology reliance.

#8 about 2 min

Designing accessible voice content and interfaces

Actionable advice for structuring conversational content and handling error recovery in voice interaction designs.

#9 about 4 min

Combating burnout through unproductive side projects

Why experimenting with unmonetized ideas helps front-end developers learn new technical logic and rediscover creative programming joy.

Matching moments

2:30 min

Synthesizer karaoke and final presentation wrap up

Rowdy Rabouw Rowdy Rabouw · LIVE

1:30 min

Overview of voice user interfaces on the web

Tobias Münch Tobias Münch · World Congress 2024

48 sec

Core components of the web speech API ecosystem

Tobias Münch Tobias Münch · World Congress 2024

3:16 min

Voice AI architecture and interactive kiosk demo

Lee Boonstra · LIVE

4:24 min

Building synthesizers and music sequencers using web audio APIs

Chris Heilmann +2 · LIVE

58 sec

Uncensored voice assistants and neural controlled web browser accessibility

Chris Heilmann +2 · LIVE

Upcoming sessions on this topic

Open session

World Congress 2026 North America

September 24, 2026 · 11:20–11:25

Outdoor Stage

Finding the Edges: Testing, Evaluating, and Monitoring Voice AI Agents Before Your Users Do

Matt Wyman

CEO of Okareo

Matt Wyman
Open session

World Congress 2026 North America

September 25, 2026 · 13:30–14:00

Stage 5

I Built My Own AI Wearable (So You Don't Have To — But You'll Want To)

Mike Chambers

AWS -Specialist Developer Advocate, Machine Learning

Mike Chambers
Open session

World Congress 2026 North America

September 25, 2026 · 16:50–17:20

Stage 1

Honey, look! I vibe-coded an OS!

Ian Smith

CTO of LYOS

Ian Smith
Open session

World Congress 2026 North America

September 25, 2026 · 15:30–16:00

Stage 6

Small LLM in your Browser: Huge Opportunities for Web Applications

Daniel Ostrovsky

UI/UX Architect at Payoneer | AI Architect | Full Cycle Development Expert | Public Speaker | Open Source Contributor |

Daniel Ostrovsky
Open session

World Congress 2026 North America

September 25, 2026 · 10:20–10:50

Stage 5

Physical AI: 5 Things You Can Build That Aren't Another Chatbot

Vini Senger

Senior Technical Evangelist for Startups

Vini Senger
Open session

World Congress 2026 North America

September 25, 2026 · 11:00–11:30

Outdoor Stage

Vibe Coding Accessibility

Karl Groves

Focused on actively fixing accessibility

Karl Groves