CAPSTONE 1 · ADD-A-FEATURE

Karaoke Mode

A sing-along feature for Spotify, designed around the person who says no. Vocal attenuation, synchronized lyrics, and a permission model that treats refusal as a first-class path.

ROLE — UX DESIGNER, SOLE PRACTITIONER PROGRAM — DESIGNLAB UX ACADEMY (75 HRS) WINDOW — DEC 2025 → APR 2026 METHOD — MODERATED LO-FI (N=5) → ASYNC HI-FI
01

Origin

The feature was selected by one criterion: personal stake. The mentor's directive — choose something the designer cares about and would use — resolved a brainstorm into a single candidate. Karaoke was recent, real, and unserved. The brief is three paragraphs. The mentor's approval note calls it a simple brief and a great idea. Both statements are correct.

[ MENTOR NOTE · 2025-12-29 ]"Rather simple, to be honest" — and approved anyway, as a great idea worth building.
Project brief submission
FIG.01 — PROJECT BRIEF · SUBMITTED 12/29/2025
02

Process, taken straight

A first capstone is for learning the process, not renegotiating it. The curriculum ran on rails by choice: brief, competitive analysis, interview guide, interviews, synthesis, definition, wireframes, mockups, testing. Deviation is a tool. It was deliberately left in the drawer.

03

The competitive gap

No major streaming service ships an integrated karaoke mode. YouTube wins the use case by accident of structure — music videos with lyric overlays uploaded by third parties. The gap is real; the analysis confirmed rather than redirected. The mentor's extension note — audit competitors for accessibility features — is logged as method for future analyses.

Streaming competitor SWOT analysis
FIG.02 — COMPETITIVE SWOT · TIDAL, APPLE MUSIC, YOUTUBE, AMAZON MUSIC
04

Research under friction

The interview guide went through one same-day iteration loop: mentor question-adds, changes bolded for review speed, approved within hours. The field period did not move at that speed. Recruiting and scheduling moderated sessions stretched across weeks. The response was instrumental, not heroic: when scheduling fails, change the instrument. The async pivot entered the toolkit here and never left it.

[ MENTOR NOTE · 2026-01-12 ]Needs iteration — five question-adds, marked up for a fast re-read. V2 approved the same day.
Interview guide V2
FIG.03 — INTERVIEW GUIDE V2
05

Honest data

Most recruited participants were not Spotify users. The submission note states the limitation plainly; the mentor's approval resolves it — interest in karaoke, not platform loyalty, is the variable under test. The panel was accepted as imperfect and useful. Pretending otherwise would have produced prettier data and a weaker study.

Affinity map from user interviews
FIG.04 — AFFINITY MAP · 2026-03-10
06

The insight: vocal shame

Interviews confirmed the founding hypothesis and sharpened it. Users love music, want to sing, and restrict singing to safe spaces — the car, the shower. The inhibitor is not capability; it is the fear of being heard. The design brief that emerged: build a low-pressure environment that bridges desire and fear. Solo-first, not gated behind group sessions — a ruling the mentor endorsed explicitly.

Personas: The Social Curator and Closet Vocalist
FIG.05 — PERSONAS · THE SOCIAL CURATOR & THE "CLOSET" VOCALIST · 2026-03-10
07

Design for the person who says no

Microphone-permission denial was specified as a first-class path from conception, not patched in after testing. The reasoning is layered: privacy anxiety deserves respect; loud rooms corrupt pitch scoring; microphones break; a user on a call may still want lyrics. Real-world karaoke does not grade its singers. The no-mic path mimics the YouTube experience — adjustable vocals, synchronized lyrics, no judgment. Flexibility was the principle; the permission fork is its implementation.

Lo-fi wireframes: mic permission fork
FIG.06 — LO-FI FRAMES · MIC PERMISSION FORK, SING-ALONG AND NO-MIC VARIANTS
08

A feature from a participant's mouth

The signature control — partial vocal attenuation rather than binary mute — originated in an interview. The original vocal, dipped but present, functions as an audio cue for lyric recall. If the system already separates stems, attenuation is free. The feature was specified, then nearly lost: a pre-test document review on 04/01 caught its omission from the hi-fi build and restored it. The review ritual paid for itself in one catch.

[ MENTOR NOTE · 2026-03-11 ]Approved — don't reinvent the wheel; a community Spotify UI kit exists, use it.
Feature set including the Vocal Dip slider
FIG.07 — FEATURE SET · VOCAL DIP SLIDER
09

Borrowed wheel, hand-matched

The screens were built inside Spotify's design language, not against it. A community UI kit supplied components; the real lyric-tab screenshot supplied truth. Spacing, tracking, and kerning were matched by hand-overlay against the live app. One editorial swap: the demo track moved from The Beatles to Creed — partly meme, partly method, because studying the live app surfaced the music-video background pattern that real karaoke convention expects, and the new track demonstrated it.

Hi-fi mockups V5 in Spotify's design language
FIG.08 — HI-FI MOCKUPS V5 · SPOTIFY LYRIC TAB REFERENCE AT FAR LEFT · 2026-04-04
10

The animation wall

The mockups earned a challenge from the mentor: without motion, the prototype reads as a slideshow. What followed is documented in five moves. A tutorial-built slider failed — full-max or nothing. The mentor's workaround, screen-recording the live app into Figma, was rejected on technical grounds: recordings lack the green active-word treatment that defines the karaoke read. A captions-style approach was attempted. Figma Make — then brand new — was trialed and abandoned within a day. The prototype shipped without full motion, deferred to a pre-portfolio pass, with the mentor's concurrence. Perfectionism lost the schedule argument, on purpose.

Iteration variants of the volume and pitch stack
FIG.09 — ITERATIONS · ALTERNATE VOLUME/PITCH STACK VARIANTS · 2026-04-09 → 04-14
Mentor thread on the animation workaround
FIG.10 — MENTOR THREAD · ANIMATION WORKAROUND EXCHANGE · 04/11–04/13
11

Testing, two instruments

The lo-fi round ran as designed: five moderated remote think-aloud sessions, ~30 minutes each. Users navigated entry points without help; the pitch-tracking representation and the permissions-denied flow needed work — the missing interstitial screens were built in response. The hi-fi round ran on the async instrument after scheduling friction returned. A Maze deployment also ran; its data no longer exists — the free tier permits one project, and creating the next capstone's test destroyed this one's record. The lesson matured two capstones later into owned instrumentation.

Usability test results
FIG.11 — LO-FI REPORT · HI-FI FORM RESULTS
12

What it seeded

The augmented workflow began here: stream-of-consciousness drafts plus data plus rubric, edited into deliverables by machine, verified and submitted by hand. The mentor named it on the record — a good mix between handmade and AI work — and twice recommended the tool that would eventually replace the whole stack. Three practices left this project alive: the async pivot, ship-over-perfectionism, and the habit of auditing tools by their failures. All three became infrastructure.