Designed for one very specific user: me.
I watch a lot of English content — lectures, design talks, meetings, YouTube. My English is fine, but listening at full speed for an hour is real cognitive load, and the moment something important lands, I'm pausing and scrubbing back. I wanted subtitles that just exist, for everything, without asking whether the platform provides them.
The tools I tried all had the same three gaps:
- Platform-bound — browser extensions translate YouTube's caption track. No track, no captions. Meetings, livestreams, and podcasts are out.
- Text-layer translation — most tools transcribe first, then translate the text, adding a full extra step of latency.
- Traditional Chinese as an afterthought — output tuned for Simplified Chinese, with mainland phrasing like 人工智能 where Taiwan says 人工智慧.
So the product definition wrote itself: capture the system's own audio, translate the speech itself rather than a transcript, and make Taiwan-flavored Traditional Chinese the first-class output — not a checkbox.