Testing ASR for Video Subtitles
A small Korean-language comparison showed why transcript coverage, timestamps, and manual review matter as much as model output.
Topic / Media pipelines
Field notes on separating transcription, translation, vision, and rendering into inspectable media stages.
I treat media work as a chain of artifacts instead of one opaque model call. I record how I separate stages, inspect failures, and keep experiments honest.
Browse the Media pipelines tag →
A small Korean-language comparison showed why transcript coverage, timestamps, and manual review matter as much as model output.
A video-subtitle pipeline that uses Cloudflare for coordination and a local machine for private, CPU-heavy media work.
A follow-cam design that combines face recognition with person tracking, appearance, pose, segmentation review, and hybrid framing.
I adapted public ideas from an AI-assisted broadcast workflow into a cautious follow-cam experiment for my bias.
Reusable skills are helping me give agents clearer working context, spend less time repeating process, and keep the important decisions reviewable.
I separated transcription, translation, and subtitle rendering to make a personal video translation pipeline easier to debug.