Skip to main content

Overview

Some interviews and courses already have captions burned into the picture. Open Hard-subtitle OCR (Beta) in the Learn panel, pick the caption language, and tap Start. You get a timestamped list without running speech-to-text.
Hard-subtitle OCR settings
Timestamped captions after OCR

Pain Point

The picture already has clear captions, but speech-to-text trips on accents and background noise. I want the words that are already on screen.

Key Features

  • Hard-subtitle OCR (Beta) tab at the top of the Learn panel
  • Chinese, English, Japanese, French, German, and Spanish
  • Timestamped caption list on the right after recognition, for review and seeking
  • Reads on-screen text only; no speech required

Highlights

When captions are already in the frame, reading them is faster and steadier than listening.

Use Cases

  • Interviews and courses with a clean lower-third caption area
  • Heavy accents or noisy rooms where speech-to-text fails
  • Quick checks of on-screen wording against timestamps

Tips

  • Extra text near the captions lowers accuracy. Stylized or title-heavy videos do worse
  • This is still Beta. If results are messy, fall back to speech-to-text or a cleaner source

Custom Transcription Engine

Smart Subtitle Segmentation

Chapter Deep Reading


Upgrade to Pro: unlock more subtitle and study tools → Pricing