Help Centre

Breaking down a video

Caption-based breakdowns, on-device speech-to-text, and what the local model costs in download and time.

Applies to Novus Learn 0.1.0

Two routes to a transcript

Copy link

If the video already publishes captions, Novus Learn uses them: it is fast, accurate and needs no model. If it does not, you can transcribe an audio or video file you already have using a speech model that runs inside this browser, in a same-origin worker. The file is never uploaded.

What the local model costs

Copy link
  • A one-time model download the first time you transcribe. It is large, and it is cached afterwards.
  • Real processing time on your own hardware, which varies enormously between devices.
  • Nothing else. There is no per-minute charge, no key, and no upload.

How a transcript becomes study material

Copy link

Once there is a transcript, the pipeline is the one every other source uses. Claims are extracted from the transcript's own sentences, and each one carries its timestamp rather than a line number, so inspecting the evidence takes you to the moment in the recording where it was said.

Accuracy, stated plainly

Copy link

Automatic transcription makes mistakes, especially with names, technical terms, overlapping speech and strong background noise. Because every claim links back to its timestamp, a suspicious sentence can be checked against the audio in a few seconds. Treat the transcript as a working copy, not a record.

Consent version 2026-08-21.1

Cookie preferences