Local captions

Speech is processed by your Caption Local service. Choose your audio input, start captions, then Stop and download your transcript.

Checking model…

Choose your audio and output

Start with six-second captions. Shorter intervals can reduce accuracy and capacity. English translation means speech translated into English.

Capture and save

Ready

Stop and let pending audio finish before downloading. Audio is not saved; closing the tab can lose pending audio and unsaved text.

Send captions to caption.ninja

Optional: caption text will travel through the selected relay (caption.ninja’s public relay by default). Audio stays on the inference host. Paste the editor’s private automatic-caption room below. The editor reviews these captions before publishing to its separate public room.

Use a private relay

A private relay needs its own room token, separate from the inference service token. The editor needs the source viewing token and output publishing token. Links carry the relay address, never tokens. Turn sharing off before changing these settings.

Sharing off

Connection settings and capture diagnostics