Skip to content
Endless Documentation

Transcripts

Turn any recording, podcast, or video link into searchable text — with speakers, timestamps, and clip-ready exports — and repurpose it into clips, posts, and subtitles.

A transcript is the searchable, editable text version of an audio or video file. Add a recording or paste a link, and Endless returns the full text — at high quality, broken into speakers and timestamped down to the word — so you can read it, search it, and turn it into something else.

Most people don't want the text for its own sake. They want what comes next: the clip, the thread, the subtitle file, the quote they can't find by scrubbing a two-hour stream. A transcript is the fastest path from "we recorded it" to "we published it."

The clip can be literal. Where the built-in clip skill is switched on — it's a skill, so a workspace or an individual can have it off — chat can cut a moment out of the recording, reframe it to vertical or square, burn the captions in from these word timings, and file the finished video to your assets. See turn one recording into a week of content.

The page describes itself as "Turn a recording, podcast, or YouTube link into searchable text, with speakers, timestamps, and clip-ready exports."

A finished transcript: a header listing the source file, who uploaded it, the duration, word count and speaker count, the quality, the language and its sharing state; below it the text itself, each turn labelled with a speaker name and a timestamp; and on the right a speaker breakdown with each person's share of the talking, and a list of chapters

What you get

Searchable text

The full recording as clean, readable text. Find any phrase in seconds instead of scrubbing the timeline.

Speakers

At high quality, each turn is attributed to a speaker. Rename them once and it updates everywhere, exports included.

Word-level timestamps

At high quality, every word is anchored to its moment in the audio. Select a line to jump the player straight to it.

Chapters and a title

A title is written for the transcript as it's created, and chapters marking the topic shifts arrive a moment later.

Clip-ready exports

Plain text, speaker text, JSON, and subtitle files (SRT / VTT) — ready for your editor, player, or workflow.

What you can do with it

Transcripts are the raw material for almost any content workflow. A few of the most common:

Media & broadcast

Transcribe a full radio show, TV segment, or live stream, then pull timestamped moments into short clips for Instagram, TikTok, and YouTube.

Podcasts

Turn a long episode into vertical clips and quote threads — and ship show notes without re-listening.

Clipping & editing

Cut the moment here — vertical, captions burnt in — where the clip skill is on. Or export SRT and finish it in Premiere or any subtitle-aware tool.

Agencies & social

Repurpose one interview or webinar into posts, carousels, and threads across multiple client accounts.

Education

Convert recorded classes and lectures into summaries, carousels, and study notes.

Newsrooms

Search long interviews for the exact quote, attribute it to the right speaker, and clip it fast.

Pair transcripts with Chat to go from text to output in one move — ask for a thread, a set of clip ideas with timestamps, or a script in your brand's tone. Endless can search your transcripts, read one in full, and even make a new one from a link mid-conversation; mentioning a specific transcript with @ hands it the right one instead of making it go looking.

Add a recording

Drag an audio or video file onto the uploader, or select "choose a file" to pick one — you can add several at once. To use a link instead, paste it into the link box and select "Transcribe". Pasting a link anywhere on the page (as long as you aren't typing in a field) drops it into that box for you.

Pick a quality, where you get the choice

For a YouTube or TikTok link, a quality toggle appears: "Fast" uses the video's existing captions and is ready in seconds, "High quality" transcribes the audio instead, adding speaker labels and exact word timings. Fast is preselected, and the rate each option is billed at is shown beside it — transcription is priced by the length of the recording, and fast is several times cheaper.

Everything else has no choice to make. Instagram and X links only have the captions route, so they always run fast. Uploads and direct media links always run at high quality, whatever the toggle last said.

Watch the queue

Recordings are transcribed one at a time, and you can keep adding more while it works. Each item shows where it is — "Queued", then "Compressing" (large files are shrunk in your browser before they're uploaded), "Uploading", "Transcribing", "Saving" — and finishes with "Done" and an "Open" action. A failure doesn't stop the rest: that item gets a retry, and the others carry on.

Read, search, and export

Open the transcript to read it by speaker or as plain text, jump around by selecting any line, search across the whole thing, and export in the format you need.

Pasting the same link twice doesn't transcribe it twice. Endless recognizes a source it has already done — including the same video in a different URL shape — and opens the transcript you already have, at no further cost.

What you can transcribe

TypeAccepted
AudioAny audio file; the uploader names MP3, WAV, and M4A
VideoAny video file; the uploader names MP4, MOV, and WEBM
LinksYouTube, Instagram, TikTok, X, and direct links to a media file

Files must be under 2 GB each — the app says so when one isn't: "Files must be under 2 GB. Try a shorter clip, or compress it first." For something longer, split it into segments or compress it before uploading.

Reading a transcript

The transcript page opens with the essentials in a short list — Source, Uploaded, Duration, Quality, Language, Sharing — above the text itself. If the source was a file you uploaded, its name is a download link.

Search inside it

"Search transcript…" finds every match and steps through them with Enter and Shift+Enter. Selecting any line moves the player to that moment.

Speakers panel

Each speaker is listed with their share of the talking. Rename one and the new name replaces it everywhere — the text, the panel, and every export.

Chapters panel

Chapters are written for you shortly after transcription finishes — until they land it says "Chapters appear a moment after this recording finishes transcribing." Select one to jump there.

Language

When a captioned source offers more than one caption language, the Language row becomes a picker and Endless re-fetches the transcript in the language you choose.

Two views sit above the text: "Speakers", which keeps each turn attributed, and "Plain", which reads as paragraphs. There's a copy button for the whole thing, and a playbar with playback speed and 10-second skips.

Export formats

Every transcript exports in five formats from the "Export" menu:

FormatFileBest for
Plain text.txtCopy-paste, drafts, feeding into another tool
Speaker text.txtInterviews and meetings — keeps who said what
JSON.jsonWords plus timing, for custom tooling
SubRip.srtSubtitles for most editors and players
WebVTT.vttCaptions for web video

When it doesn't work

A transcription that fails leaves the item in your list rather than disappearing — it reads "Couldn't transcribe. Try again.", and opening it explains: "Couldn't transcribe this source. Delete it, or paste the link again to retry." Nothing is charged for a transcription that didn't produce a transcript; the charge happens only once the text exists. See Credits and spending.

Who can make one

Transcripts can be switched off. It's one of the features an admin can turn off for the whole workspace, a team, or one person — the description in the settings reads "Turn recordings, podcasts, and links into searchable text." When it's off, the page isn't reachable and Endless can't make a transcript in chat either.

Creating a transcript is a content action, so it needs a role that can create content — a viewer, who reads and nothing else, can't start one. See Roles and permissions.

Next

On this page