Generate an attributed transcript
Generate a speaker-labeled transcript from a processed media recording, with per-segment speaker attribution, timestamps, and optional word-level timing.
The Attributed Transcript API turns a processed media recording into a speaker-labeled transcript. It returns timestamped segments, each attributed to the speaker who said it, with optional word-level timing and confidence scores.
Notes Agent produces speaker-attributed transcripts automatically for every meeting it records. Use this API when you want the same transcript for a media asset you have already uploaded, such as a webinar, interview, or a recording captured outside Notes Agent.
Just want to try it? You can generate an attributed transcript in the dashboard with no setup. See Generate an attributed transcript from the dashboard.
Note: You can generate AI features only for existing media. Upload and process your video first, then generate an attributed transcript on the ready media.
Before you begin
Make sure you have the following:
- A FastPix account with an active workspace. See Activate your account.
- A media asset that has finished processing, so you have its
mediaId.
Generate an attributed transcript with the API
You can trigger an attributed transcript and fetch the speaker-labeled result programmatically with the In-Video AI API. For the endpoints, request and response schemas, and the video.media.ai.attributed_transcript.ready webhook, see the API reference: Trigger attributed transcript and Get attributed transcript.
Generate an attributed transcript from the dashboard
If you prefer a no-code option, you can generate an attributed transcript for any media from the FastPix dashboard. No access token, webhook, or API call is required.
- In the FastPix dashboard, go to Video > Media.
- Click the media you want to transcribe to open its detail page.
- In the left sidebar, under In-Video AI, click Transcript.
- Click Generate Transcript.
- Wait while FastPix processes the media. The transcript shows a Processing state until it is ready.
- Review the transcript. Every line is labeled by speaker and time-coded, and you can search for a phrase and seek the player to any moment.
Limits and considerations
- All timestamps are in seconds from the start of the recording, so you can use them for playback navigation and captions.
- The
wordsarray is optional and can benullfor a segment. - Speaker labels depend on how the transcript is generated. When you call the API directly, speakers are labeled generically, such as
speaker_1andspeaker_2. When the media was recorded by Notes Agent, calling theGETattributed transcript endpoint returns the real participant names of the people who spoke in the meeting.
What’s next
- Notes Agent generates attributed transcripts automatically for recorded meetings.
- API reference: Trigger attributed transcript and Get attributed transcript. Both endpoints document their error responses.
- In-video AI events for the webhook payload.