TikTok API

TikTok 视频转录文本(完整版) API

用 AnyAPI 自有的语音转文字转录 TikTok 视频中的语音:带时间的句子级片段、说话人标签、检测到的语言,以及可选的逐词时间,适用于 TikTok 没有发布字幕轨的视频,以及字幕轨识别错词的视频。AnyAPI 会下载视频并用 MAI-Transcribe-2 处理音频,而不是读取 TikTok 写好的任何内容,因此它能处理非英语语音,也能听准原生字幕听错的词。结果还包含视频的 ID、URL、创作者用户名和封面图。开启 hostVideo 还能拿到托管链接上的 MP4,不必担心 TikTok 带签名的短期 CDN URL 过期。如果你只需要 TikTok 自己发布的内容,并希望价格只有十分之一,请用 tiktok.video_transcript。

POST/v1/run/tiktok.video_transcript_full
可用率
90.12%
30d · 172 次调用
请求数
172
30d · 按周,近 12 周
响应时间
4.7s
中位数 · 30d

试一试

发出你的第一个请求

打开于
获取免费密钥
示例响应
免费运行只返回前 3 条结果。为密钥充值即可获得完整响应。
{
  "data": {
    "bytes": 42,
    "durationSeconds": 12.5,
    "expiresUtc": 12.5,
    "hostedUrl": "https://example.com/page",
    "id": "a1b2c3d4",
    "language": "en",
    "ownerUsername": "alex_rivera",
    "segments": [
      {
        "endSeconds": 180,
        "language": "en",
        "speaker": "example",
        "startSeconds": 180,
        "text": "A short example description of this item.",
        "words": [
          {
            "endSeconds": 180,
            "startSeconds": 180,
            "text": "A short example description of this item."
          }
        ]
      }
    ],
    "source": "audio_asr",
    "thumbnailUrl": "https://example.com/image.jpg",
    "transcript": "example",
    "url": "https://example.com/page"
  },
  "found": true,
  "reason": "not_found"
}
响应类型定义
interface TiktokVideoTranscriptFullResponse {
  data: {
    bytes?: number;
    durationSeconds?: number;
    expiresUtc?: number;
    hostedUrl?: string;
    id?: string;
    language?: string;
    ownerUsername?: string;
    segments?: {
      endSeconds: number;
      language?: string;
      speaker?: string;
      startSeconds: number;
      text: string;
      words?: {
        endSeconds?: number;
        startSeconds?: number;
        text: string;
      }[];
    }[];
    source: "audio_asr" | "transcript_unavailable";
    thumbnailUrl?: string;
    transcript: string;
    url?: string;
  } | null;
  found: boolean;
  reason?: "not_found";
}

完整的参数与响应参考:这个接口的每个字段、类型和示例。

参考

请求、响应与价格

最近验证于 2026-09-29 · 可用率和延迟按 30d 统计
POST /v1/run/tiktok.video_transcript_full
curl -X POST https://api.getanyapi.com/v1/run/tiktok.video_transcript_full \
  -H "Authorization: Bearer $ANYAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"url":"https://www.tiktok.com/@thatdudecancook/video/7649086431641521421"}'
字段类型示例值
请求体
urlstring"https://www.tiktok.com/@thatdudecancook/video/7649086431641521421"TikTok 视频 URL(例如 "https://www.tiktok.com/@user/video/1234567890")。
hostVideoboolean同时存储视频,并返回一个无需 TikTok 签名 CDN URL 即可播放的托管 MP4 链接。在转录费用之外作为附加项扣费。
wordTimestampsboolean在每个片段内返回逐词时间。词只带时间;识别器给短语而不是单词打分,所以没有逐词置信度可返回。
响应
dataobject
data.bytesintegerSize of the hosted MP4 in bytes.
data.durationSecondsnumberVideo duration in seconds.
data.expiresUtcnumberWhen the hosted link stops working. UTC epoch timestamp in seconds (Unix time). Multiply by 1000 for a JS Date in milliseconds.
data.hostedUrlstringHosted MP4 link, returned only when the request set hostVideo. It plays without TikTok's signed CDN URL and without any cookie, and it stops working at expiresUtc.
data.idstringTikTok video id.
data.languagestringDetected spoken language of the audio (BCP-47 style code, e.g. "en").
data.ownerUsernamestringCreator handle, without the leading @.
data.segmentsobject[]Timed transcript segments in playback order, one per sentence, so a segment locates a specific line in the video rather than a whole speaker turn. Populated whenever the provider has data for the entity.
data.segments[].endSecondsnumberSegment end offset in seconds, taken from the last word it contains.
data.segments[].languagestringDetected language for this segment.
data.segments[].speakerstringSpeaker label for this segment, stable within one response and meaningless across responses. Telling voices apart is a guess, not an identification, and the label is not a name.
data.segments[].startSecondsnumberSegment start offset in seconds, taken from the first word it contains.
data.segments[].textstringText of this segment.
data.segments[].wordsobject[]Per-word timings for this segment, returned only when the request set wordTimestamps. Words carry no confidence score: the recognizer scores a phrase rather than a word.
data.segments[].words[].endSecondsnumberWord end offset in seconds.
data.segments[].words[].startSecondsnumberWord start offset in seconds.
data.segments[].words[].textstringThe recognized word, in display form with its own punctuation.
data.source"audio_asr" | "transcript_unavailable"How the text was produced. "audio_asr" means the words come from speech recognition over the video's audio, never from a caption track TikTok published - for TikTok's own captions, use tiktok.video_transcript. "transcript_unavailable" means recognition did not complete for this video, so the transcript is empty for that reason rather than because the video has no speech in it; the rest of the record is still what we resolved, and no audio time is charged. Populated whenever the provider has data for the entity.
data.thumbnailUrlstringCover image for the video. A signed, short-lived TikTok CDN URL, often served as HEIC rather than JPEG, so fetch it promptly and transcode if you need broad browser support.
data.transcriptstringFull spoken-word transcript, recognized from the video's audio track. Populated whenever the provider has data for the entity.
data.urlstringCanonical URL of the video this transcript came from.
foundboolean
reason"not_found"Present only when `found` is false, and says why there is no result. `not_found`: the source states the target does not exist, or returned nothing for it. A `found: false` answer is a successful call, not an error, and `costUsd` is what it actually cost.
价格
每次请求价格(上限)USDUS$0.095
价格(/千分钟音频)USDUS$5.94

常见问题

关于 TikTok 视频转录文本(完整版) API

AnyAPI 的 TikTok 视频转录文本(完整版) API 只需一次 POST 请求 /v1/run/tiktok.video_transcript_full,就以统一格式的 JSON 返回 TikTok 数据。用 AnyAPI 自有的语音转文字转录 TikTok 视频中的语音:带时间的句子级片段、说话人标签、检测到的语言,以及可选的逐词时间,适用于 TikTok 没有发布字幕轨的视频,以及字幕轨识别错词的视频。AnyAPI 会下载视频并用 MAI-Transcribe-2 处理音频,而不是读取 TikTok 写好的任何内容,因此它能处理非英语语音,也能听准原生字幕听错的词。结果还包含视频的 ID、URL、创作者用户名和封面图。开启 hostVideo 还能拿到托管链接上的 MP4,不必担心 TikTok 带签名的短期 CDN URL 过期。如果你只需要 TikTok 自己发布的内容,并希望价格只有十分之一,请用 tiktok.video_transcript。无论哪个数据源提供服务,AnyAPI 都返回同一套统一的 schema。费用:每千分钟音频 US$5.94,单次请求最多 US$0.095,以美元计价,无需订阅,也没有每月最低消费。过去 30 天,通过 AnyAPI 发起的 TikTok 视频转录文本(完整版) API 调用中有 90.1% 成功,响应时间中位数为 4.7 秒,共测量 172 次调用。

费用:每千分钟音频 US$5.94,单次请求最多 US$0.095,以美元计价,无需订阅,也没有每月最低消费。你只需为一个美元钱包充值,每次调用从中扣费。部分由钱包付费的失败请求会产生处理费用,我们按成本原价转嫁,不加价;错误响应会显示收取的金额。

用 AnyAPI 自有的语音转文字转录 TikTok 视频中的语音:带时间的句子级片段、说话人标签、检测到的语言,以及可选的逐词时间,适用于 TikTok 没有发布字幕轨的视频,以及字幕轨识别错词的视频。AnyAPI 会下载视频并用 MAI-Transcribe-2 处理音频,而不是读取 TikTok 写好的任何内容,因此它能处理非英语语音,也能听准原生字幕听错的词。结果还包含视频的 ID、URL、创作者用户名和封面图。开启 hostVideo 还能拿到托管链接上的 MP4,不必担心 TikTok 带签名的短期 CDN URL 过期。如果你只需要 TikTok 自己发布的内容,并希望价格只有十分之一,请用 tiktok.video_transcript。响应是统一格式的 JSON,外层结构与每个 AnyAPI 接口相同,所以解析另一个接口只需换一个 URL,别的都不用改。

过去 30 天,通过 AnyAPI 发起的 TikTok 视频转录文本(完整版) API 调用中有 90.1% 成功,响应时间中位数为 4.7 秒,共测量 172 次调用。这些是 AnyAPI 对经过网关的流量的自有测量,持续重新计算,并非公开承诺的服务等级目标。