You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Copy file name to clipboardExpand all lines: openapi.json
+29-5Lines changed: 29 additions & 5 deletions
Original file line number
Diff line number
Diff line change
@@ -178165,7 +178165,7 @@
178165
178165
},
178166
178166
"/v1/run/tiktok.video_transcript_full": {
178167
178167
"post": {
178168
-
"description": "Transcribe the spoken audio of a TikTok video with AnyAPI's own speech-to-text: timed sentence-level segments, speaker labels, detected language, and optional per-word timings, for videos TikTok publishes no caption track for and for videos whose caption track gets the words wrong. AnyAPI downloads the video and runs the audio through MAI-Transcribe-2 rather than reading anything TikTok wrote, which is why it handles non-English speech and hears words the native captions mishear. The answer also carries the video's id, URL, creator handle, and cover image. If you only want whatever TikTok itself published and you want it for a tenth of the price, tiktok.video_transcript is that.\n\n**Price:** billed per audio minute - \\$1.50 per 1,000 requests base + \\$6.00 per 1,000 audio minutes, capped at \\$95.00 per 1,000 requests.\n\n**Routing:** one lane serves this API today, so a failed attempt has nowhere to fail over to. Payment outcome follows the selected rail's settlement policy.\n\n**Catalog:** [TikTok Video Transcript (AnyAPI speech to text) pricing and uptime](https://getanyapi.com/api/tiktok/video-transcript-full) - live USD price, lane routing, and measured 30-day uptime. [Every TikTok endpoint](https://getanyapi.com/api/tiktok).",
178168
+
"description": "Transcribe the spoken audio of a TikTok video with AnyAPI's own speech-to-text: timed sentence-level segments, speaker labels, detected language, and optional per-word timings, for videos TikTok publishes no caption track for and for videos whose caption track gets the words wrong. AnyAPI downloads the video and runs the audio through MAI-Transcribe-2 rather than reading anything TikTok wrote, which is why it handles non-English speech and hears words the native captions mishear. The answer also carries the video's id, URL, creator handle, and cover image. Turn on hostVideo to also get the MP4 on a hosted link that plays without TikTok's signed, short-lived CDN URL expiring on you. If you only want whatever TikTok itself published and you want it for a tenth of the price, tiktok.video_transcript is that.\n\n**Price:** billed per audio minute - \\$1.50 per 1,000 requests base + \\$6.00 per 1,000 audio minutes, capped at \\$95.00 per 1,000 requests.\n\n**Routing:** one lane serves this API today, so a failed attempt has nowhere to fail over to. Payment outcome follows the selected rail's settlement policy.\n\n**Catalog:** [TikTok Video Transcript (AnyAPI speech to text) pricing and uptime](https://getanyapi.com/api/tiktok/video-transcript-full) - live USD price, lane routing, and measured 30-day uptime. [Every TikTok endpoint](https://getanyapi.com/api/tiktok).",
"description": "Also store the video and return a hosted MP4 link that plays without TikTok's signed CDN URL. Charged as an extra on top of the transcript.",
178244
+
"type": "boolean"
178245
+
},
178241
178246
"preferLatencyUnderMs": {
178242
178247
"description": "Optional; omit it and routing is unchanged, with the cheapest source serving. Prefer sources whose typical response time (median over the trailing 30 days, as published on this endpoint's lane health) is under this many milliseconds; among those, the cheapest serves. This can raise your price: when the cheapest source misses the target, a faster and dearer one serves, and you are quoted and charged its price. If no source is that fast the request is still served, by whichever source offers the best speed for its price - it is never refused for being slow. Sources we have not timed are tried last. This is a preference, not a guarantee: the median describes past requests and is not a ceiling on this one, and it excludes any wait this request itself asks for. On a paginated walk it applies to the first page only: later pages stay with the source that page chose, at the price it was quoted.",
178243
178248
"minimum": 1,
@@ -178298,11 +178303,26 @@
178298
178303
},
178299
178304
{
178300
178305
"properties": {
178306
+
"bytes": {
178307
+
"description": "Size of the hosted MP4 in bytes.",
178308
+
"minimum": 0,
178309
+
"type": "integer"
178310
+
},
178301
178311
"durationSeconds": {
178302
178312
"description": "Video duration in seconds.",
178303
178313
"minimum": 0,
178304
178314
"type": "number"
178305
178315
},
178316
+
"expiresUtc": {
178317
+
"description": "When the hosted link stops working. UTC epoch timestamp in seconds (Unix time). Multiply by 1000 for a JS Date in milliseconds.",
178318
+
"type": "number"
178319
+
},
178320
+
"hostedUrl": {
178321
+
"description": "Hosted MP4 link, returned only when the request set hostVideo. It plays without TikTok's signed CDN URL and without any cookie, and it stops working at expiresUtc.",
178322
+
"format": "uri",
178323
+
"type": "string",
178324
+
"x-anyapi-domain": "media.getanyapi.com"
178325
+
},
178306
178326
"id": {
178307
178327
"description": "TikTok video id.",
178308
178328
"type": "string"
@@ -178379,7 +178399,11 @@
178379
178399
"x-anyapi-must-populate": true
178380
178400
},
178381
178401
"source": {
178382
-
"description": "How the text was produced. Always \"audio_asr\" on this endpoint: the words come from speech recognition over the audio, not from a caption track TikTok published. For TikTok's own captions, use tiktok.video_transcript. Populated whenever the provider has data for the entity.",
178402
+
"description": "How the text was produced. \"audio_asr\" means the words come from speech recognition over the video's audio, never from a caption track TikTok published - for TikTok's own captions, use tiktok.video_transcript. \"transcript_unavailable\" means recognition did not complete for this video, so the transcript is empty for that reason rather than because the video has no speech in it; the rest of the record is still what we resolved, and no audio time is charged. Populated whenever the provider has data for the entity.",
0 commit comments