Skip to content
onelayer.
All tools

Tool detail

transcribe

Transcribes the spoken audio of a YouTube, TikTok or Instagram video with Whisper large-v3, returning full text with sentence-level timestamps.

Categories: Speech audio

Use when

  • transcribe a YouTube, TikTok or Instagram video's spoken audio
  • get a transcript for a video with no caption track or non-English audio

Not for

  • not caption scraping - runs real speech recognition, so no subtitle track is required

Inputs

GET request: no parameters declared in this schema; provider's own listing example shows a url field for the video link.

Admitted input schema

{
  "type": "object",
  "properties": {},
  "required": [],
  "additionalProperties": false
}

Endpoint

Method
GET
URL
https://api.x-402.online/v1/transcribe

Payment

Protocol
x402
Listed price
0.02 USDC per call on Base
Payment method
exact
Network
eip155:8453
Asset
0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913

Listed terms are catalog metadata. The caller supplies inputs and calls the tool outside Onelayer. If payment is required, the caller uses a compatible wallet to pay.

Source

Record
GET https://api.x-402.online/v1/transcribe
Retrieved
2026-09-30T20:37:07.382Z
Catalog file SHA-256
6837ccddf886f0007cf3727ea9d27faaa5047b51b4f425345c554ce1c271a524
Catalog built
2026-09-30T22:31:58.034Z