Tool detail
transcribe
Transcribes the spoken audio of a YouTube, TikTok or Instagram video with Whisper large-v3, returning full text with sentence-level timestamps.
Categories: Speech audio
Use when
- transcribe a YouTube, TikTok or Instagram video's spoken audio
- get a transcript for a video with no caption track or non-English audio
Not for
- not caption scraping - runs real speech recognition, so no subtitle track is required
Inputs
GET request: no parameters declared in this schema; provider's own listing example shows a url field for the video link.
Admitted input schema
{
"type": "object",
"properties": {},
"required": [],
"additionalProperties": false
}Endpoint
- Method
GET- URL
https://api.x-402.online/v1/transcribe
Payment
- Protocol
- x402
- Listed price
- 0.02 USDC per call on Base
- Payment method
- exact
- Network
- eip155:8453
- Asset
- 0x833589fCD6eDb6E08f4c7C32D4f71b54bdA02913
Listed terms are catalog metadata. The caller supplies inputs and calls the tool outside Onelayer. If payment is required, the caller uses a compatible wallet to pay.
Source
- Directory
- Coinbase x402 Bazaar
- Record
- GET https://api.x-402.online/v1/transcribe
- Retrieved
- 2026-09-30T20:37:07.382Z
- Catalog file SHA-256
6837ccddf886f0007cf3727ea9d27faaa5047b51b4f425345c554ce1c271a524- Catalog built
- 2026-09-30T22:31:58.034Z