Deepgram / Transcribe an audio or video file to text
Deepgram tool deepgram-listen on looot: input fields, $0.000072 per result, output shape, and code to run it with curl, JavaScript or Python.
Speech to text: transcribes a pre-recorded audio or video file from its URL with Deepgram. Send url; model, diarize, topics, intents and redact set options. Returns the transcript with word timings, plus duration and model info.
- Tool id:
deepgram-listen - Provider: Deepgram
- Job: Transcribe an audio or video file to text (
audio.transcribe) - Price: $0.000072 per result. A call that fails at the provider costs $0.
Inputs
| Name | Type | Required | Description |
|---|---|---|---|
url |
string | yes | |
callback |
string | no | |
callback_method |
string | no | |
custom_intent |
string | no | |
custom_intent_mode |
string | no | |
custom_topic |
string | no | |
custom_topic_mode |
string | no | |
detect_entities |
boolean | no | |
detect_language |
string | no | |
diarize |
boolean | no | |
diarize_model |
string | no | |
dictation |
boolean | no | |
encoding |
string | no | |
extra |
string | no | |
filler_words |
boolean | no | |
intents |
boolean | no | |
keyterm |
array | no | |
keywords |
string | no | |
language |
string | no | |
measurements |
boolean | no | |
mip_opt_out |
boolean | no | |
model |
string | no | |
multichannel |
boolean | no | |
numerals |
boolean | no | |
paragraphs |
boolean | no | |
profanity_filter |
boolean | no | |
punctuate |
boolean | no | |
redact |
string | no | |
replace |
string | no | |
search |
string | no | |
sentiment |
boolean | no | |
smart_format |
boolean | no | |
summarize |
string | no | |
tag |
string | no | |
topics |
boolean | no | |
utt_split |
number | no | |
utterances |
boolean | no | |
version |
string | no |
Output
Shape of the run’s result, checked against 3 real answers:
{ results: { channels: { alternatives: { ... }[] }[] }, metadata: { models: string[], sha256, created, channels, duration, model_info: Record<string, { ... }>, request_id, transaction_key } }
Run it
Every call needs your API token in LOOOT_TOKEN; Sign in shows how to get one. Each run also needs a new idempotency key, so a retry never pays twice. With wait: 30 the answer comes back inline when the run ends within 30 seconds. Otherwise you get the running run back: poll GET /v1/runs/<runId>.
Example input with placeholder values:
curl -X POST "https://api.looot.ai/v1/runs" \
-H "Authorization: Bearer $LOOOT_TOKEN" \
-H "Content-Type: application/json" \
-H "Idempotency-Key: $(uuidgen)" \
-d '{"endpointId":"deepgram-listen","input":{"url":"https://example.com"},"wait":30}'const response = await fetch("https://api.looot.ai/v1/runs", {
method: "POST",
headers: {
Authorization: `Bearer ${process.env.LOOOT_TOKEN}`,
"Content-Type": "application/json",
"Idempotency-Key": crypto.randomUUID(),
},
body: JSON.stringify({
endpointId: "deepgram-listen",
input: {
url: "https://example.com",
},
wait: 30,
}),
});
const run = await response.json();
console.log(run.status, run.result);import os
import uuid
import requests
response = requests.post(
"https://api.looot.ai/v1/runs",
headers={
"Authorization": f"Bearer {os.environ['LOOOT_TOKEN']}",
"Idempotency-Key": str(uuid.uuid4()),
},
json={
"endpointId": "deepgram-listen",
"input": {
"url": "https://example.com",
},
"wait": 30,
},
timeout=90,
)
run = response.json()
print(run["status"], run.get("result"))To let looot pick among every provider of this job instead, send job:audio.transcribe as endpointId; see the job page.