---
title: "Turn text into spoken audio"
description: "Run job:audio.speech.generate through the looot API: 1 provider, from $0.00008 per result. Inputs, prices, output shape and code for curl, JavaScript and Python."
sidebar:
  hidden: true
---

{/* Generated by scripts/generate-api-reference.mjs from data/api-reference.json. Do not edit. */}

Turn text into spoken audio.

Job id `audio.speech.generate`, in Generative AI (Media generation). As of 2026-10-01, 1 provider serve it through 1 tool. Run `job:audio.speech.generate` and looot picks one of them; with `fallback` on, a miss moves on to the next. See [Jobs](/concepts/jobs).

## Inputs

This job has one tool, so it takes that tool's fields. Any key you send passes through unchanged.

| Name | Type | Required | Description |
| --- | --- | --- | --- |
| `language_code` | string | no | Optional ISO 639-1 language code, for example en or fr. |
| `model_id` | string | no | Voice model. eleven_multilingual_v2 and eleven_v3 cost $0.08 per 1,000 characters; eleven_flash_v2_5 and eleven_turbo_v2_5 cost $0.04. Defaults to eleven_multilingual_v2. |
| `text` | string | yes | The words to speak, up to 5,000 characters (the eleven_v3 limit, the lowest of the four models). |
| `voice_settings` | object | no | Optional: stability, similarity_boost, style, use_speaker_boost, speed. |
| `output_format` | string | no | Audio format. Defaults to mp3_22050_32; mp3_44100_64 gives better quality. |
| `voice_id` | string | no | ElevenLabs voice id. Defaults to Rachel (21m00Tcm4TlvDq8ikWAM). |

## Providers and prices

| Provider | Tool | Price |
| --- | --- | --- |
| [ElevenLabs](/providers/elevenlabs) | [`elevenlabs-text-to-speech-audio`](/reference/tools/elevenlabs-text-to-speech-audio) | $0.00008 per result |

A call that fails at the provider costs $0. See [What's free](/money/free-and-failures).

## Output

Shape of the run's `result`, the expected shape, not yet checked against real answers:

```txt
{ media: { mediaId, contentType, byteLength, sha256, expiresAt, downloadPath } }
```

## Run it

Every call needs your API token in `LOOOT_TOKEN`; [Sign in](/get-started/sign-in#use-the-token-in-scripts-and-agents) shows how to get one. Each run also needs a new idempotency key, so a retry never pays twice. With `wait: 30` the answer comes back inline when the run ends within 30 seconds. Otherwise you get the running run back: poll `GET /v1/runs/<runId>`.

Example input with placeholder values:

<CodeGroup>

```bash curl
curl -X POST "https://api.looot.ai/v1/runs" \
  -H "Authorization: Bearer $LOOOT_TOKEN" \
  -H "Content-Type: application/json" \
  -H "Idempotency-Key: $(uuidgen)" \
  -d '{"endpointId":"job:audio.speech.generate","input":{"text":"<text>"},"wait":30}'
```

```js JavaScript
const response = await fetch("https://api.looot.ai/v1/runs", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.LOOOT_TOKEN}`,
    "Content-Type": "application/json",
    "Idempotency-Key": crypto.randomUUID(),
  },
  body: JSON.stringify({
    endpointId: "job:audio.speech.generate",
    input: {
      text: "<text>",
    },
    wait: 30,
  }),
});
const run = await response.json();
console.log(run.status, run.result);
```

```python Python
import os
import uuid

import requests

response = requests.post(
    "https://api.looot.ai/v1/runs",
    headers={
        "Authorization": f"Bearer {os.environ['LOOOT_TOKEN']}",
        "Idempotency-Key": str(uuid.uuid4()),
    },
    json={
        "endpointId": "job:audio.speech.generate",
        "input": {
            "text": "<text>",
        },
        "wait": 30,
    },
    timeout=90,
)
run = response.json()
print(run["status"], run.get("result"))
```

</CodeGroup>

To pin one provider, send its tool id as `endpointId` instead of `job:audio.speech.generate`.
