Skip to content
looot docs
Esc
↑↓navigate↵open⌘Jpreview
On this page

Apify / Read a run's dataset items

Apify tool apify-dataset-items on looot: input fields, $0.0001 per call, output shape, and code to run it with curl, JavaScript or Python.

Reads the items a finished Apify actor run wrote to its dataset. Call it after apify-actor-run or apify-meta-ads-library-search and send dataset_id (the run’s defaultDatasetId); limit, offset, fields and omit shape the output.

Inputs

Name Type Required Description
clean boolean no true (default) strips empty/hidden fields from each item. Example: true
fields string no Comma-separated field allow-list. Example: “title,url”
limit integer no Items per page. Example: 10
offset integer no Items to skip. Example: 0
omit string no Comma-separated fields to leave out. Example: “rawHtml”
dataset_id string yes defaultDatasetId from an apify-actor-run (or apify-meta-ads-library-search) result – never guess or enumerate one.

Output

Shape of the run’s result, checked against 1 real answer:

{ message }[]

Run it

Every call needs your API token in LOOOT_TOKEN; Sign in shows how to get one. Each run also needs a new idempotency key, so a retry never pays twice. With wait: 30 the answer comes back inline when the run ends within 30 seconds. Otherwise you get the running run back: poll GET /v1/runs/<runId>.

Example input with placeholder values:

curl -X POST "https://api.looot.ai/v1/runs" \
  -H "Authorization: Bearer $LOOOT_TOKEN" \
  -H "Content-Type: application/json" \
  -H "Idempotency-Key: $(uuidgen)" \
  -d '{"endpointId":"apify-dataset-items","input":{"dataset_id":"<dataset_id>","omit":"rawHtml","clean":true,"limit":10,"fields":"title,url","offset":0},"wait":30}'
const response = await fetch("https://api.looot.ai/v1/runs", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.LOOOT_TOKEN}`,
    "Content-Type": "application/json",
    "Idempotency-Key": crypto.randomUUID(),
  },
  body: JSON.stringify({
    endpointId: "apify-dataset-items",
    input: {
      dataset_id: "<dataset_id>",
      omit: "rawHtml",
      clean: true,
      limit: 10,
      fields: "title,url",
      offset: 0,
    },
    wait: 30,
  }),
});
const run = await response.json();
console.log(run.status, run.result);
import os
import uuid

import requests

response = requests.post(
    "https://api.looot.ai/v1/runs",
    headers={
        "Authorization": f"Bearer {os.environ['LOOOT_TOKEN']}",
        "Idempotency-Key": str(uuid.uuid4()),
    },
    json={
        "endpointId": "apify-dataset-items",
        "input": {
            "dataset_id": "<dataset_id>",
            "omit": "rawHtml",
            "clean": True,
            "limit": 10,
            "fields": "title,url",
            "offset": 0,
        },
        "wait": 30,
    },
    timeout=90,
)
run = response.json()
print(run["status"], run.get("result"))

To let looot pick among every provider of this job instead, send job:web.scrape.job.results as endpointId; see the job page.

Was this page helpful?