Pull structured data off web pages, and run bulk extract jobs from the terminal
Connect Diffbot and Luumen can read page structure, article fields, list pages, and job posting data without leaving the terminal. It can also check your credit balance and the state of any bulk job. When you need to process a list of URLs, Luumen prepares the bulk extract job and shows you the parameters before anything runs.
35 tools: 25 read, 10 write. Reads answer instantly. Writes require approval by default. Everything is logged.
What a governed Diffbot run looks like inside Luumen.
Authorize once with API token. Luumen lists the scopes each action needs before you approve the connection, and credentials never appear in the chat.
Read actions answer immediately. Anything that writes — create bulk extract job, create or update custom api, create bulk enhance job, delete custom api, and more — is shown as a plan and requires approval by default, including the 4 actions classified as destructive. Administrators configure that per tool, so you decide exactly which actions can ever run unattended.
You decide. Actions are granted per agent, skill, and team, and per environment — production is not staging. Read access can be broad while writes stay narrow.
Every call to Diffbot — read or write, approved or declined — is recorded with the actor, the input, and the result, and can be linked to the ticket or change record.
Connect in minutes. Every action scoped, approved, and audited from day one.