Browse fal.ai models, check pricing, and run inference jobs as approved steps
Connect fal.ai and Luumen can search the model catalog, read unit pricing, estimate what a batch will cost, and follow queued requests to completion. When you want work to actually run, it submits the job, uploads input files, or cancels a stuck request — each change previewed and approved first. Auth is a fal.ai API key.
12 tools: 7 read, 5 write. Reads answer instantly. Writes require approval by default. Everything is logged.
What a governed Fal.ai run looks like inside Luumen.
Authorize once with API token. Luumen lists the scopes each action needs before you approve the connection, and credentials never appear in the chat.
Read actions answer immediately. Anything that writes — cancel queue request, run model (sync), submit async inference job, subscribe to async inference job, and more — is shown as a plan and requires approval by default, including the 1 action classified as destructive. Administrators configure that per tool, so you decide exactly which actions can ever run unattended.
You decide. Actions are granted per agent, skill, and team, and per environment — production is not staging. Read access can be broad while writes stay narrow.
Every call to Fal.ai — read or write, approved or declined — is recorded with the actor, the input, and the result, and can be linked to the ticket or change record.
Connect in minutes. Every action scoped, approved, and audited from day one.