switchboard
f

fal

by fal

Runs generative image, video, and audio models through a queue-based inference API, Python and JS clients, a CLI, and a hosted MCP server, with webhooks and streaming output.

4
Skills
API Key
Auth
Yes
Streaming
Yes
Push

Skills

Asynchronous Inference

Submits model requests to the fal queue, then polls status or fetches results when long-running generations complete.

Webhook Callbacks

Notifies a developer endpoint when a queued request finishes so applications avoid polling for generation results.

Streaming Inference

Returns progressive model output as it is generated, with WebSocket and realtime endpoints for interactive use cases.

Hosted MCP Server

Lets AI assistants search models, check schemas and pricing, run inference, and upload files via mcp.fal.ai.

Content & Mediagenerative-mediaimage-generationvideo-generationinference-apiremote-mcpserverless-gpufal
Visit Agent
fal-ai
Runs generative image, video, and audio models through a queue-based inference API, Python and JS clients, a CLI, and a hosted MCP server, with webhooks and streaming output.
fields
namefal
providerfal
urlhttps://fal.ai/docs
categoriescontent-media
accessapi · cli · mcp
authapiKey
streamingtrue
pushtrue
verifiedtrue
tagsgenerative-media, image-generation, video-generation, inference-api, remote-mcp, serverless-gpu, fal
skills
queue-inferenceAsynchronous InferenceSubmits model requests to the fal queue, then polls sta…
webhooksWebhook CallbacksNotifies a developer endpoint when a queued request fin…
streaming-inferenceStreaming InferenceReturns progressive model output as it is generated, wi…
fal-mcpHosted MCP ServerLets AI assistants search models, check schemas and pri…