Shotstack Edit — programmatic video editing via JSON timeline. Submit a `timeline` (tracks, clips, transitions) and `output` spec (format, resolution, fps), receive the rendered video URL. Supports video, image, audio, text, and HTML clips. Billed per second of rendered video duration.
Models & APIs
Filter the unified API by capability, provider and release status.
Shotstack Ingest — fetch a media file from a public URL and make it available as a Shotstack source for use in render timelines. Returns a source id that can be referenced as a clip asset. Billed per request.
Inspect a public media URL through the Shotstack probe contract.
Shotstack Probe — inspect a media file (video, audio, or image) by URL and return its properties: codec, resolution, bitrate, duration, streams, and container format. Synchronous. Billed per request.
Read the status of a Shotstack render by render ID.
Shotstack Status — look up the status of a previously submitted Shotstack render by task_id. Returns queue/render progress and the output URL when complete. Billed per request.
Shotstack Template Create — save a reusable video template from `timeline` + `output` JSON. Templates can be rendered repeatedly with different merge field values via shotstack/template-render. Billed per request.
Shotstack Template Delete — delete a template by id. Billed per request.
Shotstack Template Get — retrieve a single template by id. Returns the template name, timeline, and output spec. Billed per request.
Shotstack Template List — list all templates owned by the authenticated Shotstack account. Returns template ids, names, and definitions. Billed per request.
Shotstack Template Render — render a video from a saved template by replacing merge field placeholders. Async render → poll until done → receive OSS-hosted MP4. Billed per second of rendered video duration.
Shotstack Template Update — update an existing template's name and/or definition (timeline + output). Billed per request.
Create an async Social Boost job for a product code, target URL, and quantity.
Fetch a Social Boost job by id.
Speech to Text: Speech-to-text model for transcribing audio into text. Ideal for meeting notes, accessibility, and voice-driven automation.
Run a multi-step natural-language browser automation on any website and return a structured result. Billed per agent step; failed runs are free.
Create a remote Chromium session and return CDP connection details (cdp_url, base_url) for direct Playwright/CDP control. Billed per minute of session time.
Retrieve a saved TinyFish browser context profile.
VEO 3.1 — high-quality text-to-video generation surface for SkillBoss users. Photorealistic motion, up to 8-second clips.
VEO 3.1 Fast — cost-optimized text-to-video generation surface for SkillBoss users. Faster + cheaper per second for drafts and high volume.