Text to speech
$0.1400/1k chars · $0.0025 minimum/callEleven v3
Generate multi-speaker dialogue audio from scripted text with voice and delivery controls.
Loading tools…
Generate images, video, speech, music and sound effects through ready-to-run APIs from live model providers.
Use verified tools with pay-per-use pricing. Provider access is included. Tools that work in your own apps ask you to connect that account.
Connect your first agent to receive $1 trial credit.
Text to speech
$0.1400/1k chars · $0.0025 minimum/callGenerate multi-speaker dialogue audio from scripted text with voice and delivery controls.
Image editing
$0.0070/megapixelEdit one or more images from a text instruction with configurable output size and format.
Image generation
$0.0420/megapixelGenerate images from text prompts with configurable aspect ratio, resolution and output format.
Image generation
$0.0042/megapixelGenerate images from text prompts with configurable size, inference steps and output count.
Image generation
$0.0280/imageGenerate or edit images from text prompts with aspect-ratio and output-format controls.
Image generation
$0.0049/imageGenerate images from text prompts with configurable aspect ratio and output count.
Video generation
$0.1960/video secondGenerate video from an input image and text prompt with duration and resolution controls.
Video generation
$0.0150/1k tokensGenerate video with duration, resolution and native audio controls.
Music generation
$0.0280/minuteGenerate music from text prompts with duration and instrumental controls.
Music generation
$0.2100/minuteGenerate music from text prompts with duration and composition controls.
Sound effect generation
$0.1680/minuteGenerate sound effects from text prompts with duration and prompt-influence controls.
Text to speech
$0.0700/1k chars · $0.0025 minimum/callGenerate speech audio from text with a selected voice and output format.
Video generation
$0.1960/video secondGenerate video from one to seven reference images and a tagged text prompt with duration and aspect-ratio controls.
Video generation
$0.1960/video secondGenerate video from a text prompt with duration, resolution and aspect-ratio controls.
Video generation
$0.1120/video secondGenerate video from a text prompt with bounded duration, resolution and aspect-ratio controls.
Video generation
$0.0034/1k tokensGenerate video from text prompts with duration, resolution and aspect-ratio controls.
Video generation
$0.0150/1k tokensGenerate video with duration, resolution and native audio controls.
Video generation
$0.0150/1k tokensGenerate video with duration, resolution and native audio controls.
Brief your agent around the outcome you want, then let it plan the work and select the right Scrollport capabilities and tools.
Start with the result you want rather than a model or API. Tell the agent who the content is for, where it will appear, what it should achieve and which source material it must use.
Ask the agent to turn a broad request into a short production plan before it generates anything. Approve the concept first, create a low-cost draft next, and produce the final assets only after the direction is settled.
Give the agent specific feedback about what is not working, such as pacing, composition, tone or fidelity to the source material. Change one important variable at a time so the next result tests a clear improvement.
This example lets your agent plan a multi-stage media job and choose the appropriate Scrollport tools.
Example prompt to give your agent
Use Scrollport to help me create [MEDIA ASSET] for [AUDIENCE AND CHANNEL]. The goal is [OBJECTIVE]. Use the attached [SOURCE MATERIAL] and follow these requirements: [MUST INCLUDE]. Avoid: [MUST NOT INCLUDE]. First propose a short production plan and ask me for anything essential that is missing. Then choose the most suitable Scrollport capability and tool for each stage. Show me the price before every paid run and create one low-cost draft before the final version. Return the finished [FORMAT] plus any reusable source assets.