How-to · 3 steps
How to generate video
with Monid.
Give this to your agent
$set up https://monid.ai/SKILL.md, then make me a five second clip of <shot> on the cheapest model that can do it
and let it take it from there.
#RunWhat came backCost
1Cheapest, 5s480P clip in 15 seconds$0.125 ✓
2Same model, 10sexactly double, to the cent$0.25 ✓
3Image to video720p from one still$0.21 ✓
4Second model, 5ssame prompt, 1.7x the price$0.21 ✓
5With audiospeech and ambience included$0.20 ✓
6The premium tiernot run, catalog rate$1.16 ✓
7A 4K secondone second of the top tier$0.42 ✓
8Totalfive clips, all billed as quoted$0.995 ✓
five clips · four providers · one eleven-fold spreadagent running$0.995
Real run, 2026-09-24. Prompts and outputs in the research folder.
Step 1
Set up Monid.
One line. It installs the CLI and asks for an API key from app.monid.ai. New accounts start with $1.00.
Say this to your agent
>set up https://monid.ai/SKILL.md
Step 2
Pick what you need. Say it in one sentence.
Click a job. Paste the sentence, fill in the brackets.
The cheapest clip that works
Say this to your agent
>generate a five second clip of <shot> with minimax h3-max-turbo at 480P on monid, passing the parameters in the request BODY: every video endpoint here takes a body, and sent as query params none of them run at all. tell me the billed seconds so i can check the price against the rate
What the agent runsdone
minimax /v1/video/minimax-h3-max-turbo$0.025 / s
minimax /v1/video/minimax-h3 · not run$0.038 / s
kling /v1/video/kling-2.5-turbo-t2v$0.042 / s
gemini /v1/video/omni-flash-t2v$0.04 / s
alibaba /v1/video/wan3.0 · not run$0.05 / s
What came backfive real clips · 2026-09-24
The same five-second clipUS dollars
what came backa playable 480P clip, five billed seconds, in fifteen seconds of wall clock
the spreadthe identical five-second prompt is $0.125 here and over $1.40 on the top audio tier
the health flagthe cheapest endpoint is marked degraded in the catalog and still answered first try
plain inspect crashes on four of these modelsit dies mid-pricing with a type error and prints no input schema at all, so an agent told to inspect before running stalls on a model that works; ask for JSON output instead and it answers
a dead job can still say COMPLETEDtwo runs came back with a run status of COMPLETED over a provider response of 500 and a body reading failed. They correctly billed nothing, but anything branching on the run status reports a success that has no file behind it
the rulepick the tier by what the shot needs, because the model choice is an eleven-fold decision
$0.125five seconds at $0.025
What a second actually costs
Say this to your agent
>run the same prompt again at ten seconds on the same model and show me both bills side by side, so i can see whether the price is per second or per call
What the agent runsdone
minimax /v1/video/minimax-h3-max-turbo$0.025 / s
minimax /v1/video/minimax-hailuo-2.3 · not run$0.28 / bundle
gemini /v1/video/omni-flash-t2v$0.04 / s
What came backfive real clips · 2026-09-24
Double the durationUS dollars
the proofsame model, same prompt, twice the duration, exactly twice the bill
why it mattersa per-second rate means a thirty-second cut is six times a five-second one, not a little more
the exceptionone model is sold as fixed bundles, so six seconds is the floor and five is not offered
the auto trapsetting duration to auto pre-holds a full thirty seconds of charge, and it does it on three of the models here rather than one, so never leave the duration unset
$0.25ten seconds at $0.025
Turn a still into a shot
Say this to your agent
>animate <image url> into a five second clip with kling 2.6 image to video at 720p on monid, audio off, passing the parameters in the request BODY rather than as query params. use an image host that serves requests with no user agent: wikimedia returns 403 to the provider's fetcher and the failure takes a minute to surface as an unhelpful message about getting the contents of the file
What the agent runsdone
kling /v1/video/kling-2.6-i2v$0.042 / s
minimax /v1/video/minimax-h3-max-turbo$0.025 / s
gemini /v1/video/omni-flash-i2v · not run$0.04 / s
alibaba /v1/video/wan2.7-i2v · not run$0.10 / s
kling /v1/video/kling-3.0-i2v · not run$0.084 / s
What came backfive real clips · 2026-09-24
what came backa clip from a single still, in 54 to 64 seconds of wall clock across two runs
the image host is the hard parttwo runs failed before one worked: the most obvious public image host returns 403 to any client sending no user agent, which is what the provider's server-side fetcher is, and the error it produces names the file rather than the refusal
and 720p is not what arrivesa request for 720p came back 1172 by 784, so the resolution parameter sets the price tier rather than the geometry: measure the file, do not trust the request
the roundingthe output ran 5.041 seconds and the bill was for five, because it bills whole requested seconds
the cheap alternativethe cheapest model also does image to video with first and last frame, at 480P for $0.025 a second
the 4K jumpthe same job on the top tier at 4K is $0.42 a second, so five seconds is $2.10
$0.21five seconds at $0.042
A clip that comes with sound
Say this to your agent
>generate <shot> with gemini omni-flash on monid and include the audio cue in the prompt. tell me whether the rate changes with duration
What the agent runsdone
gemini /v1/video/omni-flash-t2v$0.04 / s
kling /v1/video/kling-2.6-t2v · not run$0.14 / s with audio
kling /v1/video/kling-3.0-turbo-t2v · not run$0.112 / s
elevenlabs /v1/text-to-speech · not runto $0.22 / call
What came backfive real clips · 2026-09-24
what came backa 360p clip with speech, music and ambience generated together, in 34 seconds
flat is the norm, not the exceptionan earlier version of this page called this the only model with one rate across every duration. It is the other way round: nearly every endpoint here is flat per second, and the bundle-priced one is the lone exception
and it is not the cheapest way to get soundanother model does audio at 720P for $0.10 a second, and a turbo tier does it at $0.112, so the $0.14 1080p tier this page used to name as the cheapest alternative is neither the cheapest nor the only one
the link trapthe signed download link expires in one hour while the file lives seven days: fetch it, do not save the url
$0.20five seconds at $0.04
When the premium tier is worth it
Say this to your agent
>price a five second clip on the premium token-billed tier before running anything, and tell me what it would cost against the cheapest model. confirm with me before you run it
What the agent runsdone
bytedance seedance-2.5 · not run$10.70 / 1M tokens
kling /v1/video/kling-3.0-i2v · not run$0.42 / s at 4K
alibaba /v1/video/wan2.7-videoedit · not run$0.10 / s, both ways
What it returnsshape of the answer
the pricea five second clip is $1.16 at 720p, which is its default and also its ceiling, or $0.50 at 480p: about one and a half times the four real clips on this page put together
the number this page used to printan earlier version said $3.50 a clip. That figure is not a clip price at all: it is the per-million-token RATE of the CHEAPEST tier, which the catalog prints as $3.5 / 1M. Read the unit next to a token price before you turn it into a clip price
the unitit bills by tokens, but duration moves the bill far more than resolution does: across this tier resolution spans 2.3x and duration spans 7.5x
the hidden route, with a caveatone aggregator lists a flat $0.44 per call while its model field can route the job elsewhere; on our second look its documented models were the previous generation rather than this tier, so read the model list before you trust the flat price
the editing trapone edit endpoint bills the input seconds as well as the output, so a five second edit costs like ten
$1.16five seconds at 720p · catalog rate
Step 3
Take the cheapest route with Monid.
Read the pricing unit before the price. Almost everything here bills per second, so the duration you ask for is the decision.
Cheapest seconds · $0.025
H3 Max Turbo 480P$0.125five seconds
Omni Flash 360p$0.20audio included
Better shots · cheapest first
Kling 2.5 turbo$0.21
H3 Max 768P$0.4
Kling 2.6 audio$0.7
Wan 3.0 1080P$1.00
Kling 3.0 at 4K$2.1
Fetch it
Download the file$0.00
Judge the output$0.042keep or regenerate · catalog rate
Back to you
rowrendered bycost
1a still was enough$0.125
2Kling 2.5 turbo$0.335
3a still was enough$0.125
4Kling 2.6 audio$0.825
5Kling 3.0 at 4K$2.225
5 of 5 rows$3.635
Say this to your agent
>make these shots on the cheapest model that can do each one, and show me the cost and billed seconds per shot: <paste shot list>
“ask for the unit”per second and per bundle are different products; one has no five-second option at all
“name the duration”ten seconds is exactly double five on every per-second model tested
“fetch, do not link”one provider's download link dies in an hour while the file lives a week
“confirm before premium”a five second clip on the token-billed tier is $1.16, about one and a half times every cheap clip here put together
“read the unit on a token price”a per-million-token rate is not a clip price; reading one as the other put a number on this page that was three times too high
“say where the parameters go”every video endpoint here takes a body, and as query params none of them run
One key. 1,700+ tools.
Video generation is one shelf of it.
The cheapest secondsminimax · from $0.025 / s
Motion, audio, 4Kkling · $0.042 to $0.42 / s
One flat rate, audio includedgemini · $0.04 to $0.31 / s
Generate and editalibaba · $0.05 to $0.28 / s
Premium, billed by tokensbytedance · $10.70 / 1M tokens
The voice over itelevenlabs · to $0.22 / call
Read the briefcontext.dev · from $0.0009
What is working alreadytikhub · $0.0015 / call
Judge the outputtypesafe · $0.042 / call
Reference footageapify · from $0.00045 / result
+ 1,700 more across 55 providers
Browse the catalog →Read more
View all →
MiniMax H3 Max on Monid
The model behind the cheapest seconds here, and what it can and cannot frame.

The best social media scraping API in 2026
Before you generate a shot, seeing what already works is the cheaper call.

LLM gateway or MCP gateway
Why one key across every model matters more once the models bill by the second.







