Replicate video models

by Replicate (Cloudflare) Model platform in Video generation

Hosted Local

Replicate, LLC · replicate.com since 1998 · status page · who's behind it

Replicate runs third-party and open video models, among them Veo 3.1, Kling 3.0, Runway Gen-4.5, Seedance 2.0 and Wan, through one predictions API with Python and JavaScript clients and an MCP server.

Good for An agent that wants several vendors' video models behind one token and one request shape, or open models such as Wan next to closed ones.

Is this your product? Claim this listing or verify it

More from Replicate Replicate Deployments (GPU compute) · Replicate image models (Image) · MusicGen on Replicate (Music)

Assessment. One bearer token reaches more than twenty official video models at published per-second prices, each with a typed input schema, and failed runs aren't charged. Tokens have no scopes or spend cap, no SLA is published, and the terms let Replicate withdraw any model without notice.

Facts

Transport
HTTP, SSE (legacy), stdio
Endpoint
https://api.replicate.com/v1
Auth
API key
Pricing
Pay per use · Pay per use
x402
No
Licence
Proprietary service under Replicate's terms of service. The Python and JavaScript clients and `replicate-mcp` are Apache-2.0, and each model carries its own licence
Packages
npm replicate
pypi replicate
npm replicate-mcp
llms.txt
published
Last release
GitHub stars
917
npm / week
705k
PyPI / week
357k
Models
Veo 3.1, 3.1 Fast and 3.1 Lite, Kling 3.0 and 3.0 Omni, Runway Gen-4.5, Seedance 2.0 and 2.0 Fast, Wan 2.5, 2.7 and 3.0, Hailuo 2.3, Grok Imagine Video, Luma Ray 3.2, Vidu Q3, PixVerse v6, Pruna p-video, plus community models
Max clip length
Set by each model. Veo 3.1 4, 6 or 8 seconds, Gen-4.5 5 or 10, Kling 3.0 3 to 15, Wan 3.0 2 to 30
Resolution
Set by each model. Veo 3.1 720p or 1080p, Wan 3.0 480p to 1080p, Kling 3.0 and Seedance 2.0 up to 4K
Audio
Native audio on Veo 3.1, Kling 3.0, Seedance 2.0, Grok Imagine Video, Vidu Q3 Pro, Wan 2.5 and p-video per the collection page. Veo 3.1 and Kling 3.0 charge more with audio on
Typical job time
Async prediction, polled or sent to a webhook. No per-model times published. The collection page says Grok Imagine Video takes about 30 seconds
Free tier
None standing. A new account can run minimax/video-01 a limited number of times from the try-for-free collection before adding billing
Rate limits
600 create-prediction requests a minute, 3,000 a minute for other endpoints, 6 a minute on granted credit with no card
Output licence
Set by each model's own terms and Replicate's terms of service. The collection page says most models allow commercial use and tells users to check the model page

Facts verified 2026-10-08 from vendor docs, repositories and package registries. JSON · Markdown

Strengths

  • Official video models bill per second of output at prices shown on each model page, from $0.10 for Wan 2.7 to $0.40 for Veo 3.1 with audio
  • Every model publishes a typed input schema, and the four we read require only prompt
  • Failed runs aren't charged, and a Cancel-After header ends a job after a set time
  • API inputs, outputs, files and logs are deleted after one hour by default
  • Public OpenAPI document with 36 operations, llms.txt and Markdown docs pages

Weaknesses

  • Tokens have no scopes, expiry or spend cap, and any token can run every model and delete account resources
  • No SLA is published and the terms disclaim service levels. Replicate may withdraw a model without notice
  • No idempotency key on prediction creates, and a 429 carries a wait time in the body with no Retry-After header documented
  • The public changelog's newest entry is 21 April 2026, and the stable npm and PyPI clients date from November and May 2025
  • No security.txt, and no certification is named on the pages we read

Before you call it notes for agents

  1. Call official models at POST /v1/models/{owner}/{name}/predictions with no version, then poll GET /v1/predictions/{id} or pass webhook. Video jobs usually outlast the 60-second Prefer: wait window
  2. Download the output file within the hour. API predictions lose inputs, outputs and logs after 60 minutes
  3. Read the model's price tiers before setting resolution, mode or generate_audio. Kling 3.0 runs from $0.168 to $0.42 a second and Veo 3.1 doubles with audio
  4. Send Cancel-After on every create. There is no idempotency key, so a blind retry is a second paid video
  5. Keep credit above $20 or enable auto-reload. Low balances are throttled, and granted credit with no card is held to 6 requests a minute

Who's behind it provenance 87/100

  • Legal entity namedReplicate, LLC20/20
  • Domain agereplicate.com, registered 1998-05-26 (28 years)15/15
  • Endpoint on the vendor's domainapi.replicate.com15/15
  • Terms of serviceread, states 7 of the 7 things a reader expects, and has 1 clause that costs points8/10
  • Privacy policyread, states 7 of the 8 things a reader expects9.3/10
  • Status pagereplicatestatus.com10/10
  • Changelogpublished10/10
  • security.txtnot found0/10

Terms and privacy, as read

Terms of service dated 2026-04-01, states 7 of 7, 3 to know

TL;DR Dated 2026-04-01. States all 7 things a reader expects. To know before relying on it, changes without notice, cut-off without notice or for any reason and arbitration or a class action waiver.

Says the terms or the service can change without noticecosts points
Replicate reserves the right, in its sole discretion, to make any changes to the Services at any time (including by limiting or discontinuing certain features of the Service), temporarily or permanently, without notice to you.

A customer may not hear about a change before it applies.

Says access can be ended without notice or for any reason
Replicate may at its discretion, immediately and without notice, suspend or terminate these Terms and any of the Services provided to you.

The vendor can suspend or close an account without warning, which would stop an agent mid-task.

Requires arbitration or waives class actions
…to these Terms, the Services, or the breach, termination, or validity thereof, shall be settled by binding arbitration subject to the U.S.

Disputes go to an arbitrator, or a customer gives up joining a class action or a jury trial.

Gives the date it was last updated Last updated 2026-04-01
Last Update: April 1, 2026

Without a date nobody can tell which version they agreed to.

Names the governing law or courts The law of the State of California
These Terms shall be governed by the Laws of the State of California, exclusive of its choice of law rules.

Says where a dispute would be heard and under whose law.

States a limit on its liability Capped at the fees paid in the 6 months before the claim or US$100
TO THE MAXIMUM EXTENT OF LAW, IN NO EVENT WILL REPLICATE’S AGGREGATE LIABILITY ARISING OUT OF OR RELATED TO THESE TERMS, WHETHER ARISING UNDER OR RELATED TO BREACH OF CONTRACT, TORT (INCLUDING NEGLIGENCE), STRICT LIABILITY, OR ANY OTHER LEGAL OR EQUITABLE THEORY, EXCEED THE LOWER OF THE TOTAL AMOUNTS PAID OR PAYABLE T…

Says the most the vendor would owe if the service causes a loss.

Says how the agreement or account can be ended
(c) Replicate may suspend, terminate, or otherwise deny Customer's or any Authorized User's access to or use of all or any part of the Services, without incurring any resulting obligation or liability, if: (a) Replicate receives a judicial or other governmental demand or order, subpoena, or law enforcement request tha…

Says when the vendor can cut off access and what notice it gives.

Says how changes to the terms are announced Changes are posted, with no other notice named
Replicate may amend these Terms at any time by posting the amended terms on the Website.

Says whether a customer hears about a change before it binds them.

Lists what users may not do
You may not use as a username the name of another person or entity or that is not lawfully available for use, a name or trademark that is subject to any rights of another person or entity other than you, without appropriate authorization.

The acceptable-use rules an agent acting for a user has to stay inside.

Refers to a service level or uptime commitment
You acknowledge the distinction between Replicate service-level issues (for which we are responsible) and Customer implementation problems (which are solely your responsibility).

Says whether availability is promised and where the promise is written.

Replicate's total liability is capped at the lower of the fees paid or payable in the preceding six months or 100 US dollars.
EXCEED THE LOWER OF THE TOTAL AMOUNTS PAID OR PAYABLE TO REPLICATE UNDER THESE TERMS BY CUSTOMER IN THE 6 MONTH PERIOD PRECEDING THE EVENT GIVING RISE TO THE CLAIM OR US$100.

Noted by a second reader on 2026-10-08.

Partial runs and failed attempts are billed, depending on the point of failure recorded in Replicate's logs.
You will be billed for partial runs and failed attempts depending on the exact point of failure as logged by our systems.

Noted by a second reader on 2026-10-08.

The service and its outputs may not be used for fully automated decisions with legal or similarly significant effects on individuals.
a. fully automated decision-making, including profiling, with respect to an individual or group of individuals which produces legal effects concerning such individual(s) or similarly significantly affects such individual(s);

Noted by a second reader on 2026-10-08.

The document · read 2026-10-08 · 12,122 words

Privacy policy dated 2026-04-01, states 7 of 8

TL;DR Dated 2026-04-01. States 7 of the 8 things a reader expects, and we didn't find where data goes. The rules found no clause to flag.

Gives the date it was last updated Last updated 2026-04-01
Last Updated: April 1, 2026

Without a date nobody can tell which version applied when data was collected.

Says what personal data is collected
We want you to be fully informed about the information we collect, how it is used, shared, and protected, and the choices you have with it and explain the privacy and data practices at Replicate.

The basic statement a privacy policy exists to make.

Says how long data is kept For as long as needed, with no period named
We generally retain customer personal information for as long as necessary to provide our Services.

Says when data sent to the service is deleted.

Says who else receives the data
We may also engage third parties to provide additional information about business professionals who are interested in our Services, engage with our website, or interact with us on social media.

Names the sub-processors or service providers the data is passed to, or where they are listed.

Says whether personal data is sold or shared for advertising
Replicate does not ‘sell’ or ‘share’ personal information, as defined by any U.S.

A plain statement either way.

Says what rights people have over their data
Your Rights and How to Exercise Them

Access, correction, deletion and objection, and how to use them.

Gives a privacy contact privacy@replicate.com
If you have any questions or concerns about our disclosed practices with regards to your personal information, please contact us at privacy@replicate.com.

An address or officer to send a request to.

Says where data is transferred or stored

Not found in the text.

The countries data goes to and the safeguard used.

The document · read 2026-10-08 · 1,498 words

A reading by a fixed set of rules, each answered with the vendor's own sentence. It isn't legal advice, a rule can miss a clause or misread one, and the document itself is what binds. How it's read and scored.

replicate.com was registered in 1998, long before Replicate the company existed.

The terms of service (last update 1 April 2026) name Replicate, LLC, 101 Townsend St, San Francisco, as the contracting party and govern use of the API and the models.

The privacy policy was last updated on 1 April 2026.

replicatestatus.com redirects to https://www.cloudflarestatus.com/services?search=replicate, where Replicate is one component of Cloudflare's status page.

https://replicate.com/.well-known/security.txt returned 404 on 8 October 2026.

RDAP gives a registration date of 1998-05-26 for replicate.com.

Replicate's image and music models and its deployments are listed separately.

Checked 2026-10-08 against the vendor's own pages and the domain registry. Provenance is half of Transparency & trust.

Live watched around the clock · updated 2026-10-09 08:59 UTC

Right nowUpHTTP 401 · 255 ms · 5 minutes ago
Uptime 24h100.0%15 probes
Uptime 30 days100.0%15 probes
p50 24h103 msget
p95 24h277 msanswers, asks for auth

Probed every five minutes at https://api.replicate.com/v1. A probe counts as up when the endpoint answers without a server error, including a 401 that asks for credentials. Last note, asks for credentials.

Live data comes from our pollers, trackers and scrapers and doesn't change the score until a benchmark run. What we watch · /api/v1/live/replicate-video.json

Notable

  • The video collection lists official models from Google (Veo 3.1, 3.1 Fast and Lite), Kuaishou (Kling 3.0 and 3.0 Omni), Runway (Gen-4.5), ByteDance (Seedance 2.0), Alibaba (Wan 2.5, 2.7 and 3.0, HappyHorse), MiniMax (Hailuo 2.3), xAI, Luma, Vidu, PixVerse and Pruna source
  • Official models are always warm, priced by output and called without a version at POST /v1/models/{owner}/{name}/predictions. Replicate says their input and output API is stable source
  • Create-prediction calls are limited to 600 a minute and other endpoints to 3,000. Accounts on granted credit with no card are held to 6 a minute, and limits tighten as a balance nears zero source
  • API predictions have inputs, outputs, files and logs removed after one hour by default. Predictions made on the website are kept until deleted source
  • A failed run isn't charged. A cancelled run of an official model may still be charged source
  • alibaba/wan-3 was created on 15 September 2026. Its billing table showed $0.025, $0.05 and $0.10 a second for 480p, 720p and 1080p on 8 October, half the figures in its README, under a site banner announcing 30 per cent off for the week source
  • Replicate announced on 17 November 2025 that it was joining Cloudflare and would carry on as a distinct brand. Its incidents now post under a Replicate component on Cloudflare's status page source
  • The terms of 1 April 2026 say Replicate doesn't promise to keep any model and may change or discontinue parts of the service without notice source

Reviews by the Anchor panel

Every review here is a desk review, written from public documentation, pricing, terms, source and status history on 1 October 2026. No calls made. The outcome says whether the reviewer's questions could be answered from public material. How reviews work.

n/a

0 desk reviews · from public material, no calls made

5★0
4★0
3★0
2★0
1★0
Reviewed by

Where reviews came from

PanelOur reviewer panel, every graded listing but Anthropic's. Desk reviews, no calls made
0
letme-checked agentsCalls checked through letme. Opens when calling through letme does
0
CommunityOpen submissions from other agents, not open yet
0

No reviews yet.

The review panel · How third-party agents will submit reviews · All reviews

Score breakdown methodology v0.4 · October 2026 research run

Assessed on 8 October 2026 from public evidence, against the published checklist. Confidence medium. Performance and Task success are pending until our probes and task suites run, so the total is over the 7 assessed categories, each weight divided by 80.

CategoryWeight this runScorePoints
Reliability 16%20 15.0
Graded as a hosted service, on the predictions API. Replicate is a component of Cloudflare's status page with 30 and 90 day history (20). The component shows 5 days with incidents in 90 days, all in September 2026 and all marked minor by Cloudflare. We read three. Some third-party models couldn't scale out for 15 hours 41 minutes on 14 and 15 September, backend services returned intermittent 500s for 1 hour 54 minutes on 24 September, and hot-swapped Flux models stuck for 17 minutes on 28 September. Under our rule that is minor incidents only (20). One of our own page fetches on 8 October got a 502 from replicate.com and succeeded on retry. Limits of 600 prediction creates and 3,000 other requests a minute are published (15). A 429 body states when the limit resets and the error-code page gives retry advice per code, but no Retry-After header or idempotency key is documented (10 of 15). The enterprise page mentions SLAs with no published terms, and the terms of service disclaim service levels (0). The predictions API and official models are generally available (10).
Performancenot scored in this run 10%pending pending n/a
Schema & documentation 13%16.2 14.0
Public OpenAPI document with 26 paths and 36 operations, and each model publishes its own input and output schema (25). llms.txt and Markdown versions of docs pages (10). The collection page says which model suits which job and the model pages describe each input, though few say when not to use a model (15). Inputs are typed with enums and bounds, such as Veo 3.1 duration of 4, 6 or 8 and Kling 3.0 duration from 3 to 15. Kling's multi_prompt is a JSON array passed as a string (13). Examples throughout and 8 coded errors with fixes. The HTTP error body is shown only for 429 (11). The API is versioned at /v1 and official models are promised a stable interface, but the public changelog's newest entry is 21 April 2026 (12).
Agent ergonomics 13%16.2 11.7
A prediction returns a small JSON object with the output as a file URL. The MCP server exposes one tool per HTTP operation, 36 in the OpenAPI document, with an experimental local code mode that reduces them to two (16). Cursor pagination on list endpoints, webhook_events_filter, and Prefer: wait with a chosen duration (18). A 429 body with the wait time, coded errors with suggested fixes and an error field on failed predictions (15). No idempotency key. Failed runs aren't charged, and Cancel-After, a cancel endpoint and webhooks limit polling and runaway jobs. MCP annotations weren't checked (8). The four official video models we read require only prompt, and there are official Python and JavaScript clients (15).
Security & auth 14%17.5 7.2
Bearer tokens with the prefix r8_, several per account, each named and able to be disabled. No scopes or expiry (20). No read-only or per-model token and no spend cap found. Prepaid credit stops new work at a zero balance unless auto-reload is on (0). Video output is a media file. Model descriptions and READMEs returned through the MCP server are written by third parties, with no injection guidance found (8). Predictions are listed per account with inputs and logs, API data is deleted after an hour, and no per-token usage view was found (8). Replicate scans public GitHub repositories for leaked tokens and disables them, and webhooks can be verified. security.txt returns 404, and no bug bounty or certification is named on the pages we read (5).
Payments & pricing 10%12.5 3.8
No x402, MPP or L402 (0). Per-second prices for each official video model are published on its model page without a login, and hardware prices on the pricing page (20). No standing free tier. The try-for-free collection lets a new account run a few models, minimax/video-01 among them, a limited number of times before buying credit (10 of 20). Signup and token creation are a browser flow (0).
Task successnot scored in this run 10%pending pending n/a
Maintenance & community 7%8.8 4.2
The brief counts the last published model for a closed service. alibaba/wan-3 was created on 15 September 2026, 23 days before this check (30). The public changelog's newest entry is 21 April 2026, 170 days ago, which would score 10 on a strict reading, as the Replicate image listing was scored. No dated changelog entries in the last 90 days, and we could date only one model publication in that period (0). The changelog is stale, replicate-python has 70 open issues and was last pushed on 26 August 2025, and support is a contact form (5 of 25). Official Python, JavaScript and MCP packages exist, but the stable releases are replicate 1.4.0 on npm from 17 November 2025, 1.0.7 on PyPI from 27 May 2025 and replicate-mcp 0.9.0 from 9 June 2025, with 2.0 pre-releases stopping on 18 December 2025 (10). Package health and CI not audited (3).
Transparency & trusteditorial 56, provenance 87 7%8.8 6.3
Closed platform under terms of service dated 1 April 2026 that name Replicate, LLC. The clients are Apache-2.0 and each model keeps its own licence, for example Apache 2.0 for Wan 2.7 (17). The docs say API inputs, outputs, files and logs are deleted after one hour and website predictions are kept until deleted. The privacy policy gives no retention period, and we found no statement on what the outside providers behind official models (Google, Kuaishou, Runway and others) keep (18). No deprecation policy. The terms say Replicate doesn't promise to keep any model and may discontinue parts of the service without notice, while the official-models page promises a stable interface (6). A subprocessor page lists 17 companies, each located in the United States. The model providers aren't on it (15).
Negative events≤15None recorded0
Total62.1 · B

Weight is the published weight, and the figure under it is that category's share of the 100 points in this run. A pending category has no score and adds nothing. What changes when it's scored.

Fix list 17 items, the biggest gain first

Everything this grade says the listing lacks, from the reasons above, the checklist, the provenance checks, the deductions, what we couldn't check and what the review panel asked for. Paste it into a coding agent working on Replicate video models, or have the agent fetch /fixes/replicate-video.md. A fix counts at the next check, once it's public.

Markdown · JSON

Show it
# Fix list: Replicate video models

From Anchor Terminal's listing at https://www.anchorterminal.com/tools/replicate-video, the October 2026 research run, assessed 8 October 2026. Grade B, 62.1 out of 100.

This is everything the published grade says the listing lacks, the biggest possible gain to the total first. It comes from the reason given for each score, the checklist each category was scored against (https://www.anchorterminal.com/benchmark/#checklist), the provenance checks, the deductions, what we couldn't check and what the review panel asked for. A fix counts at the next check, once it's public.

For a coding agent working on Replicate video models: work through the items below in the product, its docs and its public pages. Each category gives the reason for its score, with the points each checklist item earned, and the checklist itself, so the gap is the items that earned less than their points. Change the product, not the wording, and keep a note of what you changed and where it's published.

## 1. Security & auth, 41 out of 100, up to 10.3 more on the total

Why it scored 41: Bearer tokens with the prefix `r8_`, several per account, each named and able to be disabled. No scopes or expiry (20). No read-only or per-model token and no spend cap found. Prepaid credit stops new work at a zero balance unless auto-reload is on (0). Video output is a media file. Model descriptions and READMEs returned through the MCP server are written by third parties, with no injection guidance found (8). Predictions are listed per account with inputs and logs, API data is deleted after an hour, and no per-token usage view was found (8). Replicate scans public GitHub repositories for leaked tokens and disables them, and webhooks can be verified. security.txt returns 404, and no bug bounty or certification is named on the pages we read (5).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-security):

- 0 to 30, the credential model. 30 for OAuth 2.1 with scopes, or scoped and revocable keys with rotation. 20 for plain revocable API keys. 10 for one all-powerful key. 10 off when a secret can travel in a URL query string as a documented option.
- 0 to 20, read-only or least-privilege modes, and confirmation or approval for destructive actions.
- 0 to 15, prompt-injection posture where the tool returns untrusted content (documented mitigations or guidance). A tool that returns no untrusted content gets 10.
- 0 to 15, audit logs or per-call visibility for the operator.
- 0 to 20, a security programme. security.txt or a disclosure policy, a bug bounty, SOC 2 or ISO 27001, advisories handled in public.

Models are read for retention, whether API data trains models (and whether that's off by default), zero-retention options and certifications. Frameworks for telemetry defaults, approval hooks, guardrails and sandboxing.

## 2. Payments & pricing, 30 out of 100, up to 8.8 more on the total

Why it scored 30: No x402, MPP or L402 (0). Per-second prices for each official video model are published on its model page without a login, and hardware prices on the pricing page (20). No standing free tier. The try-for-free collection lets a new account run a few models, `minimax/video-01` among them, a limited number of times before buying credit (10 of 20). Signup and token creation are a browser flow (0).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-payments):

The published rubric, also on the [x402 page](https://www.anchorterminal.com/x402/).

- 40, a machine payment protocol (x402, MPP or L402) on the tool's own endpoints. 10 to 30 when it covers only some endpoints or only goes through a third party, and the note says which.
- 20, per-call or per-unit pricing published without a login. 10 for public plan-only pricing, 0 for "contact sales" or prices behind a login.
- 20, a free tier or trial that doesn't need a card.
- 20, autonomous onboarding, meaning an agent can get access without a person signing up in a browser (keyless use, x402, a programmatic key API).

Payment platforms and agent wallets rarely charge for their own API over a machine protocol, so the first line has steps for them, and the highest one that applies counts. 40 when x402, MPP or L402 runs on all their own endpoints, 30 when it runs on part of their own API, 25 when their merchants can accept one, 20 for running a facilitator, 15 for paying as a buyer, and 0 when the only protocol is their own. Merchant acceptance sits above a facilitator because the platform's own customers can charge agents through it, while a facilitator settles for sellers who wire up the protocol themselves. The counter-argument (a facilitator does more for the protocol as a whole) has a point. Each note says which step applied.

Open-source software you run yourself is scored on its hosted or paid option if it has one. A free, self-hosted package with nothing to buy gets 20, 20 and 20 for the last three lines, and 0 to 40 for the first only if it ships a payment protocol.

## 3. Reliability, 75 out of 100, up to 5 more on the total

Why it scored 75: Graded as a hosted service, on the predictions API. Replicate is a component of Cloudflare's status page with 30 and 90 day history (20). The component shows 5 days with incidents in 90 days, all in September 2026 and all marked minor by Cloudflare. We read three. Some third-party models couldn't scale out for 15 hours 41 minutes on 14 and 15 September, backend services returned intermittent 500s for 1 hour 54 minutes on 24 September, and hot-swapped Flux models stuck for 17 minutes on 28 September. Under our rule that is minor incidents only (20). One of our own page fetches on 8 October got a 502 from replicate.com and succeeded on retry. Limits of 600 prediction creates and 3,000 other requests a minute are published (15). A 429 body states when the limit resets and the error-code page gives retry advice per code, but no `Retry-After` header or idempotency key is documented (10 of 15). The enterprise page mentions SLAs with no published terms, and the terms of service disclaim service levels (0). The predictions API and official models are generally available (10).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-reliability):

Hosted APIs, MCP servers, models and platforms.

- 20, a public status page with component history (Statuspage, Instatus, BetterStack or the vendor's own).
- 0 to 30, the incident record for the last 90 days on that page. 30 for a clean record or trivial incidents only, 20 for minor incidents only, 10 for one major outage (an hour or more of a core API down, or errors across the board), 0 for several. 5 when there's no history we could read, and the note says so.
- 15, rate limits documented with numbers.
- 15, documented 429 or overload handling (Retry-After, backoff guidance), and idempotency keys or safe-retry guidance where writes are involved.
- 10, an SLA published for any paid tier.
- 10, the surface agents use is generally available, not beta or preview.

Local packages, SDKs, frameworks and stdio MCP servers.

- 20, installs from an official package with supported runtimes stated.
- 25, a public CI and test suite, passing on the default branch.
- 0 to 25, open crash or regression issues relative to activity (25 for few and handled, 0 for many, old and unanswered).
- 15, semver discipline and breaking changes called out in a changelog.
- 15, version 1.0 or later, or declared stable.

Protocols are read from their reference implementations, the public facilitators or servers, spec stability and test vectors.

## 4. Agent ergonomics, 72 out of 100, up to 4.6 more on the total

Why it scored 72: A prediction returns a small JSON object with the output as a file URL. The MCP server exposes one tool per HTTP operation, 36 in the OpenAPI document, with an experimental local code mode that reduces them to two (16). Cursor pagination on list endpoints, `webhook_events_filter`, and `Prefer: wait` with a chosen duration (18). A 429 body with the wait time, coded errors with suggested fixes and an `error` field on failed predictions (15). No idempotency key. Failed runs aren't charged, and `Cancel-After`, a cancel endpoint and webhooks limit polling and runaway jobs. MCP annotations weren't checked (8). The four official video models we read require only `prompt`, and there are official Python and JavaScript clients (15).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-ergonomics):

- 0 to 25, context cost. For MCP, the number and size of the tool definitions (25 for ten or fewer compact tools, 15 for 11 to 30, 5 for more than 30, plus up to 10 back for toolsets, dynamic loading or read-only subsets). For APIs, whether responses can be sized (field selection, limits, summaries).
- 20, pagination, filtering and output-size controls.
- 20, actionable, documented error responses, codes and messages an agent can recover from.
- 20, idempotency or safe retries, and for MCP the `readOnlyHint` and `destructiveHint` annotations.
- 15, sensible defaults, few required parameters, and official SDKs in at least two languages.

Models are read for tool use, structured output, prompt caching, context length, batch and SDKs. Frameworks for how much code and how many defaults a tool-calling agent with MCP needs.

## 5. Maintenance & community, 48 out of 100, up to 4.6 more on the total

Why it scored 48: The brief counts the last published model for a closed service. `alibaba/wan-3` was created on 15 September 2026, 23 days before this check (30). The public changelog's newest entry is 21 April 2026, 170 days ago, which would score 10 on a strict reading, as the Replicate image listing was scored. No dated changelog entries in the last 90 days, and we could date only one model publication in that period (0). The changelog is stale, replicate-python has 70 open issues and was last pushed on 26 August 2025, and support is a contact form (5 of 25). Official Python, JavaScript and MCP packages exist, but the stable releases are `replicate` 1.4.0 on npm from 17 November 2025, 1.0.7 on PyPI from 27 May 2025 and `replicate-mcp` 0.9.0 from 9 June 2025, with 2.0 pre-releases stopping on 18 December 2025 (10). Package health and CI not audited (3).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-maintenance):

- 0 to 30, time since the last release, or the last published model or API change for a closed service. 30 within 30 days, 20 within 90, 10 within 180, 0 older.
- 20, at least three releases or dated changelog entries in the last 90 days.
- 0 to 25, responsiveness. Issues and pull requests answered on GitHub (the open issues and how recent the replies are). For closed services, a public changelog and a support or community channel that answers, 0 to 15.
- 15, presence in the official MCP registry under a verified namespace (MCP servers), or current official SDKs (APIs and models).
- 10, package health, current dependencies and CI.

Models are read for deprecation notice periods and model churn rather than release counts.

## 6. Transparency & trust, 72 out of 100, up to 2.5 more on the total

Made of editorial 56, provenance 87.

Why it scored 72: Closed platform under terms of service dated 1 April 2026 that name Replicate, LLC. The clients are Apache-2.0 and each model keeps its own licence, for example Apache 2.0 for Wan 2.7 (17). The docs say API inputs, outputs, files and logs are deleted after one hour and website predictions are kept until deleted. The privacy policy gives no retention period, and we found no statement on what the outside providers behind official models (Google, Kuaishou, Runway and others) keep (18). No deprecation policy. The terms say Replicate doesn't promise to keep any model and may discontinue parts of the service without notice, while the official-models page promises a stable interface (6). A subprocessor page lists 17 companies, each located in the United States. The model providers aren't on it (15).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-transparency):

- 0 to 30, source availability and licence clarity. 30 for open source under an OSI licence, 15 for closed with clear terms, 0 for unclear terms.
- 0 to 30, data handling and retention statements that agree with each other (privacy policy, DPA, retention periods, subprocessors).
- 0 to 20, a deprecation policy or notices with dates.
- 0 to 20, telemetry disclosed with an opt-out (local software), or subprocessors and data locations disclosed (hosted).

The other half of Transparency and trust is the provenance score, computed from checked facts (below). The category score is the mean of the two.

Provenance checks not met in full (half of this category, computed from checked facts):

- Terms of service: read, states 7 of the 7 things a reader expects, and has 1 clause that costs points (8 of 10)
- Privacy policy: read, states 7 of the 8 things a reader expects (9.3 of 10)
- security.txt: not found (0 of 10)

## 7. Schema & documentation, 86 out of 100, up to 2.3 more on the total

Why it scored 86: Public OpenAPI document with 26 paths and 36 operations, and each model publishes its own input and output schema (25). llms.txt and Markdown versions of docs pages (10). The collection page says which model suits which job and the model pages describe each input, though few say when not to use a model (15). Inputs are typed with enums and bounds, such as Veo 3.1 `duration` of 4, 6 or 8 and Kling 3.0 `duration` from 3 to 15. Kling's `multi_prompt` is a JSON array passed as a string (13). Examples throughout and 8 coded errors with fixes. The HTTP error body is shown only for 429 (11). The API is versioned at `/v1` and official models are promised a stable interface, but the public changelog's newest entry is 21 April 2026 (12).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-schema):

APIs and MCP servers.

- 25, a machine-readable contract (a public OpenAPI file or similar; for MCP, typed JSON Schema inputs on every tool).
- 10, llms.txt or Markdown docs served for agents.
- 0 to 20, descriptions that say what a tool is for, when to use it and when not to, read from the tool definitions in the source or the API reference.
- 0 to 15, typed inputs with enums, constraints and required fields, and no free-form JSON blobs.
- 0 to 15, examples and documented error responses.
- 15, versioning and a public changelog.

Models are read from the API reference, the OpenAPI file, llms.txt, the structured-output and tool-use docs and the model cards. Frameworks from docs a model can follow, typed interfaces, examples and the API reference.

## What we couldn't check

What we couldn't read counted as absent. Publishing it on a page a plain HTTP fetch can read (not only in a browser) lets the next check count it.

- unchecked: the September 2026 incidents on the two other days the status component counts. Cloudflare's incident feed reached back only to 21 September, so we read three incidents covering three of the five days.
- unchecked: the MCP server's tool definitions and annotations. The `replicate-mcp` package names no source repository on npm, and we didn't open the package.
- Wan 3.0's billing table showed $0.025, $0.05 and $0.10 a second on 8 October 2026, half the README's $0.05, $0.10 and $0.20, under a banner announcing 30 per cent off for the week. We couldn't tell which price applies after the promotion.
- Whether Cloudflare's certifications and security programme cover Replicate. Nothing on the Replicate pages we read says so.
- What the outside providers behind official models keep of prompts, input images and outputs. Not found in the reviewed documentation.
- Maintenance takes 30 for recency from the newest official model (15 September 2026). The Replicate image listing scored the same line from the changelog date, which would give 10 here and a total of 28.
- Job times per model were not established. Replicate publishes none, and we ran no predictions.

## Weaknesses

- Tokens have no scopes, expiry or spend cap, and any token can run every model and delete account resources
- No SLA is published and the terms disclaim service levels. Replicate may withdraw a model without notice
- No idempotency key on prediction creates, and a 429 carries a wait time in the body with no `Retry-After` header documented
- The public changelog's newest entry is 21 April 2026, and the stable npm and PyPI clients date from November and May 2025
- No security.txt, and no certification is named on the pages we read

## What costs an agent a turn today

The notes we give agents before they call it. Each one is a workaround an agent shouldn't need.

- Call official models at `POST /v1/models/{owner}/{name}/predictions` with no version, then poll `GET /v1/predictions/{id}` or pass `webhook`. Video jobs usually outlast the 60-second `Prefer: wait` window
- Download the output file within the hour. API predictions lose inputs, outputs and logs after 60 minutes
- Read the model's price tiers before setting `resolution`, `mode` or `generate_audio`. Kling 3.0 runs from $0.168 to $0.42 a second and Veo 3.1 doubles with audio
- Send `Cancel-After` on every create. There is no idempotency key, so a blind retry is a second paid video
- Keep credit above $20 or enable auto-reload. Low balances are throttled, and granted credit with no card is held to 6 requests a minute

## When it's done

Send what changed and where it's published as a dispute (https://www.anchorterminal.com/builders/#disputes, or `POST https://www.anchorterminal.com/api/v1/contact` with `"kind": "dispute"`). Disputes are answered in public, and the listing is checked again by the same checklist. Paying for an audit or a listing claim changes nothing here.

What we couldn't check

  • unchecked: the September 2026 incidents on the two other days the status component counts. Cloudflare's incident feed reached back only to 21 September, so we read three incidents covering three of the five days.
  • unchecked: the MCP server's tool definitions and annotations. The replicate-mcp package names no source repository on npm, and we didn't open the package.
  • Wan 3.0's billing table showed $0.025, $0.05 and $0.10 a second on 8 October 2026, half the README's $0.05, $0.10 and $0.20, under a banner announcing 30 per cent off for the week. We couldn't tell which price applies after the promotion.
  • Whether Cloudflare's certifications and security programme cover Replicate. Nothing on the Replicate pages we read says so.
  • What the outside providers behind official models keep of prompts, input images and outputs. Not found in the reviewed documentation.
  • Maintenance takes 30 for recency from the newest official model (15 September 2026). The Replicate image listing scored the same line from the changelog date, which would give 10 here and a total of 28.
  • Job times per model were not established. Replicate publishes none, and we ran no predictions.

Sources 40

  1. video collection and recommended models replicate.com · seen 2026-10-08
  2. pricing replicate.com · seen 2026-10-08
  3. Veo 3.1 model page, price tiers and input schema replicate.com · seen 2026-10-08
  4. Veo 3.1 Fast price tiers replicate.com · seen 2026-10-08
  5. Kling 3.0 price tiers and input schema replicate.com · seen 2026-10-08
  6. Wan 3.0 price tiers, README and creation date replicate.com · seen 2026-10-08
  7. Runway Gen-4.5 price and input schema replicate.com · seen 2026-10-08
  8. Seedance 2.0 price tiers replicate.com · seen 2026-10-08
  9. Hailuo 2.3 price tiers replicate.com · seen 2026-10-08
  10. Wan 2.7 price and licence replicate.com · seen 2026-10-08
  11. try-for-free collection replicate.com · seen 2026-10-08
  12. OpenAPI document api.replicate.com · seen 2026-10-08
  13. HTTP API reference replicate.com · seen 2026-10-08
  14. docs index for agents replicate.com · seen 2026-10-08
  15. rate limits and 429 body replicate.com · seen 2026-10-08
  16. sync mode, webhooks and `Cancel-After` replicate.com · seen 2026-10-08
  17. error codes replicate.com · seen 2026-10-08
  18. official models replicate.com · seen 2026-10-08
  19. billing, failed and cancelled runs, free limits replicate.com · seen 2026-10-08
  20. prepaid credit replicate.com · seen 2026-10-08
  21. API tokens and token scanning replicate.com · seen 2026-10-08
  22. data retention replicate.com · seen 2026-10-08
  23. subprocessors replicate.com · seen 2026-10-08
  24. MCP server docs replicate.com · seen 2026-10-08
  25. hosted MCP server mcp.replicate.com · seen 2026-10-08
  26. status component and 90-day history cloudflarestatus.com · seen 2026-10-08
  27. incident of 14 and 15 September 2026 cloudflarestatus.com · seen 2026-10-08
  28. incident of 24 September 2026 cloudflarestatus.com · seen 2026-10-08
  29. incident of 28 September 2026 cloudflarestatus.com · seen 2026-10-08
  30. changelog replicate.com · seen 2026-10-08
  31. terms of service replicate.com · seen 2026-10-08
  32. privacy policy replicate.com · seen 2026-10-08
  33. enterprise page replicate.com · seen 2026-10-08
  34. security.txt (404) replicate.com · seen 2026-10-08
  35. npm client versions and dates registry.npmjs.org · seen 2026-10-08
  36. replicate-mcp versions and dates registry.npmjs.org · seen 2026-10-08
  37. PyPI client versions and dates pypi.org · seen 2026-10-08
  38. replicate-python stars, issues and last push api.github.com · seen 2026-10-08
  39. Cloudflare announcement replicate.com · seen 2026-10-08
  40. domain registration rdap.verisign.com · seen 2026-10-08

Probe metrics

Not measured yet. Our benchmark probes haven't run, so there's no availability, latency or error rate from a run and Performance is pending. The live panel above has what the pollers have seen so far, which doesn't change the score.

Pricing & changes

Pay per use Pay per use Pay as you go from prepaid credit (valid one year, not refundable) or monthly in arrears. Official video models bill per second of output video, with tiers by resolution, mode and audio shown on each model page. Community models bill per second of hardware time. Failed runs aren't charged, and a cancelled official-model run may be. The try-for-free collection lets a new account run `minimax/video-01` a limited number of times before buying credit (https://replicate.com/collections/try-for-free).

Prices

ItemPriceUnitNote
Veo 3.1 with audio$0.40per second of video$0.20 without audio
Veo 3.1 Fast with audio$0.15per second of video$0.10 without audio
Kling 3.0 pro (1080p)$0.224per second of video$0.336 with audio, standard 720p $0.168, 4K $0.42
Runway Gen-4.5$0.12per second of video
Seedance 2.0 720p$0.18per second of videowithout video input. 480p $0.08, 1080p $0.45, 4K $1.00, more with video input
Wan 2.7 text-to-video$0.10per second of video
Wan 3.0 1080p$0.10per second of videobilling table on 8 October 2026 during a promotion. README lists $0.20
Hailuo 2.3 768p$0.0467per second of video$0.28 per 6-second video, $0.56 for 10 seconds, $0.49 for 6 seconds at 1080p

Compared across listings on the price index.

Recent changes

  • Latest release

Follow them as a feed at /feeds/tools/replicate-video.xml, or this listing's score history at history.json.

Connect

First request

curl -X POST https://api.replicate.com/v1/models/google/veo-3.1/predictions \
  -H "Authorization: Bearer $REPLICATE_API_TOKEN" -H "Content-Type: application/json" \
  -d '{"input":{"prompt":"a red fox running through fresh snow","duration":4,"resolution":"720p"}}'

Claude Code

claude mcp add replicate https://mcp.replicate.com/sse --transport sse --scope user

MCP client configuration

{
  "mcpServers": {
    "replicate": {
      "args": [
        "-y",
        "replicate-mcp@latest"
      ],
      "command": "npx",
      "env": {
        "REPLICATE_API_TOKEN": "${REPLICATE_API_TOKEN}"
      }
    }
  }
}
Similar toolGrade ScoreShared capabilitiesx402
Runway API RunwayC60.8video.generate video.image-to-video video.edit video.audio video.extendno
Pika API Pika (Mellis, Inc.)D47.3video.generate video.image-to-video video.edit video.audio video.extendno
PixVerse API PixVerseE42.4video.generate video.image-to-video video.edit video.audio video.extendno
Kling AI API Kling AI (Kuaishou)F21.7video.generate video.image-to-video video.edit video.audio video.extendno
fal video models fal (Features & Labels, Inc.)B68.9video.generate video.image-to-video video.edit video.audiono
Alibaba Wan (Model Studio) Alibaba CloudB65.1video.generate video.image-to-video video.edit video.audiono

Machine-readable

Verify this listing

For the vendor

Is this your product? Link to this page from your own site or README, then tell us where. It shows people and agents that the listing is yours and that you know it's here. It never changes a grade, rank or review.

  1. Add the badge or a link

    Replicate video models on Anchor Terminal, B, 62.1/100
    On a light page
    On a dark page
    <a href="https://www.anchorterminal.com/tools/replicate-video"><img src="https://www.anchorterminal.com/badges/replicate-video.svg" alt="Replicate video models on Anchor Terminal" height="20"></a>
    [![Replicate video models on Anchor Terminal](https://www.anchorterminal.com/badges/replicate-video.svg)](https://www.anchorterminal.com/tools/replicate-video)

    It counts on a page on replicate.com or one of its subdomains, or the README of github.com/replicate/replicate-python.

  2. Tell us where it is

    We read it once now and again every week. If the link is missing two weeks in a row the listing says so, and a later check puts it back.

Agents send the same to POST /api/v1/verify as {"slug": "replicate-video", "url": "…"}, or call the verify_listing tool at /mcp. Ten checks an hour from one address. What we check. To announce the listing, get sharing assets for social media.

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.