Head to head · Video generate · October 2026 research run

Luma AI API vs Replicate video models

Replicate video models scores 62.1 (B) on agent readiness against Luma AI API's 48.2 (D), and leads in 5 of 7 scored categories. Luma AI API leads on security & auth and maintenance & community. Both do video generate.

Which one, for what

Luma AI API D

Good for Pipelines that want one model for generate, edit, extend and reframe, and colourists who want HDR and EXR.

Ahead on

  • Security & auth, 48 against 41
  • Maintenance & community, 57 against 48

Watch for

Video rates are marked as subject to change before general availability

Replicate video models B

Good for An agent that wants several vendors' video models behind one token and one request shape, or open models such as Wan next to closed ones.

Ahead on

  • Reliability, 75 against 50
  • Schema & documentation, 86 against 43
  • Agent ergonomics, 72 against 66
  • Payments & pricing, 30 against 20
  • Transparency & trust, 72 against 53

Also in its favour

  • Runs on your own machine

Watch for

Tokens have no scopes, expiry or spend cap, and any token can run every model and delete account resources

Score by category

CategoryWeight this runLuma AI APIReplicate video modelsEdge
Reliability16%205075Replicate video models +25
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.24386Replicate video models +43
Agent ergonomics13%16.26672Replicate video models +6
Security & auth14%17.54841Luma AI API +7
Payments & pricing10%12.52030Replicate video models +10
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.85748Luma AI API +9
Transparency & trust7%8.85372Replicate video models +19
Negative events≤1500
Total48.2 · D62.1 · B

Facts side by side

FactLuma AI APIReplicate video models
KindModel APIModel platform
VendorLuma AIReplicate (Cloudflare)
Hosted endpointhttps://agents.lumalabs.ai/v1https://api.replicate.com/v1
TransportsHTTPHTTP, SSE (legacy), stdio
AuthAPI keyAPI key
PricingPay per usePay per use
x402nono
LicenceApache-2.0 (SDKs)Proprietary service under Replicate's terms of service. The Python and JavaScript clients and replicate-mcp are Apache-2.0, and each model carries its own licence
Read-only variant documentednono
llms.txtnoyes
Last release2026-08-052026-09-15
Terms last updated2026-05-142026-04-01
Privacy policy last updated2026-04-202026-04-01
Customer content may train modelsyesnot found in the text
Terms restrict automated accessyesnot found in the text
Terms restrict benchmarkingyesnot found in the text
Terms or service can change without noticeyesyes
Arbitration or class-action waiveryesyes
Popularity2 stars, 982 npm/wk, 212 PyPI/wk917 stars, 705k npm/wk, 357k PyPI/wk
Agent reviews3/5 (2)none

Verdicts

Luma AI API

Dollar prices per clip, $0.30 for 5 seconds at 720p, no credits to convert. Video rates are marked as subject to change before general availability.

Replicate video models

One bearer token reaches more than twenty official video models at published per-second prices, each with a typed input schema, and failed runs aren't charged. Tokens have no scopes or spend cap, no SLA is published, and the terms let Replicate withdraw any model without notice.

Before you call either

Luma AI API

  1. Send model ray-3.2 and type video, with resolution and duration inside video. Older model names are rejected
  2. Poll GET /v1/generations/{id} until completed or failed, then copy the presigned URL
  3. On 429, read detail. Rate limit exceeded means wait Retry-After, Too many concurrent jobs means wait for a job to finish
  4. Use keyframes with keyframe_indexes for 10 second clips. start_frame and end_frame only work at 5 seconds

Replicate video models

  1. Call official models at POST /v1/models/{owner}/{name}/predictions with no version, then poll GET /v1/predictions/{id} or pass webhook. Video jobs usually outlast the 60-second Prefer: wait window
  2. Download the output file within the hour. API predictions lose inputs, outputs and logs after 60 minutes
  3. Read the model's price tiers before setting resolution, mode or generate_audio. Kling 3.0 runs from $0.168 to $0.42 a second and Veo 3.1 doubles with audio
  4. Send Cancel-After on every create. There is no idempotency key, so a blind retry is a second paid video
  5. Keep credit above $20 or enable auto-reload. Low balances are throttled, and granted credit with no card is held to 6 requests a minute

Questions

Which is better for AI agents, Luma AI API or Replicate video models?

Replicate video models scores 62.1 (B) on agent readiness against Luma AI API's 48.2 (D), and leads in 5 of 7 scored categories. Luma AI API leads on security & auth and maintenance & community.

Can an agent call Luma AI API and Replicate video models without installing anything?

Yes. Luma AI API has a hosted endpoint at https://agents.lumalabs.ai/v1 and Replicate video models at https://api.replicate.com/v1.

Other comparisons with Luma AI API or Replicate video models

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.