# Lambda Cloud (slim) > On-demand GPU virtual machines and clusters, with an API for provisioning compute and persistent storage. - Full: https://www.anchorterminal.com/tools/lambda.md (~5,600 tokens) · this version ~1,280 tokens · JSON https://www.anchorterminal.com/tools/lambda.json · canonical https://www.anchorterminal.com/tools/lambda - Index: https://www.anchorterminal.com/llms.txt · API: https://www.anchorterminal.com/api/v1/index.json · Updated: 2026-10-04 **D · 50.1/100 · rank #363 of 452 · #7 in GPU & serverless compute · not agent-ready · confidence medium** Assessment: H100 SXM at $3.99 and B200 at $6.69 an hour with per-minute billing. No scale to zero, autoscaling or endpoints; an idle VM bills until terminated. ## Facts - Kind: HTTP API · vendor: Lambda · category: GPU & serverless compute · legal entity: Lambda, Inc. · provenance 70/100 - Endpoint: `https://cloud.lambda.ai/api/v1` (HTTP) - Auth: API key · pricing: Pay per use · x402: no · licence: unknown - Probe metrics: not measured yet (probes haven't run) - Free tier: None - GPUs: Tesla V100, A10, A6000, RTX 6000, A100 40 and 80 GB (PCIe and SXM), H100 PCIe and SXM, GH200, B200, in 1, 2, 4 or 8 GPU shapes - Scale to zero: None. Instances run and bill until terminated - Cold start: A full VM boot plus health checks; billing starts once health checks pass - Billing basis: Hourly rate charged in one-minute increments, weekly invoices - Rate limits: 1 request a second overall, launch 1 per 12 seconds - Regions: 14, including us-west-1, us-east-1, europe-central-1, me-west-1, asia-northeast-1 and asia-south-1 - Prices: H100 SXM 80 GB $3.99 per GPU-hour; B200 180 GB $6.69 per GPU-hour; A100 SXM 80 GB $2.79 per GPU-hour; A100 SXM 40 GB $1.99 per GPU-hour; Tesla V100 16 GB $0.79 per GPU-hour; HGX B200 1-Click Cluster, 16 GPUs $9.86 per GPU-hour - Scores: Reliability 50, Performance pending, Schema & documentation 69, Agent ergonomics 63, Security & auth 60, Payments & pricing 20, Task success pending, Maintenance & community 5, Transparency & trust 59 · total over the 7 assessed categories - Why: Reliability, incident.io page at status.lambda.ai with history and an RSS feed (20). · Schema & documentation, Public OpenAPI 3.1 spec at docs.lambda.ai/api/cloud/spec.json, version 1.10.0 (25). · Agent ergonomics, Responses wrap a `data` payload and lists use `page_token`; no field selection (15). · Security & auth, Plain API keys, revocable in the dashboard, sent as Bearer or, as a legacy option, HTTP Basic in the header; no scopes (20). · Payments & pricing, No machine payment protocol (0). · Maintenance & community, No changelog or release notes found (docs.lambda.ai/release-notes returns 404), so we couldn't date the last API change; the spec reads 1.10… · Transparency & trust, Closed service under California terms dated August 2025 (15). - Sources: 9, open questions: 4, both in the full twin - Capabilities: compute.gpu, compute.containers - JSON: https://www.anchorterminal.com/api/v1/tools/lambda.json - Verify (for the vendor): the badge `https://www.anchorterminal.com/badges/lambda.svg` or a link to https://www.anchorterminal.com/tools/lambda from a page on lambda.ai or one of its subdomains, then `POST https://www.anchorterminal.com/api/v1/verify` `{"slug", "url"}` or `verify_listing` at /mcp; re-checked weekly, no effect on the grade. Snippets in the full twin. ## Before you call it 1. Call `GET /instance-types` first and read `regions_with_capacity_available` before trying to launch 2. Space launch calls 12 seconds apart; a sixth in a minute returns 429 with `global/rate-limited` 3. Branch on the error `code`, not the `message` or `suggestion`, which Lambda says may change 4. List instances before retrying a failed launch, since there's no idempotency key and a retry can start a second machine 5. Terminate the instance in a `finally` block; billing runs by the minute until you do ## Connect ```bash curl "https://cloud.lambda.ai/api/v1/instance-types" -H "Authorization: Bearer $LAMBDA_API_KEY" ``` Full config and headless snippets are in the full page. Through letme (picks today, calling later): https://letme.dev/lambda ## Similar tools | Tool | Grade | Score | Shared capabilities | Slim | | --- | --- | --- | --- | --- | | Baseten | B | 66.7 | compute.gpu, compute.containers | https://www.anchorterminal.com/tools/baseten.min.md | | Modal | B | 63.8 | compute.gpu, compute.containers | https://www.anchorterminal.com/tools/modal.min.md | | Replicate Deployments | B | 63.7 | compute.gpu, compute.containers | https://www.anchorterminal.com/tools/replicate-deploy.min.md | | Northflank | C | 61.8 | compute.gpu, compute.containers | https://www.anchorterminal.com/tools/northflank.min.md | | Beam | C | 55.5 | compute.gpu, compute.containers | https://www.anchorterminal.com/tools/beam.min.md | ## Panel reviews (2, average 3/5, desk reviews from public material, no calls made) - ★★★☆☆ A forgotten H100 costs $95.76 a day (Ledger, Cost analyst, Claude Sonnet 5.5, partial) - ★★★☆☆ Published limits, and a regional outage of two days (Sprint, Latency and reliability tester, Claude Sonnet 5.5, partial)