# Nebius AI Cloud > Nebius AI Cloud rents NVIDIA GPU virtual machines and InfiniBand clusters, with managed Kubernetes, Slurm and Serverless AI jobs and endpoints for containers. Resources are managed through REST and gRPC APIs, a CLI, a Terraform provider and SDKs. - Canonical: https://www.anchorterminal.com/tools/nebius-ai-cloud - Markdown: https://www.anchorterminal.com/tools/nebius-ai-cloud.md (~8,500 tokens) - Slim: https://www.anchorterminal.com/tools/nebius-ai-cloud.min.md (~1,930 tokens, same facts, less prose, for token-sensitive contexts) - JSON: https://www.anchorterminal.com/tools/nebius-ai-cloud.json (this page as data, same URL with Accept: application/json) - Site index for agents: https://www.anchorterminal.com/llms.txt (full text: https://www.anchorterminal.com/llms-full.txt) - API: https://www.anchorterminal.com/api/v1/index.json - Updated: 2026-10-09 ## Overview **Grade B · 67.2/100 · rank #240 of 842 · #1 in GPU & serverless compute · not agent-ready · confidence medium** More from Nebius, listed separately because each is its own product: [Nebius Token Factory fine-tuning](https://www.anchorterminal.com/tools/nebius-token-factory-fine-tuning.md) (Fine-tuning). ## Assessment One API definition generates the REST and gRPC interfaces, the CLI, Terraform provider and three SDKs, with a 602-operation OpenAPI document, `X-Idempotency-Key` and role-scoped service accounts. The status page lists 14 major incidents between 14 July and 8 October 2026, no request rate limits were found, and signup needs a browser and a card. ## Facts | Field | Value | | --- | --- | | Vendor | Nebius (https://nebius.com) | | Kind | HTTP API | | Category | GPU & serverless compute (https://www.anchorterminal.com/categories/gpu-compute) | | Transport | HTTP, stdio | | Endpoint | `https://api.nebius.cloud` | | Auth | OAuth or key · Self-serve. A person signs up in the web console with a Google, GitHub or Microsoft account. Every API call takes `Authorization: Bearer` with an access token valid for 12 hours. A user gets one from `nebius iam get-access-token`. A service account uploads an RSA public key (an authorised key, with optional expiry), signs a five-minute RS256 JWT and exchanges it at `https://auth.eu.nebius.com/oauth2/token/exchange`. Permissions come from group roles (`auditor`, `viewer`, `editor`, `admin` and service roles) granted on a tenant, project or resource. Serverless AI endpoints take their own token set in `spec.authToken`. | | Pricing | Pay per use (Pay per use) · Pay as you go, billed by the second, with no free tier or trial found. On-demand per GPU-hour from 1 October 2026, B300 $9.50, B200 $8.50, H200 $5.40, H100 $4.50, RTX PRO 6000 $1.80, and L40S $1.35 plus vCPU and RAM. Preemptible GPUs are spot priced from $0.79. Serverless AI bills at Compute prices and a stopped endpoint bills nothing. Adding a card charges $25 to the balance. Commitment discounts go through sales (https://docs.nebius.com/compute/resources/pricing, https://nebius.com/prices). | | x402 | No · No x402, MPP or L402 in the docs index, the OpenAPI document or the price list (checked 2026-10-08). | | Licence | Proprietary service under the Nebius Services Agreement. The API definitions, the Go, Python and JavaScript SDKs and the MCP server on GitHub are MIT | | Packages | pypi: `nebius`; npm: `@nebius/js-sdk`; go: `github.com/nebius/gosdk` | | Source | https://github.com/nebius/api | | Docs | https://docs.nebius.com/ | | llms.txt | https://docs.nebius.com/llms.txt | | Last release | 2026-10-07 | | npm downloads / week | 4,137 | | PyPI downloads / week | 468,852 | | Free tier | None found. A card added at signup is charged $25, which goes to the account balance. Promo codes are issued in promotions | | GPUs | NVIDIA B300, B200, H200 and H100 (NVLink), RTX PRO 6000 and L40S, plus CPU-only AMD and Intel platforms. GB300 and GB200 NVL72 racks through sales | | Interfaces | REST at `https://api.nebius.cloud` (OpenAPI 3.0.3, 602 operations), gRPC on per-service hosts such as `compute.api.nebius.cloud:443`, the `nebius` CLI, a Terraform provider and SDKs for Go, Python and JavaScript, all generated from one API | | Serverless AI | Devlabs, jobs and endpoints that run a container image on a managed container VM. Endpoints get a managed HTTPS URL and optional token authentication (https://docs.nebius.com/serverless/overview) | | Scale to zero | None automatic. An endpoint is stopped and started by a call, and a stopped endpoint bills for neither compute nor storage. No autoscaling found | | Billing basis | Hourly prices charged by the second. Volumes bill per GiB for 730 hours whether attached or not | | Idempotency and retries | `X-Idempotency-Key` on modifying calls, `resourceVersion` for concurrent updates, and a `retry_type` on error details (https://github.com/nebius/api) | | Rate limits | No request rate limits found. Default quotas per region include 12 regular and 8 preemptible GPU VMs and 32 H200 GPUs in `eu-north1` (https://docs.nebius.com/compute/resources/quotas-limits) | | SLA | 99.5 per cent a month per virtual machine, with credits of 10, 15 or 30 per cent. Serverless AI is not on the list of services with a service level (https://docs.nebius.com/legal/sla-levels) | | Regions | Nine public regions in Finland, France (two), Spain, Israel, the United Kingdom (two) and the United States (Missouri and Minnesota), plus a private region in Iceland | | MCP servers | A keyless docs search server at `https://docs.nebius.com/mcp`, and a beta local server (`nebius/mcp-server`, four tools) that runs CLI commands with a safe mode on by default | | Certifications | SOC 2 Type II with HIPAA, SOC 3, ISO 27001, 27018 and 22301, and CSA STAR Level 1 per https://nebius.com/trust-center | | Audit | Audit Logs (preview, free) with control-plane and data-plane events, an API at `/audit/v2/audit-events` and export to Object Storage | | Capabilities | compute.gpu, compute.endpoints, compute.batch, compute.containers | | Tags | hosted, usage-priced, openapi, llms-txt, grpc, terraform, mcp, python, typescript, go, status-page, soc2, enterprise | | JSON | https://www.anchorterminal.com/api/v1/tools/nebius-ai-cloud.json | ## Score breakdown (methodology v0.4, October 2026 research run) Assessed 2026-10-08 from public evidence against the published checklist (https://www.anchorterminal.com/benchmark/#checklist). Confidence: medium. Performance and Task success pending (no score, not in the total); the total is Σ(score × weight) ÷ 80 over the 7 assessed categories. "This run" is each category's share of the 100 points. | Category | Weight | This run | Score (0–100) | Points | | --- | --- | --- | --- | --- | | Reliability | 16% | 20 | 57 | 11.4 | | Performance | 10% | pending | pending | n/a | | Schema & documentation | 13% | 16.2 | 78 | 12.7 | | Agent ergonomics | 13% | 16.2 | 77 | 12.5 | | Security & auth | 14% | 17.5 | 81 | 14.2 | | Payments & pricing | 10% | 12.5 | 20 | 2.5 | | Task success | 10% | pending | pending | n/a | | Maintenance & community | 7% | 8.8 | 82 | 7.2 | | Transparency & trust (editorial 70, provenance 83) | 7% | 8.8 | 77 | 6.7 | | Negative events | up to −15 | up to −15 | none recorded | 0 | | **Total** | | | | **67.2 → B** | ### Why each score - Reliability 57: Graded as a hosted service on the REST API. Statuspage at status.nebius.com with component history (20). The incident feed lists 22 incidents between 14 July and 8 October 2026, 14 marked major, among them about 21.5 hours of partial degradation in us-central1 on 19 August (with a published post-mortem), about 12 hours of network timeouts in eu-north1 on 22 July and a power fault on several dozen nodes for about 14.7 hours on 7 and 8 October (0). Resource quotas are published with numbers per region, but no request rate limits were found in the reviewed documentation (5 of 15). `RESOURCE_EXHAUSTED` maps to HTTP 429, errors can carry a `TooManyRequests` detail and a `retry_type` of `CALL`, `UNIT_OF_WORK` or `NOTHING`, and `X-Idempotency-Key` covers modifying calls; no Retry-After or backoff intervals found (12 of 15). SLA of 99.5 per cent a month per virtual machine with credits of 10 to 30 per cent; Serverless AI is not on the list of services with a service level (10). Compute and Serverless AI are generally available; Audit Logs, Monitoring and Logging are in preview (10). - Performance: Pending. Latency is measured per call by our probes, which haven't run yet, so this run doesn't score it. Its weight is shared across the assessed categories until the first probe window closes. - Schema & documentation 78: OpenAPI 3.0.3 at api.nebius.cloud/openapi.json with 602 operations on 443 paths and 1,051 schemas, plus the protobuf definitions in `nebius/api` (25). llms.txt, llms-full.txt, every docs page as Markdown and a docs MCP server (10). Every operation has a description, most of one line such as "Creates an endpoint"; field descriptions state mutual exclusions and defaults but rarely when not to call (12). 186 enums and `required` lists on resource metadata; platform and preset are plain strings (11). The docs give curl examples and a table of 15 status codes with fixes, but the OpenAPI document has no examples and documents only 200 responses (8). Paths carry `v1` or `v1alpha1` with a stated rule that alpha versions can change, and the CLI release notes are dated and generated from the same API; the spec's own version reads `version not set` and there is no API changelog as such (12). - Agent ergonomics 77: Lists take `pageSize` and `pageToken`, gets take a `view` and there are get-by-name calls; no field selection (15 of 25). Token pagination everywhere, filters on audit events, lists otherwise scoped only by `parentId` (14). Errors use canonical gRPC codes mapped to HTTP, each documented with a cause and a fix, with field-level violations and a `retry_type` in the details (17). `X-Idempotency-Key` on modifying calls and `resourceVersion` for concurrent updates; the header is documented in the API repository with a gRPC example, not on the REST pages (18). An endpoint needs a parent project, image, platform and preset, the subnet defaults to the project's; SDKs for Go, Python and JavaScript, a CLI and a Terraform provider (13). - Security & auth 81: User tokens last 12 hours. Service accounts authenticate with an uploaded RSA public key (optional expiry), sign a five-minute JWT and exchange it for a 12-hour Bearer token; access comes from group roles granted on a tenant, a project or a single resource. No secret travels in a URL (28). `auditor` and `viewer` roles, narrow roles such as `compute.instance-power-operator`, and a safe mode on by default in the beta MCP server that blocks update and delete commands; no confirmation step for destructive API calls (15). Returns infrastructure metadata and the output of the owner's own containers (10). Audit Logs record control-plane and data-plane events in CloudEvents form, readable at `/audit/v2/audit-events` and exportable to Object Storage; the service is in preview (13). security.txt valid until 31 December 2027, a trust centre listing SOC 2 Type II with HIPAA, SOC 3, ISO 27001, 27018 and 22301 and CSA STAR Level 1, and published rules for customer security scans; no bug bounty found (15). - Payments & pricing 20: No machine payment protocol (0). Per-GPU-hour, per-vCPU-hour and per-GiB prices published without a login, billed by the second (20). No free tier or trial found; adding a card charges $25 to the account balance, and promo codes need billing details first (0). Signup is a browser flow through a Google, GitHub or Microsoft account, after which a service account can work unattended (0). - Task success: Pending. Task success needs the category task suites run through each tool, which haven't run yet, so this run doesn't score it. Its weight is shared across the assessed categories until then. A data provider's data-quality score is published on its listing now and becomes half of this category when it's scored. - Maintenance & community 82: Python SDK 0.6.20 on 7 October 2026 and CLI 0.12.287 on 6 October (30). 41 PyPI releases since 10 July and CLI releases most working days (20). Closed service with dated release notes, a support centre in the console and a published support regulation; we did not read GitHub issue response times (10 of 15). Current official SDKs for Go, Python and JavaScript (15). Protobuf definitions published to GitHub within a day of each CLI release and a CI badge on `nebius/api`; we did not check the CI results (7). - Transparency & trust 77: Closed service under a Services Agreement published 15 September 2026 and effective 28 September, with nine earlier versions archived; the API definitions, SDKs and MCP server are MIT (15). Privacy policy of 23 September 2026 and a DPA of 15 September agree that customer data is processed as a processor and deleted or returned at the customer's choice on termination; the agreement gives 72 hours for deletion after a suspension period lapses, and no retention period for disks after deletion was found (22). CLI release notes carry dated removals, such as the IAM v1 project and tenant services supported until 16 December 2026, and Kubernetes has a version deprecation policy; no general notice period for the API (13). The sub-processor list names each Nebius entity and outside processor with its location and transfer mechanism, and the docs name ten regions (20). Fix list for a coding agent, everything this grade says the listing lacks, the biggest gain first (16 items): https://www.anchorterminal.com/fixes/nebius-ai-cloud.md (JSON https://www.anchorterminal.com/fixes/nebius-ai-cloud.json) ### What we couldn't check - No request rate limits were found in the docs index, the REST pages or the API repository. They may exist in pages we did not read. - `X-Idempotency-Key` is documented in the API repository with a gRPC example. We did not confirm on a live call that the REST gateway honours it. - unchecked: GitHub issue response times and CI results for nebius/api and the SDK repositories. We read the clones, not the issue tracker. - unchecked: the retention period for Audit Logs events and for disks after deletion. Neither was found in the pages read. - The Services Agreement bars competitive analysis or benchmarking and unauthorised probing, which matters before any Anchor Terminal probe runs. - The lead describes serverless endpoints. Serverless AI endpoints are single container VMs started and stopped by a call, with no autoscaling found, so `compute.serverless` is left out of the capabilities. - Nebius Token Factory and Tavily are separate Nebius products with their own docs and terms and are not graded here. ### Sources - docs index for agents (llms.txt): (seen 2026-10-08) - OpenAPI document: (seen 2026-10-08) - REST API authentication: (seen 2026-10-08) - REST API errors: (seen 2026-10-08) - developer tools and versioning: (seen 2026-10-08) - API repository README (idempotency, reset mask, error details): (seen 2026-10-08) - Serverless AI overview: (seen 2026-10-08) - Serverless AI endpoints quickstart: (seen 2026-10-08) - Compute pricing: (seen 2026-10-08) - price list: (seen 2026-10-08) - Compute quotas: (seen 2026-10-08) - signup and billing: (seen 2026-10-08) - regions: (seen 2026-10-08) - service stages: (seen 2026-10-08) - IAM roles: (seen 2026-10-08) - Audit Logs: (seen 2026-10-08) - status incident feed: (seen 2026-10-08) - us-central1 incident of 19 August 2026: (seen 2026-10-08) - power incident of 7 October 2026: (seen 2026-10-08) - Services Agreement: (seen 2026-10-08) - privacy policy: (seen 2026-10-08) - Data Processing Agreement: (seen 2026-10-08) - sub-processor list: (seen 2026-10-08) - SLA and Compute service level: (seen 2026-10-08) - trust centre: (seen 2026-10-08) - security.txt: (seen 2026-10-08) - CLI release notes: (seen 2026-10-08) - Python SDK on PyPI: (seen 2026-10-08) - JavaScript SDK on npm: (seen 2026-10-08) - MCP server repository: (seen 2026-10-08) - RDAP for nebius.com: (seen 2026-10-08) ## Who's behind it (provenance 83/100, checked 2026-10-08) | Check | Finding | Points | | --- | --- | --- | | Legal entity named | Nebius B.V. | 20/20 | | Domain age | nebius.com, registered 2004-06-26 (22 years) | 15/15 | | Endpoint on the vendor's domain | api.nebius.cloud is not on nebius.com | 0/15 | | Terms of service | read, states 7 of the 7 things a reader expects, and has 1 clause that costs points | 8/10 | | Privacy policy | read, states 8 of the 8 things a reader expects | 10/10 | | Status page | status.nebius.com | 10/10 | | Changelog | published | 10/10 | | security.txt | valid | 10/10 | The Services Agreement (published 15 September 2026, effective 28 September 2026) names Nebius B.V. under Dutch law for customers outside the United States and Israel, with other Nebius entities for those two countries. The separate Terms of Use page covers the website. The privacy policy (23 September 2026) gives Nebius B.V., Burgerweeshuispad 101, 1076ER Amsterdam, and says data processed for customers as a processor falls under the DPA at https://docs.nebius.com/legal/dpa. The API answers at api.nebius.cloud and tokens are exchanged at auth.eu.nebius.com. nebius.cloud is a second domain that Nebius's docs name for the API and the CLI installer. security.txt at nebius.com gives security@nebius.com and expires 2027-12-31. It has no Policy field. The changelog link is the CLI release notes, which are generated from the API. No separate API changelog was found. RDAP for nebius.com gives a registration date of 2004-06-26. ### Terms and privacy, as read A reading by a fixed set of rules, each answered with the vendor's own sentence. Not legal advice. **Terms of service** (https://docs.nebius.com/legal/agreement), read 2026-10-08, dated 2026-09-28, states 7 of the 7 things a reader expects. - To know. Restricts benchmarking or competitive use (costs points). "…or improve a product or service that competes with the Services or engage in competitive analysis or benchmarking, or (d) for any illegal, unlawful, fraudulent, unfair, deceptive or other prohibited purposes (including, but not limited to, providing an illegal service, terrorism, illegal hate speech, child pornography…" - To know. Says access can be ended without notice or for any reason. "Nebius may terminate the Agreement for cause with the Services being immediately disabled and with no expenses or damages reimbursed without notice if:" - To know. Requires arbitration or waives class actions. "If the arbitrator(s) determine a party to be the prevailing party under circumstances where the prevailing party won on some but not all of the claims and counterclaims, the arbitrator may award the prevailing party an appropriate percentage of the costs and attorneys’ fees reasonably incurred by the prevailing party…" - Gives the date it was last updated. Last updated 2026-09-28. - Names the governing law or courts. The law of the State of Israel. - States a limit on its liability. Capped at the fees paid in the 12 months before the claim. - Says how changes to the terms are announced. Gives ten calendar days of notice before a change. - Also in the text (2026-10-08). Nebius may use the customer's name, logo and trademark in advertising and marketing, including customer lists and case studies, with no further consent. "The Customer hereby authorizes Nebius to use the Customer’s name, logo, trademark, trade name, and/or the name of the Customer’s software product or website for informational, advertising, and marketing purposes." - Also in the text (2026-10-08). Customer data on the platform is marked and deleted within 72 hours after the agreement ends, unless applicable law sets another storage period. "In case of termination of the Agreement the Customer Data uploaded on the resources of the Platform is marked and deleted along with resources of the Platform used by the Customer within 72 hours after termination of the Agreement unless applicable law stipulates any other storage period." - Also in the text (2026-10-08). A third party authorised to manage the services for the customer must accept the agreement, and the customer answers for all activity under its account. "If the Customer authorizes any third parties to manage the Services on behalf of the Customer, the Customer shall ensure that such third parties accept this Agreement, including the Linked Documents referred to in the Agreement." **Privacy policy** (https://docs.nebius.com/legal/privacy), read 2026-10-08, dated 2026-09-23, states 8 of the 8 things a reader expects. - Gives the date it was last updated. Last updated 2026-09-23. - Says how long data is kept. Names a period of 18 months. - Says whether personal data is sold or shared for advertising. Says it does not sell personal data. - Gives a privacy contact. Names a data protection officer. - Says where data is transferred or stored. Relies on the Data Privacy Framework. ## Live (updated 2026-10-09 10:42 UTC) - Right now: up, HTTP 404, 228 ms, checked 2026-10-09 10:42 UTC (get on `https://api.nebius.cloud`) - Uptime 24h 100.0% (33 probes) · 30 days 100.0% (33 probes) · p50 164 ms · p95 237 ms - Vendor status page: none, All Systems Operational - Always current: https://www.anchorterminal.com/api/v1/live/nebius-ai-cloud.json ## Probe metrics Not measured yet. Our benchmark probes haven't run, so there's no availability, latency or error rate from a run and Performance is pending. Live uptime, where we poll the endpoint, is under Live and doesn't change the score. ## Prices | Item | Price | Unit | Note | | --- | --- | --- | --- | | NVIDIA H200 NVLink, on demand | $5.40 | per GPU-hour | Billed per second. $4.50 before 1 October 2026 | | NVIDIA H100 NVLink, on demand | $4.50 | per GPU-hour | $3.85 before 1 October 2026 | | NVIDIA B200 NVLink, on demand | $8.50 | per GPU-hour | | | NVIDIA B300 NVLink, on demand | $9.50 | per GPU-hour | | | NVIDIA RTX PRO 6000, on demand | $1.80 | per GPU-hour | | | NVIDIA L40S, GPU only | $1.35 | per GPU-hour | vCPU ($0.01 to $0.012 an hour) and RAM ($0.0032 a GiB-hour) are billed separately | | NVIDIA H200 NVLink, preemptible | $0.79 | per GPU-hour | Minimum spot price from 8 October 2026. The spot price can change every 15 minutes | | Network SSD disk | $0.071 | per GB per month | Per GiB for 730 hours | Across all listings: https://www.anchorterminal.com/prices/index.md ## Strengths - OpenAPI 3.0.3 document at `https://api.nebius.cloud/openapi.json` with 602 operations, generated from the same protobuf definitions as the gRPC API, CLI, Terraform provider and SDKs - `X-Idempotency-Key` header for modifying calls, and a `retry_type` field on errors that says whether to retry the call - Service accounts sign in with an uploaded RSA key and receive 12-hour tokens, with roles granted per tenant, project or resource - Per-second billing with public prices, and a 99.5 per cent monthly uptime commitment per virtual machine - llms.txt, every docs page as Markdown, and a keyless docs MCP server at `https://docs.nebius.com/mcp` ## Weaknesses - Status page lists 14 incidents marked major between 14 July and 8 October 2026, including about 21 hours of partial degradation in us-central1 on 19 August - No request rate limits with numbers and no Retry-After guidance found in the reviewed documentation - Serverless AI endpoints run on one container VM that is started and stopped by hand; no autoscaling or scale to zero found - No free tier or trial found. Adding a card at signup charges $25 to the balance, and signup is a browser flow through Google, GitHub or Microsoft - The OpenAPI document has no examples, documents only 200 responses and reports its version as `version not set` ## Before you call it (notes for agents) 1. Use a service account with an authorised key, then exchange a five-minute RS256 JWT at `https://auth.eu.nebius.com/oauth2/token/exchange` for a 12-hour Bearer token 2. Send `X-Idempotency-Key` with a random UUID on every create, update and delete, since a 504 can follow a call that succeeded 3. Poll the returned operation (`/ai/v1/endpoints/operations/{id}`) until `status` is set; concurrent operations on one resource are not supported 4. Stop or delete endpoints when idle. A stopped endpoint bills nothing, a stopped Devlab or VM still bills for its disk 5. Check region support first. Serverless AI is absent from `eu-south1` and `us-north1`, and each GPU platform exists in one to four regions 6. Run the beta `nebius/mcp-server` with safe mode on (the default); `nebius_cli_execute` can run any CLI command when `SAFE_MODE=false` ## Connect Install: ```bash curl -sSL https://artifacts.nebius.cloud/cli/install.sh | bash ``` First request: ```bash curl --request GET --url 'https://api.nebius.cloud/iam/v1/profiles' --header 'Authorization: Bearer ' ``` MCP client configuration: ```json { "mcpServers": { "Nebius MCP Server": { "args": [ "--refresh-package", "nebius-mcp-server", "nebius-mcp-server@git+https://github.com/nebius/mcp-server@main" ], "command": "uvx", "env": {} } } } ``` Through letme (picks today, calling later): https://letme.dev/nebius-ai-cloud (letme picks it for compute.batch, the top-graded tool for the job, letme picks it for compute.containers, the top-graded tool for the job, letme picks it for compute.endpoints, the top-graded tool for the job, letme picks it for compute.gpu, the top-graded tool for the job). letme answers with the pick and how to call it direct; calling through letme (one key, the vendor's own price) comes later. How it works: https://www.anchorterminal.com/letme/index.md ## Similar tools Ranked by shared capabilities, then score. Same-category tools with no shared capability key are listed last. | Tool | Grade | Score | Rank | Shared capabilities | x402 | Markdown | | --- | --- | --- | --- | --- | --- | --- | | Modal | B | 63.6 | 346 | compute.gpu, compute.endpoints, compute.batch, compute.containers | no | https://www.anchorterminal.com/tools/modal.md | | Verda | B | 62.3 | 396 | compute.gpu, compute.endpoints, compute.batch, compute.containers | no | https://www.anchorterminal.com/tools/verda.md | | CoreWeave | C | 61.5 | 421 | compute.gpu, compute.containers, compute.endpoints, compute.batch | no | https://www.anchorterminal.com/tools/coreweave.md | | Northflank | C | 61.3 | 426 | compute.gpu, compute.containers, compute.batch, compute.endpoints | no | https://www.anchorterminal.com/tools/northflank.md | | Beam | C | 55.5 | 588 | compute.gpu, compute.endpoints, compute.batch, compute.containers | no | https://www.anchorterminal.com/tools/beam.md | | Cerebrium | C | 55.3 | 593 | compute.gpu, compute.endpoints, compute.batch, compute.containers | no | https://www.anchorterminal.com/tools/cerebrium.md | ## Panel reviews (0) Reviewed by the Anchor panel (https://www.anchorterminal.com/reviewers/index.md): . Desk reviews, written from public documentation, pricing, terms, source and status history on 1 October 2026. No calls made. For a desk review, the outcome says whether the reviewer's questions could be answered from public material: success, partial or failure. How reviews work: https://www.anchorterminal.com/reviews/how-it-works.md ## Notable - The REST API at api.nebius.cloud is generated from the same protobuf definitions as the gRPC API, the CLI, the Terraform provider and the Go, Python and JavaScript SDKs (source: ) - Serverless AI runs containers as Devlabs, jobs and endpoints on managed container VMs with per-second billing at Compute prices. Endpoints are started and stopped through `/ai/v1/endpoints/start` and `/ai/v1/endpoints/stop`; no autoscaling was found (source: ) - On-demand prices rose on 1 October 2026 (H200 from $4.50 to $5.40 a GPU-hour, H100 from $3.85 to $4.50), and from 8 October preemptible prices for B300, B200, H200, H100 and RTX PRO 6000 are dynamic spot prices that can change every 15 minutes (source: ) - The Services Agreement bars using the services to "engage in competitive analysis or benchmarking", and bars API requests above the specified limits and vulnerability probing without authorisation (source: ) - docs.nebius.com/llms.txt and nebius.com/llms.txt carry instructions addressed to AI agents (fetch the `.md` form, discover CLI commands with `--help`, read the product index first). Recorded as a fact (source: ) - The official MCP server is beta, runs locally over stdio and wraps the CLI with four tools; its README warns that `nebius_cli_execute` can run any CLI command when safe mode is off (source: ) - status.nebius.com lists 22 incidents between 14 July and 8 October 2026, 14 marked major (source: ) ## Compare - [Baseten vs Nebius AI Cloud](https://www.anchorterminal.com/compare/baseten-vs-nebius-ai-cloud.md): B 66.5 vs B 67.2 - [Beam vs Nebius AI Cloud](https://www.anchorterminal.com/compare/beam-vs-nebius-ai-cloud.md): C 55.5 vs B 67.2 - [Cerebrium vs Nebius AI Cloud](https://www.anchorterminal.com/compare/cerebrium-vs-nebius-ai-cloud.md): C 55.3 vs B 67.2 - [CoreWeave vs Nebius AI Cloud](https://www.anchorterminal.com/compare/coreweave-vs-nebius-ai-cloud.md): C 61.5 vs B 67.2 - [Hugging Face Inference Endpoints vs Nebius AI Cloud](https://www.anchorterminal.com/compare/hugging-face-inference-endpoints-vs-nebius-ai-cloud.md): B 64.5 vs B 67.2 - [Hyperbolic vs Nebius AI Cloud](https://www.anchorterminal.com/compare/hyperbolic-vs-nebius-ai-cloud.md): D 48 vs B 67.2 - [Koyeb vs Nebius AI Cloud](https://www.anchorterminal.com/compare/koyeb-vs-nebius-ai-cloud.md): D 46.5 vs B 67.2 - [Lambda Cloud vs Nebius AI Cloud](https://www.anchorterminal.com/compare/lambda-vs-nebius-ai-cloud.md): D 50 vs B 67.2 - [Modal vs Nebius AI Cloud](https://www.anchorterminal.com/compare/modal-vs-nebius-ai-cloud.md): B 63.6 vs B 67.2 - [Nebius AI Cloud vs Northflank](https://www.anchorterminal.com/compare/nebius-ai-cloud-vs-northflank.md): B 67.2 vs C 61.3 - [Nebius AI Cloud vs Replicate Deployments](https://www.anchorterminal.com/compare/nebius-ai-cloud-vs-replicate-deploy.md): B 67.2 vs B 63.6 - [Nebius AI Cloud vs Runpod](https://www.anchorterminal.com/compare/nebius-ai-cloud-vs-runpod.md): B 67.2 vs D 53.5 - [Nebius AI Cloud vs Thunder Compute](https://www.anchorterminal.com/compare/nebius-ai-cloud-vs-thunder-compute.md): B 67.2 vs C 56.1 - [Nebius AI Cloud vs Vast.ai](https://www.anchorterminal.com/compare/nebius-ai-cloud-vs-vast-ai.md): B 67.2 vs B 62.6 - [Nebius AI Cloud vs Verda](https://www.anchorterminal.com/compare/nebius-ai-cloud-vs-verda.md): B 67.2 vs B 62.3 ## Verify this listing For the vendor. The badge or a plain link to this page verifies the listing, from a page on nebius.com or one of its subdomains, or the README of github.com/nebius/api. It shows the listing is the vendor's and that the vendor knows it's here, and it never changes a grade, rank or review. The vendor sends the page's address to `POST https://www.anchorterminal.com/api/v1/verify` as `{"slug": "nebius-ai-cloud", "url": "…"}`, or calls the `verify_listing` tool at https://www.anchorterminal.com/mcp. We fetch the page once, then again every week; two failed checks in a row and the verification lapses, and a later pass restores it. What we check: https://www.anchorterminal.com/builders/index.md#verify HTML badge: ```html Nebius AI Cloud on Anchor Terminal ``` Markdown badge, for a README: ```markdown [![Nebius AI Cloud on Anchor Terminal](https://www.anchorterminal.com/badges/nebius-ai-cloud.svg)](https://www.anchorterminal.com/tools/nebius-ai-cloud) ``` Plain link: ```html Nebius AI Cloud on Anchor Terminal ``` ## Share this listing For the vendor. Sharing assets for social media, two PNGs of 1200 × 630 that say Nebius AI Cloud is listed on Anchor Terminal, with the vendor's logo and this page's address and no grade or score. - Dark: https://www.anchorterminal.com/assets/share/nebius-ai-cloud-dark.png - Light: https://www.anchorterminal.com/assets/share/nebius-ai-cloud-light.png