{
  "data": {
    "a": {
      "slug": "prism-inference",
      "name": "Prism Inference",
      "vendor": "Prism Technologies Inc",
      "vendorUrl": "https://prisminference.com",
      "kind": "model",
      "category": "inference",
      "summary": "Prism is a hosted inference API from Prism Technologies Inc for open-weight models, aimed at coding agents. It accepts OpenAI Chat Completions, OpenAI Responses and Anthropic Messages requests at api.prisminference.com. It launched on 24 September 2026.",
      "url": "https://www.anchorterminal.com/tools/prism-inference",
      "markdownUrl": "https://www.anchorterminal.com/tools/prism-inference.md",
      "slimMarkdownUrl": "https://www.anchorterminal.com/tools/prism-inference.min.md",
      "jsonUrl": "https://www.anchorterminal.com/api/v1/tools/prism-inference.json",
      "repo": "https://github.com/prismhq/hermes-prism-provider",
      "license": "Proprietary service under Prism's terms of service. The OpenAPI file declares `LicenseRef-Proprietary`. The Hermes provider plugin repository carries no licence file",
      "transports": [
        "http"
      ],
      "remoteUrl": "https://api.prisminference.com/v1",
      "packages": [],
      "auth": "api-key",
      "authNotes": "`Authorization: Bearer` with a key issued from account settings after sign-up at prisminference.com/signup. The inference endpoints also accept the key in `x-api-key`. A missing, invalid, expired or revoked key returns 401. No key scopes were found in the reviewed documentation. An agent can call `POST https://prisminference.com/api/agent-signups` with the owner's email and a username and receive a key once, but that key can't run inference until the owner supplies a six-digit emailed code, and it expires after 30 days. `GET /v1/models` needs no key.",
      "pricing": "usage",
      "pricingNotes": "Prepaid per-token pricing with no minimum. DeepSeek-V4.1-Flash is $0.09 in, $1.20 out and $0.06 cache read per million tokens, and Gemma 4 31B is $0.30, $0.40 and $0.15 (https://prisminference.com/pricing, matched by https://api.prisminference.com/v1/models). No free tier or trial credit was found, and a workspace without credit gets 402. A person funds the workspace in a browser. Elastic endpoints, dedicated deployments and batch are sold through sales with no published price.",
      "priceSummary": "from $0.09 / 1M in",
      "where": "hosted",
      "x402": {
        "level": "no",
        "evidence": "No x402, MPP or L402 in llms.txt, the docs index, the OpenAPI file or the pricing page, read 2026-10-08. The docs describe prepaid credit funded by a person in the browser.",
        "endpoints": []
      },
      "toolCount": null,
      "popularity": {
        "githubStars": null,
        "npmWeekly": null,
        "pypiWeekly": null,
        "asOf": "2026-10-08"
      },
      "docsUrl": "https://docs.prisminference.com",
      "rateLimitsUrl": "https://docs.prisminference.com/rate-limits",
      "llmsTxt": "https://prisminference.com/llms.txt",
      "openapi": "https://docs.prisminference.com/openapi.yaml",
      "capabilities": [
        "inference.fast",
        "inference.open-weights",
        "inference.llm"
      ],
      "tags": [
        "hosted",
        "model",
        "open-weights",
        "fast",
        "usage-priced",
        "prepaid",
        "openapi",
        "llms-txt",
        "openai-compatible",
        "anthropic-compatible",
        "zero-retention",
        "status-page",
        "new"
      ],
      "lastRelease": "2026-10-06",
      "graded": true,
      "anchor": {
        "graded": true,
        "score": 60.1,
        "grade": "C",
        "agentReady": false,
        "rank": 479,
        "ranked": true,
        "rankOf": 842,
        "categoryRank": 10,
        "methodology": "0.4",
        "run": "2026-10-01",
        "scores": {
          "ergonomics": 68,
          "maintenance": 49,
          "payments": 30,
          "reliability": 65,
          "schema": 82,
          "security": 65,
          "transparency": 61
        },
        "pending": [
          "performance",
          "tasks"
        ],
        "assessment": {
          "confidence": "medium",
          "date": "2026-10-08"
        },
        "negative": -2,
        "negativeNotes": [
          "2026-10-08. The home page shows a '99.99% Uptime SLA' tile, and the pricing page says 'No minimums, no rate limits' above the per-token table. The terms of 9 September 2026 say the services have no guaranteed uptime or service credit unless a separate written agreement says otherwise, and the docs describe per-key rate limits that return 429. No SLA document was found. A misleading claim, with the smallest deduction because the terms and docs state the real position (https://prisminference.com/, https://prisminference.com/pricing, https://prisminference.com/terms, https://docs.prisminference.com/rate-limits)."
        ],
        "verdict": "Three wire formats, a public OpenAPI 3.1 file, per-token prices in a keyless catalogue and zero data retention by default on every tier. The service launched on 24 September 2026 with two models, one of them by request, from a two-person company. No rate-limit numbers, SLA document, free tier or deprecation policy was found.",
        "bestFor": "Coding agents that want DeepSeek-V4.1-Flash at a low input price, with no retention, through whichever of the three wire formats the harness already speaks.",
        "strengths": [
          "One key works across OpenAI Chat Completions, OpenAI Responses and Anthropic Messages, with a public OpenAPI 3.1 file, llms.txt and Markdown docs",
          "Zero data retention is the default on every tier, and the privacy policy, terms and docs all say inputs and outputs are never used for training",
          "Every error carries a stable `code`, a `retryable` flag, a `fix` hint and a `docs_url`, and 429 carries `Retry-After` in seconds",
          "`GET /v1/models` answers without a key and returns context length, maximum output and per-token prices for each model",
          "DeepSeek-V4.1-Flash is listed with a 1M-token context and 384,000 output tokens at $0.09 in and $1.20 out per million"
        ],
        "weaknesses": [
          "Two models. The docs mark Gemma 4 31B as request access per organisation, while llms.txt and the keyless catalogue list it as available",
          "No rate-limit numbers are published. The docs say per-key limits exist, and the pricing page says 'no rate limits'",
          "The home page shows a '99.99% Uptime SLA' tile, while the terms say there is no guaranteed uptime without a separate written agreement",
          "No free tier found. Billing is prepaid, a person funds the workspace in a browser, and the agent sign-up key needs an emailed code",
          "The service launched on 24 September 2026. The changelog has one entry, and no deprecation policy, sub-processor list or certification was found"
        ],
        "agentNotes": [
          "Call `GET https://api.prisminference.com/v1/models` at start-up, with no key, and use only ids it returns. Expect 403 on `gemma-4-31b` without organisation access",
          "Use base URL `https://api.prisminference.com/v1` for OpenAI clients and `https://api.prisminference.com` with no `/v1` for Anthropic clients",
          "Read `error.retryable` before retrying, and wait for `Retry-After` on 429, which covers both key limits and model capacity",
          "Send `reasoning_effort: \"none\"` or `low` when latency matters. Reasoning is on by default and its tokens are billed as output",
          "Keep conversation state yourself and send `store: false` on Responses. `previous_response_id`, stored responses and hosted tools aren't supported"
        ],
        "metrics": {
          "kind": "remote",
          "measured": false
        },
        "reviewCount": 0,
        "avgRating": 0,
        "history": [
          {
            "basis": "public evidence",
            "confidence": "medium",
            "grade": "C",
            "methodology": "0.4",
            "pending": [
              "performance",
              "tasks"
            ],
            "run": "2026-10-01",
            "runLabel": "October 2026 research run",
            "score": 60.1
          }
        ],
        "editorialScores": {
          "ergonomics": 68,
          "maintenance": 49,
          "payments": 30,
          "reliability": 65,
          "schema": 82,
          "security": 65,
          "transparency": 43
        },
        "provenanceScore": 79
      },
      "connect": {
        "install": "pip install openai   # or: npm install openai, base URL https://api.prisminference.com/v1. Anthropic SDKs use https://api.prisminference.com with no /v1",
        "http": "curl \"https://api.prisminference.com/v1/chat/completions\" \\\n  -H \"Authorization: Bearer $PRISM_API_KEY\" \\\n  -H \"Content-Type: application/json\" \\\n  -d '{\"model\":\"deepseek-v4.1-flash\",\"messages\":[{\"role\":\"user\",\"content\":\"Return pong.\"}]}'",
        "claudeCode": "export ANTHROPIC_BASE_URL=https://api.prisminference.com\nexport ANTHROPIC_AUTH_TOKEN=$PRISM_API_KEY\nexport ANTHROPIC_MODEL=deepseek-v4.1-flash\nexport ANTHROPIC_SMALL_FAST_MODEL=gemma-4-31b"
      },
      "letme": {
        "capability": "https://letme.dev/inference.fast",
        "tool": "https://letme.dev/prism-inference"
      },
      "area": "models",
      "provenance": {
        "legalEntity": "Prism Technologies Inc",
        "domain": "prisminference.com",
        "domainRegistered": "2026-09-09",
        "endpointOnVendorDomain": true,
        "terms": "https://prisminference.com/terms",
        "privacy": "https://prisminference.com/privacy",
        "statusPage": "https://status.prisminference.com",
        "changelog": "https://docs.prisminference.com/changelog",
        "securityTxt": "valid",
        "checked": "2026-10-08",
        "notes": [
          "The terms and the privacy policy, both last updated 9 September 2026, name Prism Technologies Inc. The terms are governed by California law with arbitration in San Francisco.",
          "RDAP gives a registration date of 2026-09-09 for prisminference.com, with Name.com as registrar.",
          "security.txt has a Contact line (founders@prisminference.com), an Expires date of 2027-10-06 and a Canonical line. No disclosure policy or bug bounty is named.",
          "The status page runs on incident.io with one component for each model and no component for the API or the website. Its incidents feed was empty on 8 October 2026.",
          "The YC directory lists Prism in the Spring 2025 batch, in San Francisco, with a team of 2. The home page links an X account named prism_videos, from the company's earlier video product."
        ],
        "score": 79
      },
      "pageJsonUrl": "https://www.anchorterminal.com/tools/prism-inference.json",
      "live": {
        "slug": "prism-inference",
        "probe": {
          "target": "https://api.prisminference.com/v1",
          "method": "get",
          "lastAt": "2026-10-09T11:46:37.811260774Z",
          "lastOk": true,
          "lastStatus": 200,
          "lastMs": 130,
          "authRequired": false,
          "uptime24h": 100,
          "uptime30d": 100,
          "p50ms24h": 140,
          "p95ms24h": 443,
          "samples24h": 218,
          "samples30d": 218,
          "days": [
            {
              "date": "2026-10-08",
              "probes": 93,
              "ok": 93
            },
            {
              "date": "2026-10-09",
              "probes": 125,
              "ok": 125
            }
          ]
        },
        "vendorStatus": {
          "page": "https://status.prisminference.com",
          "indicator": "none",
          "summary": "All Systems Operational",
          "checkedAt": "2026-10-09T11:40:26.68885523Z"
        },
        "versions": [
          {
            "registry": "github",
            "name": "prismhq/hermes-prism-provider",
            "version": "v1.0.2",
            "released": "2026-09-16",
            "seenAt": "2026-10-08T16:26:20.638377705Z"
          }
        ],
        "githubStars": 0,
        "securityTxt": {
          "url": "https://prisminference.com/.well-known/security.txt",
          "state": "valid",
          "expires": "2027-10-06T00:00:00.000Z",
          "checkedAt": "2026-10-08T15:38:58.310470594Z"
        },
        "pages": [
          {
            "url": "https://docs.prisminference.com/changelog",
            "kind": "changelog",
            "status": 200,
            "checkedAt": "2026-10-08T18:19:15.357311357Z",
            "changedAt": "0001-01-01T00:00:00Z",
            "fingerprint": "184a0fafdb8c"
          },
          {
            "url": "https://prisminference.com/pricing",
            "kind": "pricing",
            "status": 200,
            "checkedAt": "2026-10-08T18:23:20.103073255Z",
            "changedAt": "0001-01-01T00:00:00Z",
            "fingerprint": "dffb5d16a3c8"
          },
          {
            "url": "https://prisminference.com/privacy",
            "kind": "privacy",
            "status": 200,
            "checkedAt": "2026-10-08T18:23:22.266572016Z",
            "changedAt": "0001-01-01T00:00:00Z",
            "fingerprint": "64ea7b8f1b7b"
          },
          {
            "url": "https://prisminference.com/terms",
            "kind": "terms",
            "status": 200,
            "checkedAt": "2026-10-08T18:23:24.296415884Z",
            "changedAt": "0001-01-01T00:00:00Z",
            "fingerprint": "27dd15c69f49"
          }
        ],
        "updatedAt": "2026-10-09T11:46:37.811260774Z"
      }
    },
    "answer": "SambaCloud scores 66.4 (B) on agent readiness against Prism Inference's 60.1 (C), and leads in 4 of 7 scored categories. Prism Inference leads on security \u0026 auth.",
    "b": {
      "slug": "sambanova",
      "name": "SambaCloud",
      "vendor": "SambaNova Systems, Inc.",
      "vendorUrl": "https://sambanova.ai",
      "kind": "model",
      "category": "inference",
      "summary": "SambaCloud is SambaNova's hosted inference API for open-weight models running on its own RDU processors. It answers OpenAI-style chat, completions and Responses calls and Anthropic-style Messages calls, with Python and TypeScript SDKs.",
      "url": "https://www.anchorterminal.com/tools/sambanova",
      "markdownUrl": "https://www.anchorterminal.com/tools/sambanova.md",
      "slimMarkdownUrl": "https://www.anchorterminal.com/tools/sambanova.min.md",
      "jsonUrl": "https://www.anchorterminal.com/api/v1/tools/sambanova.json",
      "repo": "https://github.com/sambanova/sambanova-python",
      "license": "Proprietary service under the SambaCloud Terms of Service. The SDKs and the OpenAPI document are Apache-2.0",
      "transports": [
        "http"
      ],
      "remoteUrl": "https://api.sambanova.ai/v1",
      "packages": [
        {
          "registry": "pypi",
          "name": "sambanova"
        },
        {
          "registry": "npm",
          "name": "sambanova"
        }
      ],
      "auth": "api-key",
      "authNotes": "Self-serve Bearer key from the SambaCloud console at https://cloud.sambanova.ai/apis after a browser sign-up. The Messages routes also take the same key in `x-api-key`. Up to 25 keys per user, shown once, with no scopes.",
      "pricing": "freemium",
      "pricingNotes": "Free tier with no payment method, at 20 requests a minute, 20 requests a day and 200,000 tokens a day per model. Linking a card moves the account to the Developer tier, billed per token through Stripe, from $0.22 in and $0.59 out per 1M tokens on gpt-oss-120b (https://cloud.sambanova.ai/plans/pricing, checked 2026-10-08).",
      "priceSummary": "from $0.22 / 1M in",
      "where": "hosted",
      "x402": {
        "level": "no",
        "evidence": "No x402, MPP or L402 in the docs index, the OpenAPI document or the pricing page (checked 2026-10-08).",
        "endpoints": []
      },
      "toolCount": null,
      "popularity": {
        "githubStars": 2,
        "npmWeekly": 76,
        "pypiWeekly": 6129,
        "asOf": "2026-10-08"
      },
      "docsUrl": "https://docs.sambanova.ai/docs/en/get-started/overview",
      "rateLimitsUrl": "https://docs.sambanova.ai/docs/en/models/rate-limits",
      "llmsTxt": "https://docs.sambanova.ai/docs/llms.txt",
      "openapi": "https://raw.githubusercontent.com/sambanova/sambanova-inference-api-spec/refs/heads/main/openapi.documented.json",
      "capabilities": [
        "inference.llm",
        "inference.open-weights",
        "inference.fast"
      ],
      "tags": [
        "hosted",
        "model",
        "open-weights",
        "free-tier",
        "no-card",
        "openapi",
        "llms-txt",
        "openai-compatible",
        "anthropic-compatible",
        "python",
        "typescript",
        "status-page",
        "soc2"
      ],
      "lastRelease": "2026-09-17",
      "graded": true,
      "anchor": {
        "graded": true,
        "score": 66.4,
        "grade": "B",
        "agentReady": false,
        "rank": 269,
        "ranked": true,
        "rankOf": 842,
        "categoryRank": 7,
        "methodology": "0.4",
        "run": "2026-10-01",
        "scores": {
          "ergonomics": 67,
          "maintenance": 77,
          "payments": 40,
          "reliability": 85,
          "schema": 81,
          "security": 58,
          "transparency": 63
        },
        "pending": [
          "performance",
          "tasks"
        ],
        "assessment": {
          "confidence": "medium",
          "date": "2026-10-08"
        },
        "negative": -2,
        "negativeNotes": [
          "2026-10-08. The models and rate-limit pages list `MiniMax-M2.7` as a production model, and the July 2026 release notes say `Mistral-Large-3-675B-Instruct-2512` was promoted to production. Neither is on the pricing page or in `/v1/models`, and the deprecations page records neither. Two points, since a keyed call was not made to confirm the ids fail (https://docs.sambanova.ai/docs/en/models/sambacloud-models, https://api.sambanova.ai/v1/models)."
        ],
        "verdict": "Per-token prices and the model list are readable without a key at `/v1/models`, and a free tier needs no card. Production models get two to three weeks' notice before removal, 13 model ids left between March and June 2026, and the docs disagree with the live catalogue on MiniMax-M2.7.",
        "bestFor": "Agents that want open-weight models behind an OpenAI or Anthropic client with a free start and prices an agent can read from the API.",
        "strengths": [
          "Public OpenAPI 3.1.1 document for all 10 operations, plus `llms.txt` and a Markdown copy of every docs page",
          "`GET /v1/models` answers without a key and returns context length, output cap and per-token prices for each model",
          "Free tier with no payment method at 20 requests a minute, 20 requests a day and 200,000 tokens a day per model",
          "One key works with the OpenAI client, the Anthropic client (`/v1/messages`, `x-api-key`) and SambaNova's own SDKs",
          "Status page with a component per model and no incident posted since 25 June 2026"
        ],
        "weaknesses": [
          "Production models get a notice of two to three weeks, and 13 model ids were removed between 9 March and 9 June 2026",
          "The docs list `MiniMax-M2.7` as a production model, while the pricing page and `/v1/models` did not list it on 8 October 2026",
          "`strict: true` on a JSON schema is accepted and has no effect, and the July 2026 notes record failing structured output on DeepSeek-V3.2",
          "Keys carry no scopes or project limits, and no key rotation guidance was found in the reviewed documentation",
          "The privacy policy is dated 27 May 2023, and no DPA, retention period or sub-processor list was found on the legal page"
        ],
        "agentNotes": [
          "Call `GET https://api.sambanova.ai/v1/models` at start-up and use only ids it returns. The docs name models the endpoint no longer lists",
          "Read `x-ratelimit-remaining-requests` and `x-ratelimit-remaining-requests-day` on every response. The free tier allows 20 requests a day per model",
          "Treat 429 `queue_full` and 503 `maintenance` as retryable after a delay, and 410 `model_deprecated` as a signal to change model",
          "Validate JSON output yourself. Schema enforcement is best effort and `strict: true` changes nothing",
          "Check `max_completion_tokens` per model. DeepSeek-V3.1 caps output at 7,168 tokens and Llama 3.3 70B at 3,072"
        ],
        "metrics": {
          "kind": "remote",
          "measured": false
        },
        "reviewCount": 0,
        "avgRating": 0,
        "history": [
          {
            "basis": "public evidence",
            "confidence": "medium",
            "grade": "B",
            "methodology": "0.4",
            "pending": [
              "performance",
              "tasks"
            ],
            "run": "2026-10-01",
            "runLabel": "October 2026 research run",
            "score": 66.4
          }
        ],
        "editorialScores": {
          "ergonomics": 67,
          "maintenance": 77,
          "payments": 40,
          "reliability": 85,
          "schema": 81,
          "security": 58,
          "transparency": 44
        },
        "provenanceScore": 81
      },
      "connect": {
        "install": "pip install sambanova   # or: npm install sambanova",
        "http": "curl https://api.sambanova.ai/v1/chat/completions \\\n  -H \"Authorization: Bearer $SAMBANOVA_API_KEY\" -H \"Content-Type: application/json\" \\\n  -d '{\"model\":\"gpt-oss-120b\",\"messages\":[{\"role\":\"user\",\"content\":\"hello\"}]}'"
      },
      "letme": {
        "capability": "https://letme.dev/inference.llm",
        "tool": "https://letme.dev/sambanova"
      },
      "area": "models",
      "provenance": {
        "legalEntity": "SambaNova Systems, Inc.",
        "domain": "sambanova.ai",
        "domainRegistered": "2017-12-18",
        "endpointOnVendorDomain": true,
        "terms": "https://sambanova.ai/cloud-end-user-license-agreement",
        "privacy": "https://sambanova.ai/privacy-policy",
        "statusPage": "https://status.sambanova.ai",
        "changelog": "https://docs.sambanova.ai/docs/en/release-notes/sambacloud",
        "securityTxt": "none",
        "checked": "2026-10-08",
        "notes": [
          "The SambaCloud Terms of Service name SambaNova Systems, Inc., a Delaware corporation at 2460 N. First St, Suite #100, San Jose, CA 95131, and define the Service as the SambaCloud platform. The page carries no date. The legal index says it was last updated on 5 February 2026.",
          "The privacy policy is the company's only one. Its last revision is dated 27 May 2023, it covers the website, communications and related services, and it does not name SambaCloud or API inputs.",
          "No DPA, SLA, acceptable use policy or sub-processor list is linked from https://sambanova.ai/legal-agreements.",
          "https://sambanova.ai/.well-known/security.txt returns 404.",
          "RDAP for sambanova.ai gives a registration date of 2017-12-18.",
          "The API answers at api.sambanova.ai and the console at cloud.sambanova.ai."
        ],
        "score": 81
      },
      "pageJsonUrl": "https://www.anchorterminal.com/tools/sambanova.json",
      "live": {
        "slug": "sambanova",
        "probe": {
          "target": "https://api.sambanova.ai/v1",
          "method": "get",
          "lastAt": "2026-10-09T11:46:39.537400402Z",
          "lastOk": true,
          "lastStatus": 405,
          "lastMs": 408,
          "authRequired": false,
          "uptime24h": 100,
          "uptime30d": 100,
          "p50ms24h": 403,
          "p95ms24h": 1118,
          "samples24h": 44,
          "samples30d": 44,
          "days": [
            {
              "date": "2026-10-09",
              "probes": 44,
              "ok": 44
            }
          ]
        },
        "vendorStatus": {
          "page": "https://status.sambanova.ai",
          "indicator": "none",
          "summary": "All Systems Operational",
          "checkedAt": "2026-10-09T11:40:38.07700378Z"
        },
        "updatedAt": "2026-10-09T11:46:39.537400402Z"
      }
    },
    "facts": [
      {
        "a": "Model API",
        "b": "Model API",
        "name": "Kind"
      },
      {
        "a": "Prism Technologies Inc",
        "b": "SambaNova Systems, Inc.",
        "name": "Vendor"
      },
      {
        "a": "https://api.prisminference.com/v1",
        "b": "https://api.sambanova.ai/v1",
        "name": "Hosted endpoint"
      },
      {
        "a": "HTTP",
        "b": "HTTP",
        "name": "Transports"
      },
      {
        "a": "API key",
        "b": "API key",
        "name": "Auth"
      },
      {
        "a": "Pay per use",
        "b": "Freemium",
        "name": "Pricing"
      },
      {
        "a": "no",
        "b": "no",
        "name": "x402"
      },
      {
        "a": "Proprietary service under Prism's terms of service. The OpenAPI file declares `LicenseRef-Proprietary`. The Hermes provider plugin repository carries no licence file",
        "b": "Proprietary service under the SambaCloud Terms of Service. The SDKs and the OpenAPI document are Apache-2.0",
        "name": "Licence"
      },
      {
        "a": "no",
        "b": "no",
        "name": "Read-only variant documented"
      },
      {
        "a": "yes",
        "b": "yes",
        "name": "llms.txt"
      },
      {
        "a": "2026-10-06",
        "b": "2026-09-17",
        "name": "Last release"
      },
      {
        "a": "2026-09-09",
        "b": "no date given",
        "name": "Terms last updated"
      },
      {
        "a": "2026-09-09",
        "b": "2023-05-27",
        "name": "Privacy policy last updated"
      },
      {
        "a": "not found in the text",
        "b": "not found in the text",
        "name": "Customer content may train models"
      },
      {
        "a": "yes",
        "b": "not found in the text",
        "name": "Terms restrict automated access"
      },
      {
        "a": "not found in the text",
        "b": "yes",
        "name": "Terms restrict benchmarking"
      },
      {
        "a": "not found in the text",
        "b": "not found in the text",
        "name": "Terms or service can change without notice"
      },
      {
        "a": "yes",
        "b": "not found in the text",
        "name": "Arbitration or class-action waiver"
      },
      {
        "a": "none",
        "b": "2 stars, 76 npm/wk, 6.1k PyPI/wk",
        "name": "Popularity"
      }
    ],
    "faq": [
      {
        "answer": "SambaCloud scores 66.4 (B) on agent readiness against Prism Inference's 60.1 (C), and leads in 4 of 7 scored categories. Prism Inference leads on security \u0026 auth.",
        "question": "Which is better for AI agents, Prism Inference or SambaCloud?"
      },
      {
        "answer": "Both need an API key.",
        "question": "Do Prism Inference and SambaCloud need an API key?"
      },
      {
        "answer": "Yes. Prism Inference has a hosted endpoint at https://api.prisminference.com/v1 and SambaCloud at https://api.sambanova.ai/v1.",
        "question": "Can an agent call Prism Inference and SambaCloud without installing anything?"
      }
    ],
    "goodFor": [
      {
        "aheadOn": [
          "Security \u0026 auth, 65 against 58"
        ],
        "also": null,
        "goodFor": "Coding agents that want DeepSeek-V4.1-Flash at a low input price, with no retention, through whichever of the three wire formats the harness already speaks.",
        "slug": "prism-inference",
        "watchFor": "Two models. The docs mark Gemma 4 31B as request access per organisation, while llms.txt and the keyless catalogue list it as available"
      },
      {
        "aheadOn": [
          "Reliability, 85 against 65",
          "Payments \u0026 pricing, 40 against 30",
          "Maintenance \u0026 community, 77 against 49"
        ],
        "also": [
          "Free to start without a card"
        ],
        "goodFor": "Agents that want open-weight models behind an OpenAI or Anthropic client with a free start and prices an agent can read from the API.",
        "slug": "sambanova",
        "watchFor": "Production models get a notice of two to three weeks, and 13 model ids were removed between 9 March and 9 June 2026"
      }
    ],
    "job": {
      "capability": "inference.fast",
      "name": "Inference fast"
    },
    "others": [
      {
        "json": "https://www.anchorterminal.com/compare/anthropic-api-vs-prism-inference.json",
        "title": "Claude API vs Prism Inference",
        "url": "https://www.anchorterminal.com/compare/anthropic-api-vs-prism-inference"
      },
      {
        "json": "https://www.anchorterminal.com/compare/anthropic-api-vs-sambanova.json",
        "title": "Claude API vs SambaCloud",
        "url": "https://www.anchorterminal.com/compare/anthropic-api-vs-sambanova"
      },
      {
        "json": "https://www.anchorterminal.com/compare/antseed-vs-prism-inference.json",
        "title": "Antseed vs Prism Inference",
        "url": "https://www.anchorterminal.com/compare/antseed-vs-prism-inference"
      },
      {
        "json": "https://www.anchorterminal.com/compare/antseed-vs-sambanova.json",
        "title": "Antseed vs SambaCloud",
        "url": "https://www.anchorterminal.com/compare/antseed-vs-sambanova"
      },
      {
        "json": "https://www.anchorterminal.com/compare/blockrun-ai-vs-prism-inference.json",
        "title": "BlockRun.AI vs Prism Inference",
        "url": "https://www.anchorterminal.com/compare/blockrun-ai-vs-prism-inference"
      },
      {
        "json": "https://www.anchorterminal.com/compare/blockrun-ai-vs-sambanova.json",
        "title": "BlockRun.AI vs SambaCloud",
        "url": "https://www.anchorterminal.com/compare/blockrun-ai-vs-sambanova"
      },
      {
        "json": "https://www.anchorterminal.com/compare/deepinfra-vs-prism-inference.json",
        "title": "DeepInfra vs Prism Inference",
        "url": "https://www.anchorterminal.com/compare/deepinfra-vs-prism-inference"
      },
      {
        "json": "https://www.anchorterminal.com/compare/deepinfra-vs-sambanova.json",
        "title": "DeepInfra vs SambaCloud",
        "url": "https://www.anchorterminal.com/compare/deepinfra-vs-sambanova"
      },
      {
        "json": "https://www.anchorterminal.com/compare/deepseek-api-vs-prism-inference.json",
        "title": "DeepSeek API vs Prism Inference",
        "url": "https://www.anchorterminal.com/compare/deepseek-api-vs-prism-inference"
      },
      {
        "json": "https://www.anchorterminal.com/compare/deepseek-api-vs-sambanova.json",
        "title": "DeepSeek API vs SambaCloud",
        "url": "https://www.anchorterminal.com/compare/deepseek-api-vs-sambanova"
      },
      {
        "json": "https://www.anchorterminal.com/compare/gemini-api-vs-prism-inference.json",
        "title": "Gemini Developer API vs Prism Inference",
        "url": "https://www.anchorterminal.com/compare/gemini-api-vs-prism-inference"
      },
      {
        "json": "https://www.anchorterminal.com/compare/gemini-api-vs-sambanova.json",
        "title": "Gemini Developer API vs SambaCloud",
        "url": "https://www.anchorterminal.com/compare/gemini-api-vs-sambanova"
      },
      {
        "json": "https://www.anchorterminal.com/compare/groq-vs-prism-inference.json",
        "title": "GroqCloud vs Prism Inference",
        "url": "https://www.anchorterminal.com/compare/groq-vs-prism-inference"
      },
      {
        "json": "https://www.anchorterminal.com/compare/groq-vs-sambanova.json",
        "title": "GroqCloud vs SambaCloud",
        "url": "https://www.anchorterminal.com/compare/groq-vs-sambanova"
      },
      {
        "json": "https://www.anchorterminal.com/compare/mistral-api-vs-prism-inference.json",
        "title": "Mistral AI API vs Prism Inference",
        "url": "https://www.anchorterminal.com/compare/mistral-api-vs-prism-inference"
      },
      {
        "json": "https://www.anchorterminal.com/compare/mistral-api-vs-sambanova.json",
        "title": "Mistral AI API vs SambaCloud",
        "url": "https://www.anchorterminal.com/compare/mistral-api-vs-sambanova"
      },
      {
        "json": "https://www.anchorterminal.com/compare/openai-api-vs-prism-inference.json",
        "title": "OpenAI API vs Prism Inference",
        "url": "https://www.anchorterminal.com/compare/openai-api-vs-prism-inference"
      },
      {
        "json": "https://www.anchorterminal.com/compare/openai-api-vs-sambanova.json",
        "title": "OpenAI API vs SambaCloud",
        "url": "https://www.anchorterminal.com/compare/openai-api-vs-sambanova"
      },
      {
        "json": "https://www.anchorterminal.com/compare/openrouter-vs-prism-inference.json",
        "title": "OpenRouter vs Prism Inference",
        "url": "https://www.anchorterminal.com/compare/openrouter-vs-prism-inference"
      },
      {
        "json": "https://www.anchorterminal.com/compare/openrouter-vs-sambanova.json",
        "title": "OpenRouter vs SambaCloud",
        "url": "https://www.anchorterminal.com/compare/openrouter-vs-sambanova"
      }
    ],
    "scores": [
      {
        "by": 20,
        "edge": "sambanova",
        "key": "reliability",
        "name": "Reliability",
        "prism-inference": 65,
        "sambanova": 85,
        "weight": 16
      },
      {
        "key": "performance",
        "name": "Performance",
        "pending": true,
        "weight": 10
      },
      {
        "by": 1,
        "edge": "prism-inference",
        "key": "schema",
        "name": "Schema \u0026 documentation",
        "prism-inference": 82,
        "sambanova": 81,
        "weight": 13
      },
      {
        "by": 1,
        "edge": "prism-inference",
        "key": "ergonomics",
        "name": "Agent ergonomics",
        "prism-inference": 68,
        "sambanova": 67,
        "weight": 13
      },
      {
        "by": 7,
        "edge": "prism-inference",
        "key": "security",
        "name": "Security \u0026 auth",
        "prism-inference": 65,
        "sambanova": 58,
        "weight": 14
      },
      {
        "by": 10,
        "edge": "sambanova",
        "key": "payments",
        "name": "Payments \u0026 pricing",
        "prism-inference": 30,
        "sambanova": 40,
        "weight": 10
      },
      {
        "key": "tasks",
        "name": "Task success",
        "pending": true,
        "weight": 10
      },
      {
        "by": 28,
        "edge": "sambanova",
        "key": "maintenance",
        "name": "Maintenance \u0026 community",
        "prism-inference": 49,
        "sambanova": 77,
        "weight": 7
      },
      {
        "by": 2,
        "edge": "sambanova",
        "key": "transparency",
        "name": "Transparency \u0026 trust",
        "prism-inference": 61,
        "sambanova": 63,
        "weight": 7
      }
    ],
    "summary": "SambaCloud scores 66.4 (B) on agent readiness against Prism Inference's 60.1 (C), and leads in 4 of 7 scored categories. Prism Inference leads on security \u0026 auth. Both do inference fast.",
    "verdicts": {
      "prism-inference": "Three wire formats, a public OpenAPI 3.1 file, per-token prices in a keyless catalogue and zero data retention by default on every tier. The service launched on 24 September 2026 with two models, one of them by request, from a two-person company. No rate-limit numbers, SLA document, free tier or deprecation policy was found.",
      "sambanova": "Per-token prices and the model list are readable without a key at `/v1/models`, and a free tier needs no card. Production models get two to three weeks' notice before removal, 13 model ids left between March and June 2026, and the docs disagree with the live catalogue on MiniMax-M2.7."
    }
  },
  "kind": "anchor.page",
  "links": {
    "api": "https://www.anchorterminal.com/api/v1/index.json",
    "html": "https://www.anchorterminal.com/compare/prism-inference-vs-sambanova",
    "json": "https://www.anchorterminal.com/compare/prism-inference-vs-sambanova.json",
    "llms": "https://www.anchorterminal.com/llms.txt",
    "markdown": "https://www.anchorterminal.com/compare/prism-inference-vs-sambanova.md",
    "slim": "https://www.anchorterminal.com/compare/prism-inference-vs-sambanova.min.md"
  },
  "markdown": "SambaCloud scores 66.4 (B) on agent readiness against Prism Inference's 60.1 (C), and leads in 4 of 7 scored categories. Prism Inference leads on security \u0026 auth. Both do inference fast.\n\n- Prism Inference: grade C, 60.1/100, rank #479 of 842. Markdown https://www.anchorterminal.com/tools/prism-inference.md · JSON https://www.anchorterminal.com/api/v1/tools/prism-inference.json\n- SambaCloud: grade B, 66.4/100, rank #269 of 842. Markdown https://www.anchorterminal.com/tools/sambanova.md · JSON https://www.anchorterminal.com/api/v1/tools/sambanova.json\n\n## Which one, for what\n\n### Prism Inference (C)\n\nGood for: Coding agents that want DeepSeek-V4.1-Flash at a low input price, with no retention, through whichever of the three wire formats the harness already speaks.\n\nAhead on:\n- Security \u0026 auth, 65 against 58\n\nWatch for: Two models. The docs mark Gemma 4 31B as request access per organisation, while llms.txt and the keyless catalogue list it as available\n\n### SambaCloud (B)\n\nGood for: Agents that want open-weight models behind an OpenAI or Anthropic client with a free start and prices an agent can read from the API.\n\nAhead on:\n- Reliability, 85 against 65\n- Payments \u0026 pricing, 40 against 30\n- Maintenance \u0026 community, 77 against 49\n\nAlso in its favour:\n- Free to start without a card\n\nWatch for: Production models get a notice of two to three weeks, and 13 model ids were removed between 9 March and 9 June 2026\n\n\n## Score by category\n\n| Category | Weight | Prism Inference | SambaCloud | Edge |\n| --- | --- | --- | --- | --- |\n| Reliability | 16% (20 this run) | 65 | 85 | SambaCloud +20 |\n| Performance | 10%, pending | pending | pending | not scored in this run |\n| Schema \u0026 documentation | 13% (16.2 this run) | 82 | 81 | Prism Inference +1 |\n| Agent ergonomics | 13% (16.2 this run) | 68 | 67 | Prism Inference +1 |\n| Security \u0026 auth | 14% (17.5 this run) | 65 | 58 | Prism Inference +7 |\n| Payments \u0026 pricing | 10% (12.5 this run) | 30 | 40 | SambaCloud +10 |\n| Task success | 10%, pending | pending | pending | not scored in this run |\n| Maintenance \u0026 community | 7% (8.8 this run) | 49 | 77 | SambaCloud +28 |\n| Transparency \u0026 trust | 7% (8.8 this run) | 61 | 63 | SambaCloud +2 |\n| Negative events | ≤15 | -2 | -2 | |\n| **Total** | | **60.1 · C** | **66.4 · B** | |\n\n## Facts side by side\n\n| Fact | Prism Inference | SambaCloud |\n| --- | --- | --- |\n| Kind | Model API | Model API |\n| Vendor | Prism Technologies Inc | SambaNova Systems, Inc. |\n| Hosted endpoint | `https://api.prisminference.com/v1` | `https://api.sambanova.ai/v1` |\n| Transports | HTTP | HTTP |\n| Auth | API key | API key |\n| Pricing | Pay per use | Freemium |\n| x402 | no | no |\n| Licence | Proprietary service under Prism's terms of service. The OpenAPI file declares `LicenseRef-Proprietary`. The Hermes provider plugin repository carries no licence file | Proprietary service under the SambaCloud Terms of Service. The SDKs and the OpenAPI document are Apache-2.0 |\n| Read-only variant documented | no | no |\n| llms.txt | yes | yes |\n| Last release | 2026-10-06 | 2026-09-17 |\n| Terms last updated | 2026-09-09 | no date given |\n| Privacy policy last updated | 2026-09-09 | 2023-05-27 |\n| Customer content may train models | not found in the text | not found in the text |\n| Terms restrict automated access | yes | not found in the text |\n| Terms restrict benchmarking | not found in the text | yes |\n| Terms or service can change without notice | not found in the text | not found in the text |\n| Arbitration or class-action waiver | yes | not found in the text |\n| Popularity | none | 2 stars, 76 npm/wk, 6.1k PyPI/wk |\n\n## Verdicts\n\n**Prism Inference.** Three wire formats, a public OpenAPI 3.1 file, per-token prices in a keyless catalogue and zero data retention by default on every tier. The service launched on 24 September 2026 with two models, one of them by request, from a two-person company. No rate-limit numbers, SLA document, free tier or deprecation policy was found.\n\n**SambaCloud.** Per-token prices and the model list are readable without a key at `/v1/models`, and a free tier needs no card. Production models get two to three weeks' notice before removal, 13 model ids left between March and June 2026, and the docs disagree with the live catalogue on MiniMax-M2.7.\n\n## Before you call either\n\n### Prism Inference\n\n1. Call `GET https://api.prisminference.com/v1/models` at start-up, with no key, and use only ids it returns. Expect 403 on `gemma-4-31b` without organisation access\n2. Use base URL `https://api.prisminference.com/v1` for OpenAI clients and `https://api.prisminference.com` with no `/v1` for Anthropic clients\n3. Read `error.retryable` before retrying, and wait for `Retry-After` on 429, which covers both key limits and model capacity\n4. Send `reasoning_effort: \"none\"` or `low` when latency matters. Reasoning is on by default and its tokens are billed as output\n5. Keep conversation state yourself and send `store: false` on Responses. `previous_response_id`, stored responses and hosted tools aren't supported\n\n### SambaCloud\n\n1. Call `GET https://api.sambanova.ai/v1/models` at start-up and use only ids it returns. The docs name models the endpoint no longer lists\n2. Read `x-ratelimit-remaining-requests` and `x-ratelimit-remaining-requests-day` on every response. The free tier allows 20 requests a day per model\n3. Treat 429 `queue_full` and 503 `maintenance` as retryable after a delay, and 410 `model_deprecated` as a signal to change model\n4. Validate JSON output yourself. Schema enforcement is best effort and `strict: true` changes nothing\n5. Check `max_completion_tokens` per model. DeepSeek-V3.1 caps output at 7,168 tokens and Llama 3.3 70B at 3,072\n\n## Questions\n\n### Which is better for AI agents, Prism Inference or SambaCloud?\n\nSambaCloud scores 66.4 (B) on agent readiness against Prism Inference's 60.1 (C), and leads in 4 of 7 scored categories. Prism Inference leads on security \u0026 auth.\n\n### Do Prism Inference and SambaCloud need an API key?\n\nBoth need an API key.\n\n### Can an agent call Prism Inference and SambaCloud without installing anything?\n\nYes. Prism Inference has a hosted endpoint at https://api.prisminference.com/v1 and SambaCloud at https://api.sambanova.ai/v1.\n\n\n## For agents\n\n- This comparison as JSON: https://www.anchorterminal.com/compare/prism-inference-vs-sambanova.json, and with the fewest tokens: https://www.anchorterminal.com/compare/prism-inference-vs-sambanova.min.md\n- Over MCP at https://www.anchorterminal.com/mcp (no key): `compare_tools {\"a\": \"prism-inference\", \"b\": \"sambanova\"}`. From a terminal: `anchor compare prism-inference sambanova`\n- Each listing in full: https://www.anchorterminal.com/api/v1/tools/prism-inference.json and https://www.anchorterminal.com/api/v1/tools/sambanova.json\n\n## Other comparisons with Prism Inference or SambaCloud\n\n- [Claude API vs Prism Inference](https://www.anchorterminal.com/compare/anthropic-api-vs-prism-inference.md)\n- [Claude API vs SambaCloud](https://www.anchorterminal.com/compare/anthropic-api-vs-sambanova.md)\n- [Antseed vs Prism Inference](https://www.anchorterminal.com/compare/antseed-vs-prism-inference.md)\n- [Antseed vs SambaCloud](https://www.anchorterminal.com/compare/antseed-vs-sambanova.md)\n- [BlockRun.AI vs Prism Inference](https://www.anchorterminal.com/compare/blockrun-ai-vs-prism-inference.md)\n- [BlockRun.AI vs SambaCloud](https://www.anchorterminal.com/compare/blockrun-ai-vs-sambanova.md)\n- [DeepInfra vs Prism Inference](https://www.anchorterminal.com/compare/deepinfra-vs-prism-inference.md)\n- [DeepInfra vs SambaCloud](https://www.anchorterminal.com/compare/deepinfra-vs-sambanova.md)\n- [DeepSeek API vs Prism Inference](https://www.anchorterminal.com/compare/deepseek-api-vs-prism-inference.md)\n- [DeepSeek API vs SambaCloud](https://www.anchorterminal.com/compare/deepseek-api-vs-sambanova.md)\n- [Gemini Developer API vs Prism Inference](https://www.anchorterminal.com/compare/gemini-api-vs-prism-inference.md)\n- [Gemini Developer API vs SambaCloud](https://www.anchorterminal.com/compare/gemini-api-vs-sambanova.md)\n- [GroqCloud vs Prism Inference](https://www.anchorterminal.com/compare/groq-vs-prism-inference.md)\n- [GroqCloud vs SambaCloud](https://www.anchorterminal.com/compare/groq-vs-sambanova.md)\n- [Mistral AI API vs Prism Inference](https://www.anchorterminal.com/compare/mistral-api-vs-prism-inference.md)\n- [Mistral AI API vs SambaCloud](https://www.anchorterminal.com/compare/mistral-api-vs-sambanova.md)\n- [OpenAI API vs Prism Inference](https://www.anchorterminal.com/compare/openai-api-vs-prism-inference.md)\n- [OpenAI API vs SambaCloud](https://www.anchorterminal.com/compare/openai-api-vs-sambanova.md)\n- [OpenRouter vs Prism Inference](https://www.anchorterminal.com/compare/openrouter-vs-prism-inference.md)\n- [OpenRouter vs SambaCloud](https://www.anchorterminal.com/compare/openrouter-vs-sambanova.md)\n",
  "meta": {
    "attribution": "Anchor Terminal (https://www.anchorterminal.com)",
    "docs": "https://www.anchorterminal.com/docs/",
    "generatedAt": "2026-10-09",
    "license": "CC-BY-4.0",
    "method": "https://www.anchorterminal.com/benchmark/",
    "methodology": "0.4",
    "openapi": "https://www.anchorterminal.com/openapi.json",
    "preview": false,
    "run": "2026-10-01",
    "runLabel": "October 2026 research run"
  },
  "page": {
    "breadcrumbs": [
      {
        "name": "Home",
        "url": "https://www.anchorterminal.com/"
      },
      {
        "name": "Compare",
        "url": "https://www.anchorterminal.com/compare/"
      },
      {
        "name": "Prism Inference vs SambaCloud",
        "url": ""
      }
    ],
    "description": "SambaCloud scores 66.4 (B) on agent readiness against Prism Inference's 60.1 (C), and leads in 4 of 7 scored categories. Prism Inference leads on security \u0026 auth. Both do inference fast. Category scores, facts, verdicts and agent notes side by side.",
    "facts": [
      "Prism Inference C 60.1",
      "SambaCloud B 66.4",
      "scores"
    ],
    "h1": "Prism Inference vs SambaCloud",
    "image": "https://www.anchorterminal.com/assets/og/compare-prism-inference-vs-sambanova.png",
    "path": "/compare/prism-inference-vs-sambanova",
    "published": "2026-10-01",
    "section": "tools",
    "title": "Prism Inference vs SambaCloud for AI agents, C 60.1 vs B 66.4",
    "toc": null,
    "updated": "2026-10-09",
    "url": "https://www.anchorterminal.com/compare/prism-inference-vs-sambanova"
  },
  "tokens": {
    "markdown": 2400,
    "slim": 680
  },
  "version": 1
}
