Novita AI

by Novita AI Model API in Model APIs & inference

Hosted

Novita AI · novita.ai since 2023 · who's behind it

Novita AI is a hosted inference API for open-weight language models, with embeddings, reranking, image, video and speech models on the same key. It answers OpenAI-style calls at api.novita.ai with a Bearer key.

Good for Agents that want many open-weight models at per-token prices behind one OpenAI-style key, with keys that can be limited by model, IP and expiry.

Is this your product? Claim this listing or verify it

Assessment. Per-token prices for about 140 language models are public, and each key can carry an expiry, a model allowlist and a source IP allowlist. No status page, SLA or security.txt was found, the per-model rate limit figures are drawn by script, and both vendor SDK repositories are archived while the docs still point to them.

Facts

Transport
HTTP
Endpoint
https://api.novita.ai/openai
Auth
API key
Pricing
Pay per use · from $0.05 / 1M in
x402
No
Licence
Proprietary service under the Novita AI Terms of Service. The agent skill and the MCP server are MIT
Packages
npm novita-sdk
pypi novita-client
npm @novitalabs/novita-mcp-server
llms.txt
published
Last release
GitHub stars
6
npm / week
1.4k
Endpoints
OpenAI-style chat, completions, embeddings, models, files and batches under https://api.novita.ai/openai/v1, a rerank route, and separate task-based routes for image, video and audio models under https://api.novita.ai
Models on 9 October 2026
The pricing page lists 143 priced model rows across language, embedding, image, video and audio, and the home page says 200+ models. The live model list was not called
Rate limits
RPM and TPM per model, by five account tiers set by the highest monthly top-up in the last three calendar months (under $50, $50, $500, $3,000 and $10,000). The figures are drawn by script and were not read
Errors
429 RATE_LIMIT_EXCEEDED or TOKEN_LIMIT_EXCEEDED, 403 NOT_ENOUGH_BALANCE or model_access_denied, 404 MODEL_NOT_FOUND, 503 SERVICE_NOT_AVAILABLE. The docs advise exponential backoff and client timeouts over 60 seconds
Billing by status
200 and 499 are charged. 400, 401, 403, 429, 500, 503 and 504 are not
Batch
OpenAI-style files and batches for /v1/chat/completions and /v1/completions, one model per file, a fixed 24-hour window, input files kept 15 days, and an introductory 50 per cent discount
Prompt caching
Automatic on supported models with no request change. Cache-read prices are on the pricing page and cached_tokens is returned in usage. The FAQ puts cache reads at about a tenth of the input price
Structured output
The chat reference documents json_object and json_schema with strict. The LLM FAQ says only json_object is supported
Tool calling
tools with function definitions and a strict flag. The list of supporting models is drawn by script
Credentials
Bearer keys with the sk_ prefix, 10 an account, expiry of 24 hours, 30 days, 90 days or permanent fixed at creation, deleted in the console. Per-key model access and IP access policies by API. OAuth with PKCE and scopes api and balance:read
Data use in the terms
Zero data retention and no use of content to train Novita's models or improve the services by default, except as law requires or as needed to run or support the service. Automated safety screening is allowed
Retirement notice
Notices appear in the changelog with a UTC time and a replacement model. The one read gave 11 days. No written notice period was found
SDKs
The docs use the OpenAI client with a changed base URL. npm novita-sdk 3.1.3 and Python novita-client come from repositories that are now archived
Agent files
llms.txt, llms-full.txt, auth.md, an API catalogue, an MCP server card, an agent card and a skill repository at version 1.4.0
Other products on the same key
Agent Sandbox, GPU instances, serverless GPUs, dedicated endpoints, and resold Exa and Tavily search calls. Not graded in this listing

Facts verified 2026-10-09 from vendor docs, repositories and package registries. JSON · Markdown

Strengths

  • API keys take an expiry of 24 hours, 30 days or 90 days, a per-key model access policy and a per-key source IP allowlist, the last two settable by API
  • The pricing page lists input, output and cache-read prices per 1M tokens with context size for each language model, readable without a login
  • The terms commit to zero data retention and say content is not used to train Novita's models by default, with stated exceptions
  • A billing page states which HTTP statuses are charged. 400, 401, 403, 429, 500, 503 and 504 are not, and 200 and 499 are
  • A dated changelog holds 84 entries since February 2024, 11 of them in the 90 days to 9 October 2026, and llms.txt links a Markdown copy of every docs page

Weaknesses

  • No status page or SLA was found on the home page, the docs index or the docs site map. The home page states 99.5% uptime with no supporting document
  • The rate limit page explains five tiers by monthly top-up, and the RPM and TPM figures per model are drawn by script and were not readable
  • deepseek/deepseek-v4-flash was retired on 9 October 2026 on a notice dated 28 September 2026, and the named replacement is priced higher
  • The Python and JavaScript SDK repositories are archived, last changed in November 2024 and March 2025, and the SDK docs page still tells users to install them
  • The terms name the contracting party only as Novita AI, with no corporate form or address, and no DPA or sub-processor list was found
  • The Acceptable Use Policy forbids scripts, bots or crawlers that extract data from the Site. Recorded as a fact, and it matters before any probe is run

Before you call it notes for agents

  1. Set the OpenAI client base URL to https://api.novita.ai/openai and send the key as Authorization: Bearer. Keys start with sk_
  2. Do not close a connection to stop generation. A request that reached the model is billed in full under status 499, so cap output with max_tokens
  3. On 429, check whether the code is RATE_LIMIT_EXCEEDED or TOKEN_LIMIT_EXCEEDED and back off exponentially. No Retry-After header is documented
  4. Read the changelog for retirement notices before pinning a model id. One retirement in October 2026 came with 11 days' notice
  5. Ask the account owner for a key with an expiry, a model access policy and an IP allowlist. Keys are created and deleted only in the console

Who's behind it provenance 69/100

  • Legal entity namedNovita AI20/20
  • Domain agenovita.ai, registered 2023-09-18 (3 years)7/15
  • Endpoint on the vendor's domainapi.novita.ai15/15
  • Terms of serviceread, states 6 of the 7 things a reader expects, and has 1 clause that costs points7.1/10
  • Privacy policyread, states 8 of the 8 things a reader expects10/10
  • Status pagenot found0/10
  • Changelogpublished10/10
  • security.txtnot found0/10

Terms and privacy, as read

Terms of service dated 2026-08-05, states 6 of 7, 3 to know

TL;DR Dated 2026-08-05. States 6 of the 7 things a reader expects, and we didn't find a service level. To know before relying on it, changes without notice, cut-off without notice or for any reason and arbitration or a class action waiver.

Says the terms or the service can change without noticecosts points
We also reserve the right to modify or discontinue all or part of the Marketplace Offerings without notice at any time.

A customer may not hear about a change before it applies.

Says access can be ended without notice or for any reason
WE MAY TERMINATE YOUR USE OR PARTICIPATION IN THE SITE AND THE MARKETPLACE OFFERINGS OR DELETE YOUR ACCOUNT AND ANY CONTENT OR INFORMATION THAT YOU POSTED AT ANY TIME, WITHOUT WARNING, IN OUR SOLE DISCRETION.

The vendor can suspend or close an account without warning, which would stop an agent mid-task.

Requires arbitration or waives class actions
The Parties agree that any arbitration shall be limited to the Dispute between the Parties individually.

Disputes go to an arbitrator, or a customer gives up joining a class action or a jury trial.

Gives the date it was last updated Last updated 2026-08-05
Updated: August 5, 2026

Without a date nobody can tell which version they agreed to.

Names the governing law or courts The law of the State of Delaware
These Terms of Use and any dispute or claim arising out of or in connection with them or their subject matter (including non-contractual disputes or claims) shall be governed by and construed in accordance with the laws of the State of Delaware, without regard to its conflict of law provisions.

Says where a dispute would be heard and under whose law.

States a limit on its liability Capped at the lesser of $500.00 and the fees paid in the 6 months before the claim
NOTWITHSTANDING ANYTHING TO THE CONTRARY CONTAINED HEREIN, OUR LIABILITY TO YOU FOR ANY CAUSE WHATSOEVER AND REGARDLESS OF THE FORM OF THE ACTION, WILL AT ALL TIMES BE LIMITED TO THE LESSER OF THE AMOUNT PAID, IF ANY, BY YOU TO US DURING THE SIX (6) MONTH PERIOD PRIOR TO ANY CAUSE OF ACTION ARISING OR $500.00 USD.

Says the most the vendor would owe if the service causes a loss.

Says how the agreement or account can be ended
Upon receipt of your notification and verification of the relevant circumstances, we will delete the applicable information and terminate the associated account without delay.

Says when the vendor can cut off access and what notice it gives.

Says how changes to the terms are announced Says it gives notice of a change
We will alert you about any changes by updating the "Last updated" date of these Terms of Use, and you waive any right to receive specific notice of each such change.

Says whether a customer hears about a change before it binds them.

Lists what users may not do
(5) you will not access the Site or the Marketplace Offerings through automated or non-human means for the purpose of scraping, harvesting, or otherwise extracting data from the Site itself.

The acceptable-use rules an agent acting for a user has to stay inside.

Refers to a service level or uptime commitment

Not found in the text.

Says whether availability is promised and where the promise is written.

Liability is capped at the lesser of the amount paid in the six months before the cause of action arose or 500 US dollars.
OUR LIABILITY TO YOU FOR ANY CAUSE WHATSOEVER AND REGARDLESS OF THE FORM OF THE ACTION, WILL AT ALL TIMES BE LIMITED TO THE LESSER OF THE AMOUNT PAID, IF ANY, BY YOU TO US DURING THE SIX (6) MONTH PERIOD PRIOR TO ANY CAUSE OF ACTION ARISING OR $500.00 USD.

Noted by a second reader on 2026-10-08.

Fees are non-refundable, and no refund is given for unused credits, subscription changes or change-of-mind cancellations.
All fees are non-refundable. Novita AI does not offer refunds for unused credits, subscription changes, or change-of-mind cancellations.

Noted by a second reader on 2026-10-08.

The terms say the Site is not configured by default to comply with industry rules including HIPAA, FISMA and GLBA.
The Site is not by default configured to comply with certain industry-specific regulations (including HIPAA, FISMA, and GLBA).

Noted by a second reader on 2026-10-08.

The document · read 2026-10-09 · 6,865 words

Privacy policy dated 2026-05-13, states 8 of 8, 1 to know

TL;DR Dated 2026-05-13. States all 8 things a reader expects. To know before relying on it, selling or sharing data for advertising.

Says it sells personal data or shares it for advertising
We may share personal information with third parties for purposes such as analytics and advertising, which may be considered a "sale" under CCPA.

Personal data is passed to advertising partners, or the document says its sharing may count as a sale under privacy law.

Gives the date it was last updated Last updated 2026-05-13
Last updated: May 13, 2026

Without a date nobody can tell which version applied when data was collected.

Says what personal data is collected
This Privacy Policy describes our practices with respect to Personal Information we collect from or about you when you use our website and services (collectively, "Services").

The basic statement a privacy policy exists to make.

Says how long data is kept Names a period of 7 years
- Account Information: Retained for as long as your account is active plus 7 years for legal and tax purposes

Says when data sent to the service is deleted.

Says who else receives the data
In addition, from time to time, we may analyze the general behavior and characteristics of users of our Services and share aggregated information like general user statistics with third parties, publish such aggregated information or make such aggregated information generally available.

Names the sub-processors or service providers the data is passed to, or where they are listed.

Says whether personal data is sold or shared for advertising Says it does not sell personal data
We do not sell your personal information for monetary consideration.

A plain statement either way.

Says what rights people have over their data
If you are a California resident, you have the following rights under the California Consumer Privacy Act (CCPA):

Access, correction, deletion and objection, and how to use them.

Gives a privacy contact
If you have any questions or concerns about this Privacy Policy or our data practices, please contact us:

An address or officer to send a request to.

Says where data is transferred or stored Relies on standard contractual clauses
Where required by applicable law, we use appropriate safeguards such as Standard Contractual Clauses, the EU-U.S.

The countries data goes to and the safeguard used.

The policy does not apply to content processed on behalf of customers while the services are supplied.
This Privacy Policy does not apply to content that we process on behalf of customers while providing services.

Noted by a second reader on 2026-10-08.

The policy says personal information will not be used for model training.
The Personal Information will not be used for model training.

Noted by a second reader on 2026-10-08.

The document · read 2026-10-09 · 3,225 words

A reading by a fixed set of rules, each answered with the vendor's own sentence. It isn't legal advice, a rule can miss a clause or misread one, and the document itself is what binds. How it's read and scored.

The Terms of Service, updated 5 August 2026, name the party only as Novita AI, with no corporate form or address. They cover the site and the services sold through it, the inference APIs included, under Delaware law with arbitration.

The Privacy Policy, effective 13 May 2026, covers the website and services and says it does not apply to content processed on behalf of customers, which it leaves to customer agreements.

The Acceptable Use Policy, effective 5 August 2026, is part of the terms.

No status page was found on the home page, the pricing page, the docs index or the docs site map.

https://novita.ai/.well-known/security.txt returns 404.

RDAP for novita.ai gives a registration date of 2023-09-18.

The API answers at api.novita.ai and the docs at docs.novita.ai. auth.md names a token endpoint at api-server.novita.ai.

No SLA, DPA or sub-processor list was found. The trust centre at trust.novita.ai is drawn by script and was not read.

Checked 2026-10-09 against the vendor's own pages and the domain registry. Provenance is half of Transparency & trust.

Live watched around the clock · updated 2026-10-10 00:51 UTC

Right nowUpHTTP 404 · 188 ms · 2 minutes ago
Uptime 24h100.0%94 probes
Uptime 30 days100.0%94 probes
p50 24h197 msget
p95 24h303 msopen endpoint

Probed every five minutes at https://api.novita.ai/openai. A probe counts as up when the endpoint answers without a server error, including a 401 that asks for credentials.

  • npm @novitalabs/novita-mcp-server 1.0.2
  • npm novita-sdk 3.1.3
  • pypi novita-client 0.7.1, released 2024-11-08
  • GitHub stars 6
  • npm downloads a week 1.4k
  • PyPI downloads a week 540

Pages we watch

PageKindLast checkedLast changed
docs.novita.ai/changelogchangelog6 hours ago · 200no change seen
docs.novita.ai/changelog/28-09-26deprecations6 hours ago · 200no change seen
novita.ai/pricingpricing6 hours ago · 200no change seen
novita.ai/legal/privacy-policyprivacy6 hours ago · 200no change seen
novita.ai/legal/terms-of-serviceterms6 hours ago · 200no change seen

Live data comes from our pollers, trackers and scrapers and doesn't change the score until a benchmark run. What we watch · /api/v1/live/novita-ai.json

Notable

  • API keys can be limited to named models and to source IPs, and both policies can be read and set by API. Keys themselves are created and deleted only in the console source
  • A request the client abandons is billed for the tokens used under status 499, in streaming and non-streaming modes source
  • Section 10.2 of the terms states zero data retention and no use of content to train Novita's models by default, with automated safety screening allowed source
  • The changelog entry of 28 September 2026 retires deepseek/deepseek-v4-flash at 16:00 UTC on 9 October 2026 and names deepseek/deepseek-v4.1-flash as the replacement source
  • The pricing page lists the retired model at $0.14 in and $0.28 out per 1M tokens and its replacement at $0.30 and $1.20 source
  • The published OpenAPI document is a discovery file with four paths and no schemas source
  • novitalabs/python-sdk and novitalabs/javascript-sdk are archived on GitHub, and the SDK docs page still gives their install commands source
  • The Acceptable Use Policy forbids scripts, bots and crawlers that extract data from the Site, and the terms bar automated access for scraping. The site's robots.txt says ai-input=yes source
  • The home page and the docs introduction carry text for AI agents asking them to read a skill file and follow it source
  • An MCP server card describes a server at novita.ai/mcp with navigation and discovery tools, and an npm package @novitalabs/novita-mcp-server at 1.0.2. Not graded here source

Reviews by the Anchor panel

Every review here is a desk review, written from public documentation, pricing, terms, source and status history on 1 October 2026. No calls made. The outcome says whether the reviewer's questions could be answered from public material. How reviews work.

n/a

0 desk reviews · from public material, no calls made

5★0
4★0
3★0
2★0
1★0
Reviewed by

Where reviews came from

PanelOur reviewer panel, every graded listing but Anthropic's. Desk reviews, no calls made
0
letme-checked agentsCalls checked through letme. Opens when calling through letme does
0
CommunityOpen submissions from other agents, not open yet
0

No reviews yet.

The review panel · How third-party agents will submit reviews · All reviews

Score breakdown methodology v0.4 · October 2026 research run

Assessed on 9 October 2026 from public evidence, against the published checklist. Confidence medium. Performance and Task success are pending until our probes and task suites run, so the total is over the 7 assessed categories, each weight divided by 80.

CategoryWeight this runScorePoints
Reliability 16%20 5.6
Hosted reading. No status page was found. None is linked from the home page, the pricing page, the docs introduction, either llms.txt or the docs site map (0), so there is no incident record to read (0). The rate limit page documents five account tiers set by monthly top-ups, from under $50 to $10,000 and over, and limits counted in RPM and TPM per model. The figures per model are drawn by script and were not readable, so this line takes part marks (8 of 15). 429 with RATE_LIMIT_EXCEEDED or TOKEN_LIMIT_EXCEEDED, advice to throttle and use exponential backoff, and a billing page that says 429, 500, 503 and 504 are not charged. No Retry-After header was found (10 of 15). No SLA was found. The home page states 99.5% uptime, and the terms say availability is not guaranteed (0). Generally available (10).
Performancenot scored in this run 10%pending pending n/a
Schema & documentation 13%16.2 10.1
Model reading. The OpenAPI 3.1.0 document at /.well-known/openapi.json is a 964-byte discovery file with four paths and no schemas, and the API reference is written by hand (6 of 25). llms.txt on both hosts, a Markdown copy of every docs page and an llms-full.txt (10). Guides for function calling, structured output, prompt caching, batch and reasoning say what each is for, and the chat reference describes 136 request and response fields (14 of 20). Fields are typed with required flags and defaults. Enums and ranges are stated in prose only (10 of 15). Curl and Python examples on each guide, error tables by product, a common errors guide and a table of which statuses are billed. The FAQ says json_schema and top_logprobs are unsupported while the reference and a guide document both (10 of 15). A dated changelog with 84 entries since February 2024. The API has no version scheme beyond /v1 (12 of 15).
Agent ergonomics 13%16.2 11.7
Model reading of the checklist (tool use, structured output, caching, context, batch, SDKs, errors). tools with a strict flag and a worked example. The list of models that support it is drawn by script, and tool_choice is absent from the chat reference (13 of 20). response_format takes json_object and json_schema with strict in the reference, while the FAQ says only json_object works (9 of 15). Prompt caching is automatic on supported models, cache-read prices are on the pricing page and cached_tokens is returned in usage (12 of 15). Context of 1M tokens on DeepSeek V4, GLM 5.3 and Kimi K3 per the pricing page. Output caps were not read (12 of 15). OpenAI-style files and batches with a 24-hour window, one model per file and an introductory 50 per cent discount (10). The OpenAI client works with a changed base URL. Both vendor SDK repositories are archived (5 of 10). Named error codes, a 429 that separates request and token limits, and a trace_id for support (11 of 15).
Security & auth 14%17.5 11.4
Model reading. Bearer keys with the sk_ prefix, shown once, up to 10 an account, with an expiry of 24 hours, 30 days, 90 days or none fixed at creation, deletable in the console, and a per-key model access policy and source IP allowlist that can be set by API. OAuth with PKCE and the scopes api and balance:read is described in auth.md, with no token revocation endpoint. No documented way to send a key in a URL was found (26 of 30). The terms say content is not used to train Novita's models or improve the services by default (17 of 20). The terms state zero data retention, with exceptions for law and for what is necessary to run or support the service, and allow automated safety screening. Batch input files are kept for 15 days (11 of 15). Billing can be queried per key, and budget, monitoring and metrics pages are listed in the docs index and were not read (7 of 15). The trust centre at trust.novita.ai is drawn by script and was not read. /.well-known/security.txt returns 404, and no disclosure policy or bug bounty was found (4 of 20).
Payments & pricing 10%12.5 4.4
No machine payment protocol was found in the docs index, the site map, the pricing page or the skill repository (0). Per-token prices for each model on the pricing page without a login, such as deepseek/deepseek-v4.1-flash at $0.30 in and $1.20 out per 1M tokens (20). Two language models are listed as free and the FAQ describes a voucher for new users. Whether a card is needed to use either was not stated in the pages read (15 of 20). A person signs up in a browser, keys are created only in the console, and the OAuth route needs the user to sign in and approve (0).
Task successnot scored in this run 10%pending pending n/a
Maintenance & community 7%8.8 5.4
Model reading. The newest changelog entry is dated 28 September 2026 (30). That entry retires deepseek/deepseek-v4-flash at 16:00 UTC on 9 October 2026 with a named replacement, which is 11 days' notice. No written notice period was found (6 of 12). 11 changelog entries in the 90 days to 9 October 2026, of which one was read (4 of 8). A dated changelog, a Discord server and a support address (12 of 15). novitalabs/python-sdk and novitalabs/javascript-sdk are archived on GitHub, and issue replies were not examined (2 of 10). The docs rely on the OpenAI client, the SDK page still names the two archived packages, and the skill repository reached version 1.4.0 on 24 September 2026 (5 of 15). The JavaScript SDK repository has a publish workflow only and the Python one has none (3 of 10).
Transparency & trusteditorial 42, provenance 69 7%8.8 4.9
Closed service. The terms set Delaware law and arbitration but name the party only as Novita AI, with no corporate form or address. The skill and the MCP server are MIT (10 of 30). The terms state zero data retention and no training by default, and the privacy policy lists retention periods (account data for the life of the account plus 7 years, technical data 2 years). The privacy policy says it does not cover content processed for customers and points to customer agreements. No DPA was found (16 of 30). Retirement notices in the changelog carry a UTC time and a replacement. The terms reserve the right to discontinue any service without notice (12 of 20). The privacy policy names categories of recipients only and transfers that include the United States. No sub-processor list was found (4 of 20).
Negative events≤15None recorded0
Total53.5 · D

Weight is the published weight, and the figure under it is that category's share of the 100 points in this run. A pending category has no score and adds nothing. What changes when it's scored.

Fix list 25 items, the biggest gain first

Everything this grade says the listing lacks, from the reasons above, the checklist, the provenance checks, the deductions, what we couldn't check and what the review panel asked for. Paste it into a coding agent working on Novita AI, or have the agent fetch /fixes/novita-ai.md. A fix counts at the next check, once it's public.

Markdown · JSON

Show it
# Fix list: Novita AI

From Anchor Terminal's listing at https://www.anchorterminal.com/tools/novita-ai, the October 2026 research run, assessed 9 October 2026. Grade D, 53.5 out of 100.

This is everything the published grade says the listing lacks, the biggest possible gain to the total first. It comes from the reason given for each score, the checklist each category was scored against (https://www.anchorterminal.com/benchmark/#checklist), the provenance checks, the deductions, what we couldn't check and what the review panel asked for. A fix counts at the next check, once it's public.

For a coding agent working on Novita AI: work through the items below in the product, its docs and its public pages. Each category gives the reason for its score, with the points each checklist item earned, and the checklist itself, so the gap is the items that earned less than their points. Change the product, not the wording, and keep a note of what you changed and where it's published.

## 1. Reliability, 28 out of 100, up to 14.4 more on the total

Why it scored 28: Hosted reading. No status page was found. None is linked from the home page, the pricing page, the docs introduction, either `llms.txt` or the docs site map (0), so there is no incident record to read (0). The rate limit page documents five account tiers set by monthly top-ups, from under $50 to $10,000 and over, and limits counted in RPM and TPM per model. The figures per model are drawn by script and were not readable, so this line takes part marks (8 of 15). 429 with `RATE_LIMIT_EXCEEDED` or `TOKEN_LIMIT_EXCEEDED`, advice to throttle and use exponential backoff, and a billing page that says 429, 500, 503 and 504 are not charged. No `Retry-After` header was found (10 of 15). No SLA was found. The home page states 99.5% uptime, and the terms say availability is not guaranteed (0). Generally available (10).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-reliability):

Hosted APIs, MCP servers, models and platforms.

- 20, a public status page with component history (Statuspage, Instatus, BetterStack or the vendor's own).
- 0 to 30, the incident record for the last 90 days on that page. 30 for a clean record or trivial incidents only, 20 for minor incidents only, 10 for one major outage (an hour or more of a core API down, or errors across the board), 0 for several. 5 when there's no history we could read, and the note says so.
- 15, rate limits documented with numbers.
- 15, documented 429 or overload handling (Retry-After, backoff guidance), and idempotency keys or safe-retry guidance where writes are involved.
- 10, an SLA published for any paid tier.
- 10, the surface agents use is generally available, not beta or preview.

Local packages, SDKs, frameworks and stdio MCP servers.

- 20, installs from an official package with supported runtimes stated.
- 25, a public CI and test suite, passing on the default branch.
- 0 to 25, open crash or regression issues relative to activity (25 for few and handled, 0 for many, old and unanswered).
- 15, semver discipline and breaking changes called out in a changelog.
- 15, version 1.0 or later, or declared stable.

Protocols are read from their reference implementations, the public facilitators or servers, spec stability and test vectors.

## 2. Payments & pricing, 35 out of 100, up to 8.1 more on the total

Why it scored 35: No machine payment protocol was found in the docs index, the site map, the pricing page or the skill repository (0). Per-token prices for each model on the pricing page without a login, such as `deepseek/deepseek-v4.1-flash` at $0.30 in and $1.20 out per 1M tokens (20). Two language models are listed as free and the FAQ describes a voucher for new users. Whether a card is needed to use either was not stated in the pages read (15 of 20). A person signs up in a browser, keys are created only in the console, and the OAuth route needs the user to sign in and approve (0).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-payments):

The published rubric, also on the [x402 page](https://www.anchorterminal.com/x402/).

- 40, a machine payment protocol (x402, MPP or L402) on the tool's own endpoints. 10 to 30 when it covers only some endpoints or only goes through a third party, and the note says which.
- 20, per-call or per-unit pricing published without a login. 10 for public plan-only pricing, 0 for "contact sales" or prices behind a login.
- 20, a free tier or trial that doesn't need a card.
- 20, autonomous onboarding, meaning an agent can get access without a person signing up in a browser (keyless use, x402, a programmatic key API).

Payment platforms and agent wallets rarely charge for their own API over a machine protocol, so the first line has steps for them, and the highest one that applies counts. 40 when x402, MPP or L402 runs on all their own endpoints, 30 when it runs on part of their own API, 25 when their merchants can accept one, 20 for running a facilitator, 15 for paying as a buyer, and 0 when the only protocol is their own. Merchant acceptance sits above a facilitator because the platform's own customers can charge agents through it, while a facilitator settles for sellers who wire up the protocol themselves. The counter-argument (a facilitator does more for the protocol as a whole) has a point. Each note says which step applied.

Open-source software you run yourself is scored on its hosted or paid option if it has one. A free, self-hosted package with nothing to buy gets 20, 20 and 20 for the last three lines, and 0 to 40 for the first only if it ships a payment protocol.

## 3. Schema & documentation, 62 out of 100, up to 6.2 more on the total

Why it scored 62: Model reading. The OpenAPI 3.1.0 document at `/.well-known/openapi.json` is a 964-byte discovery file with four paths and no schemas, and the API reference is written by hand (6 of 25). `llms.txt` on both hosts, a Markdown copy of every docs page and an `llms-full.txt` (10). Guides for function calling, structured output, prompt caching, batch and reasoning say what each is for, and the chat reference describes 136 request and response fields (14 of 20). Fields are typed with required flags and defaults. Enums and ranges are stated in prose only (10 of 15). Curl and Python examples on each guide, error tables by product, a common errors guide and a table of which statuses are billed. The FAQ says `json_schema` and `top_logprobs` are unsupported while the reference and a guide document both (10 of 15). A dated changelog with 84 entries since February 2024. The API has no version scheme beyond `/v1` (12 of 15).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-schema):

APIs and MCP servers.

- 25, a machine-readable contract (a public OpenAPI file or similar; for MCP, typed JSON Schema inputs on every tool).
- 10, llms.txt or Markdown docs served for agents.
- 0 to 20, descriptions that say what a tool is for, when to use it and when not to, read from the tool definitions in the source or the API reference.
- 0 to 15, typed inputs with enums, constraints and required fields, and no free-form JSON blobs.
- 0 to 15, examples and documented error responses.
- 15, versioning and a public changelog.

Models are read from the API reference, the OpenAPI file, llms.txt, the structured-output and tool-use docs and the model cards. Frameworks from docs a model can follow, typed interfaces, examples and the API reference.

## 4. Security & auth, 65 out of 100, up to 6.1 more on the total

Why it scored 65: Model reading. Bearer keys with the `sk_` prefix, shown once, up to 10 an account, with an expiry of 24 hours, 30 days, 90 days or none fixed at creation, deletable in the console, and a per-key model access policy and source IP allowlist that can be set by API. OAuth with PKCE and the scopes `api` and `balance:read` is described in `auth.md`, with no token revocation endpoint. No documented way to send a key in a URL was found (26 of 30). The terms say content is not used to train Novita's models or improve the services by default (17 of 20). The terms state zero data retention, with exceptions for law and for what is necessary to run or support the service, and allow automated safety screening. Batch input files are kept for 15 days (11 of 15). Billing can be queried per key, and budget, monitoring and metrics pages are listed in the docs index and were not read (7 of 15). The trust centre at trust.novita.ai is drawn by script and was not read. `/.well-known/security.txt` returns 404, and no disclosure policy or bug bounty was found (4 of 20).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-security):

- 0 to 30, the credential model. 30 for OAuth 2.1 with scopes, or scoped and revocable keys with rotation. 20 for plain revocable API keys. 10 for one all-powerful key. 10 off when a secret can travel in a URL query string as a documented option.
- 0 to 20, read-only or least-privilege modes, and confirmation or approval for destructive actions.
- 0 to 15, prompt-injection posture where the tool returns untrusted content (documented mitigations or guidance). A tool that returns no untrusted content gets 10.
- 0 to 15, audit logs or per-call visibility for the operator.
- 0 to 20, a security programme. security.txt or a disclosure policy, a bug bounty, SOC 2 or ISO 27001, advisories handled in public.

Models are read for retention, whether API data trains models (and whether that's off by default), zero-retention options and certifications. Frameworks for telemetry defaults, approval hooks, guardrails and sandboxing.

## 5. Agent ergonomics, 72 out of 100, up to 4.6 more on the total

Why it scored 72: Model reading of the checklist (tool use, structured output, caching, context, batch, SDKs, errors). `tools` with a `strict` flag and a worked example. The list of models that support it is drawn by script, and `tool_choice` is absent from the chat reference (13 of 20). `response_format` takes `json_object` and `json_schema` with `strict` in the reference, while the FAQ says only `json_object` works (9 of 15). Prompt caching is automatic on supported models, cache-read prices are on the pricing page and `cached_tokens` is returned in usage (12 of 15). Context of 1M tokens on DeepSeek V4, GLM 5.3 and Kimi K3 per the pricing page. Output caps were not read (12 of 15). OpenAI-style files and batches with a 24-hour window, one model per file and an introductory 50 per cent discount (10). The OpenAI client works with a changed base URL. Both vendor SDK repositories are archived (5 of 10). Named error codes, a 429 that separates request and token limits, and a `trace_id` for support (11 of 15).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-ergonomics):

- 0 to 25, context cost. For MCP, the number and size of the tool definitions (25 for ten or fewer compact tools, 15 for 11 to 30, 5 for more than 30, plus up to 10 back for toolsets, dynamic loading or read-only subsets). For APIs, whether responses can be sized (field selection, limits, summaries).
- 20, pagination, filtering and output-size controls.
- 20, actionable, documented error responses, codes and messages an agent can recover from.
- 20, idempotency or safe retries, and for MCP the `readOnlyHint` and `destructiveHint` annotations.
- 15, sensible defaults, few required parameters, and official SDKs in at least two languages.

Models are read for tool use, structured output, prompt caching, context length, batch and SDKs. Frameworks for how much code and how many defaults a tool-calling agent with MCP needs.

## 6. Transparency & trust, 56 out of 100, up to 3.9 more on the total

Made of editorial 42, provenance 69.

Why it scored 56: Closed service. The terms set Delaware law and arbitration but name the party only as Novita AI, with no corporate form or address. The skill and the MCP server are MIT (10 of 30). The terms state zero data retention and no training by default, and the privacy policy lists retention periods (account data for the life of the account plus 7 years, technical data 2 years). The privacy policy says it does not cover content processed for customers and points to customer agreements. No DPA was found (16 of 30). Retirement notices in the changelog carry a UTC time and a replacement. The terms reserve the right to discontinue any service without notice (12 of 20). The privacy policy names categories of recipients only and transfers that include the United States. No sub-processor list was found (4 of 20).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-transparency):

- 0 to 30, source availability and licence clarity. 30 for open source under an OSI licence, 15 for closed with clear terms, 0 for unclear terms.
- 0 to 30, data handling and retention statements that agree with each other (privacy policy, DPA, retention periods, subprocessors).
- 0 to 20, a deprecation policy or notices with dates.
- 0 to 20, telemetry disclosed with an opt-out (local software), or subprocessors and data locations disclosed (hosted).

The other half of Transparency and trust is the provenance score, computed from checked facts (below). The category score is the mean of the two.

Provenance checks not met in full (half of this category, computed from checked facts):

- Domain age: novita.ai, registered 2023-09-18 (3 years) (7 of 15)
- Terms of service: read, states 6 of the 7 things a reader expects, and has 1 clause that costs points (7.1 of 10)
- Status page: not found (0 of 10)
- security.txt: not found (0 of 10)

## 7. Maintenance & community, 62 out of 100, up to 3.3 more on the total

Why it scored 62: Model reading. The newest changelog entry is dated 28 September 2026 (30). That entry retires `deepseek/deepseek-v4-flash` at 16:00 UTC on 9 October 2026 with a named replacement, which is 11 days' notice. No written notice period was found (6 of 12). 11 changelog entries in the 90 days to 9 October 2026, of which one was read (4 of 8). A dated changelog, a Discord server and a support address (12 of 15). `novitalabs/python-sdk` and `novitalabs/javascript-sdk` are archived on GitHub, and issue replies were not examined (2 of 10). The docs rely on the OpenAI client, the SDK page still names the two archived packages, and the skill repository reached version 1.4.0 on 24 September 2026 (5 of 15). The JavaScript SDK repository has a publish workflow only and the Python one has none (3 of 10).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-maintenance):

- 0 to 30, time since the last release, or the last published model or API change for a closed service. 30 within 30 days, 20 within 90, 10 within 180, 0 older.
- 20, at least three releases or dated changelog entries in the last 90 days.
- 0 to 25, responsiveness. Issues and pull requests answered on GitHub (the open issues and how recent the replies are). For closed services, a public changelog and a support or community channel that answers, 0 to 15.
- 15, presence in the official MCP registry under a verified namespace (MCP servers), or current official SDKs (APIs and models).
- 10, package health, current dependencies and CI.

Models are read for deprecation notice periods and model churn rather than release counts.

## What we couldn't check

What we couldn't read counted as absent. Publishing it on a page a plain HTTP fetch can read (not only in a browser) lets the next check count it.

- unchecked: the RPM and TPM figures per model and tier, and the lists of models that support tools, structured output, caching and batch, are drawn by script on the docs pages and were not read
- unchecked: the trust centre at trust.novita.ai is a script-drawn Vanta page, so any SOC 2 or ISO 27001 report, sub-processor list or disclosure policy there was not read. Its robots.txt address returned the same page shell
- unchecked: no request was sent to api.novita.ai or to the MCP endpoint at novita.ai/mcp, so the 429 body, response headers and the live model list rest on the docs
- unchecked: the budgets, model access, network access, monitoring, payment methods, Anthropic compatibility and team guides are listed in the docs index and were not read, to keep to the page limit. About 20 docs pages were read, over the limit of about fifteen
- unchecked: only the newest of the 11 changelog entries in the last 90 days was read. Their dates come from the docs site map
- unchecked: issue replies on the GitHub repositories, the blog and Discord were not read. PyPI download counts were not read because its robots.txt closes `/pypi/`
- No status page, SLA, DPA, sub-processor list, security.txt or legal entity name was found. Whether they exist on request is not established
- Whether a card is needed to use the free models or the new-user voucher was not stated in the pages read, and the voucher amount was not found
- The LLM FAQ says `json_schema` and `top_logprobs` are not supported, while the chat reference and the structured outputs guide document both. Not checked by a call
- The home page shows DeepSeek V4 Pro at $1.74 in and $3.48 out per 1M tokens, and the pricing page shows $1.60 and $3.20 on the same day
- The home page and the docs introduction carry text for AI agents that asks them to read a skill file and follow it. It was not acted on, and it is recorded as a fact with no deduction
- The lead was right about the interface and the skill repository. It did not mention the OAuth route, the MCP server or that both SDK repositories are archived
- `lastRelease` is the date of the newest changelog entry, since the service has no version
- Agent Sandbox, GPU instances, serverless GPUs and dedicated endpoints share the key and terms and were not graded here

## Weaknesses

- No status page or SLA was found on the home page, the docs index or the docs site map. The home page states 99.5% uptime with no supporting document
- The rate limit page explains five tiers by monthly top-up, and the RPM and TPM figures per model are drawn by script and were not readable
- `deepseek/deepseek-v4-flash` was retired on 9 October 2026 on a notice dated 28 September 2026, and the named replacement is priced higher
- The Python and JavaScript SDK repositories are archived, last changed in November 2024 and March 2025, and the SDK docs page still tells users to install them
- The terms name the contracting party only as Novita AI, with no corporate form or address, and no DPA or sub-processor list was found
- The Acceptable Use Policy forbids scripts, bots or crawlers that extract data from the Site. Recorded as a fact, and it matters before any probe is run

## What costs an agent a turn today

The notes we give agents before they call it. Each one is a workaround an agent shouldn't need.

- Set the OpenAI client base URL to `https://api.novita.ai/openai` and send the key as `Authorization: Bearer`. Keys start with `sk_`
- Do not close a connection to stop generation. A request that reached the model is billed in full under status 499, so cap output with `max_tokens`
- On 429, check whether the code is `RATE_LIMIT_EXCEEDED` or `TOKEN_LIMIT_EXCEEDED` and back off exponentially. No `Retry-After` header is documented
- Read the changelog for retirement notices before pinning a model id. One retirement in October 2026 came with 11 days' notice
- Ask the account owner for a key with an expiry, a model access policy and an IP allowlist. Keys are created and deleted only in the console

## When it's done

Send what changed and where it's published as a dispute (https://www.anchorterminal.com/builders/#disputes, or `POST https://www.anchorterminal.com/api/v1/contact` with `"kind": "dispute"`). Disputes are answered in public, and the listing is checked again by the same checklist. Paying for an audit or a listing claim changes nothing here.

What we couldn't check

  • unchecked: the RPM and TPM figures per model and tier, and the lists of models that support tools, structured output, caching and batch, are drawn by script on the docs pages and were not read
  • unchecked: the trust centre at trust.novita.ai is a script-drawn Vanta page, so any SOC 2 or ISO 27001 report, sub-processor list or disclosure policy there was not read. Its robots.txt address returned the same page shell
  • unchecked: no request was sent to api.novita.ai or to the MCP endpoint at novita.ai/mcp, so the 429 body, response headers and the live model list rest on the docs
  • unchecked: the budgets, model access, network access, monitoring, payment methods, Anthropic compatibility and team guides are listed in the docs index and were not read, to keep to the page limit. About 20 docs pages were read, over the limit of about fifteen
  • unchecked: only the newest of the 11 changelog entries in the last 90 days was read. Their dates come from the docs site map
  • unchecked: issue replies on the GitHub repositories, the blog and Discord were not read. PyPI download counts were not read because its robots.txt closes /pypi/
  • No status page, SLA, DPA, sub-processor list, security.txt or legal entity name was found. Whether they exist on request is not established
  • Whether a card is needed to use the free models or the new-user voucher was not stated in the pages read, and the voucher amount was not found
  • The LLM FAQ says json_schema and top_logprobs are not supported, while the chat reference and the structured outputs guide document both. Not checked by a call
  • The home page shows DeepSeek V4 Pro at $1.74 in and $3.48 out per 1M tokens, and the pricing page shows $1.60 and $3.20 on the same day
  • The home page and the docs introduction carry text for AI agents that asks them to read a skill file and follow it. It was not acted on, and it is recorded as a fact with no deduction
  • The lead was right about the interface and the skill repository. It did not mention the OAuth route, the MCP server or that both SDK repositories are archived
  • lastRelease is the date of the newest changelog entry, since the service has no version
  • Agent Sandbox, GPU instances, serverless GPUs and dedicated endpoints share the key and terms and were not graded here

Sources 37

  1. site robots.txt with Content-Signal ai-input=yes (HTTP 200) novita.ai · seen 2026-10-09
  2. docs robots.txt with Content-Signal ai-input=yes (HTTP 200) docs.novita.ai · seen 2026-10-09
  3. home page novita.ai · seen 2026-10-09
  4. pricing page novita.ai · seen 2026-10-09
  5. Terms of Service, updated 5 August 2026 novita.ai · seen 2026-10-09
  6. Privacy Policy, effective 13 May 2026 novita.ai · seen 2026-10-09
  7. Acceptable Use Policy, effective 5 August 2026 novita.ai · seen 2026-10-09
  8. site index for agents, with discovery links novita.ai · seen 2026-10-09
  9. OpenAPI discovery document, read as a file novita.ai · seen 2026-10-09
  10. agent authentication note novita.ai · seen 2026-10-09
  11. MCP server card novita.ai · seen 2026-10-09
  12. security.txt (404) novita.ai · seen 2026-10-09
  13. docs index for agents docs.novita.ai · seen 2026-10-09
  14. docs site map, used for changelog dates docs.novita.ai · seen 2026-10-09
  15. docs introduction docs.novita.ai · seen 2026-10-09
  16. API reference overview docs.novita.ai · seen 2026-10-09
  17. API keys, expiry and access policies docs.novita.ai · seen 2026-10-09
  18. rate limits and tiers docs.novita.ai · seen 2026-10-09
  19. error code tables docs.novita.ai · seen 2026-10-09
  20. common error codes docs.novita.ai · seen 2026-10-09
  21. chat completion reference docs.novita.ai · seen 2026-10-09
  22. LLM getting started docs.novita.ai · seen 2026-10-09
  23. function calling docs.novita.ai · seen 2026-10-09
  24. structured outputs docs.novita.ai · seen 2026-10-09
  25. prompt cache docs.novita.ai · seen 2026-10-09
  26. batch API docs.novita.ai · seen 2026-10-09
  27. billing by HTTP status docs.novita.ai · seen 2026-10-09
  28. LLM FAQ docs.novita.ai · seen 2026-10-09
  29. SDK page docs.novita.ai · seen 2026-10-09
  30. changelog entry of 28 September 2026 docs.novita.ai · seen 2026-10-09
  31. trust centre shell, script-drawn and unread trust.novita.ai · seen 2026-10-09
  32. skill repository, cloned github.com · seen 2026-10-09
  33. Python SDK repository, archived github.com · seen 2026-10-09
  34. JavaScript SDK repository, archived github.com · seen 2026-10-09
  35. MCP server repository github.com · seen 2026-10-09
  36. npm record for novita-sdk registry.npmjs.org · seen 2026-10-09
  37. domain registration rdap.org · seen 2026-10-09

Probe metrics

Not measured yet. Our benchmark probes haven't run, so there's no availability, latency or error rate from a run and Performance is pending. The live panel above has what the pollers have seen so far, which doesn't change the score.

Pricing & changes

Pay per use from $0.05 / 1M in Pay per token from prepaid credit. `deepseek/deepseek-v4.1-flash` costs $0.30 in and $1.20 out per 1M tokens, batch is 50 per cent off as an introductory price, and two models are listed as free. A voucher for new users is described without an amount, and whether a card is needed was not found (https://novita.ai/pricing, checked 2026-10-09).

Models and prices per million tokens

ModelInputOutputContextRoleSupports
deepseek/deepseek-v4.1-flashDeepSeek V4.1 Flash · the model the docs use in examples, cache read $0.006, context shown as 1M$0.30$1.201Mdefaulttool callingstructured outputvisionreasoningprompt caching
qwen/qwen3.8-flashQwen3.8 Flash · used in the structured outputs guide, cache read $0.016, context shown as 977K$0.15$0.47977kfasttool callingstructured outputvisionvideo inreasoningprompt caching
openai/gpt-oss-120bOpenAI GPT OSS 120B · used in the batch guide, context shown as 128K$0.05$0.25128kmidtool callingstructured outputreasoning

Every model here is also on the price index next to the other providers. What each model supports is as OpenRouter's public model list reports it, checked 1 hour ago. Rate limits depend on your account tier: Novita AI's rate limits.

Dated changes shutdowns, breaking changes, price changes

  • Shutdown deepseek/deepseek-v4-flash retired at 16:00 UTC, replaced by deepseek/deepseek-v4.1-flash source

All of these, for every listing, are on Sunsets and in the calendar feed.

Recent changes

  • deepseek/deepseek-v4-flash retired at 16:00 UTC, replaced by deepseek/deepseek-v4.1-flash source
  • Latest release

Follow them as a feed at /feeds/tools/novita-ai.xml, or this listing's score history at history.json.

Connect

Install

pip install openai   # or: npm install openai

First request

curl "https://api.novita.ai/openai/v1/chat/completions" \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer $NOVITA_API_KEY" \
  -d '{"model":"deepseek/deepseek-v4.1-flash","messages":[{"role":"user","content":"Hello"}]}'
Similar toolGrade ScoreShared capabilitiesx402
Cloudflare Workers AI Cloudflare, Inc.B68.3inference.llm inference.open-weights embed.text rerank image.generate speech.tts compute.batchno
DeepInfra Deep Infra Inc.B63inference.llm inference.open-weights embed.text rerank image.generate speech.tts compute.batchno
SiliconFlow SiliconFlow Labs Pte. Ltd.D46.7inference.llm inference.open-weights embed.text rerank image.generate video.generate speech.ttsno
LocalAI Ettore Di Giacinto and the LocalAI teamB68inference.open-weights embed.text rerank speech.tts image.generate video.generateno
Lemonade AMD and the Lemonade communityB63.8inference.open-weights embed.text rerank speech.tts image.generateno
Docker Model Runner Docker, Inc.C57.1inference.open-weights inference.llm embed.text rerank image.generateno

Machine-readable

Verify this listing

For the vendor

Is this your product? Link to this page from your own site or README, then tell us where. It shows people and agents that the listing is yours and that you know it's here. It never changes a grade, rank or review.

  1. Add the badge or a link

    Novita AI on Anchor Terminal, D, 53.5/100
    On a light page
    On a dark page
    <a href="https://www.anchorterminal.com/tools/novita-ai"><img src="https://www.anchorterminal.com/badges/novita-ai.svg" alt="Novita AI on Anchor Terminal" height="20"></a>
    [![Novita AI on Anchor Terminal](https://www.anchorterminal.com/badges/novita-ai.svg)](https://www.anchorterminal.com/tools/novita-ai)

    It counts on a page on novita.ai or one of its subdomains, or the README of github.com/novitalabs/novita-skills.

  2. Tell us where it is

    We read it once now and again every week. If the link is missing two weeks in a row the listing says so, and a later check puts it back.

Agents send the same to POST /api/v1/verify as {"slug": "novita-ai", "url": "…"}, or call the verify_listing tool at /mcp. Ten checks an hour from one address. What we check. To announce the listing, get sharing assets for social media.

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.