{
  "fixes": {
    "slug": "text-generation-webui",
    "name": "TextGen",
    "listing": "https://www.anchorterminal.com/tools/text-generation-webui",
    "markdown": "# Fix list: TextGen\n\nFrom Anchor Terminal's listing at https://www.anchorterminal.com/tools/text-generation-webui, the October 2026 research run, assessed 8 October 2026. Grade E, 45.1 out of 100.\n\nThis is everything the published grade says the listing lacks, the biggest possible gain to the total first. It comes from the reason given for each score, the checklist each category was scored against (https://www.anchorterminal.com/benchmark/#checklist), the provenance checks, the deductions, what we couldn't check and what the review panel asked for. A fix counts at the next check, once it's public.\n\nFor a coding agent working on TextGen: work through the items below in the product, its docs and its public pages. Each category gives the reason for its score, with the points each checklist item earned, and the checklist itself, so the gap is the items that earned less than their points. Change the product, not the wording, and keep a note of what you changed and where it's published.\n\n## 1. Security \u0026 auth, 40 out of 100, up to 10.5 more on the total\n\nWhy it scored 40: Read with the tool checklist. One optional key from `--api-key`, off by default, sent as a Bearer token or `x-api-key` and never in a query string, with a separate `--admin-key` for loading and unloading models and LoRAs. No scopes or per-client keys, and a key changes only with a restart (14 of 30). Without `--listen` or `--public-api` the API binds 127.0.0.1, rejects other Host headers and limits CORS to localhost (since v4.9). The chat has a Confirm tool calls setting, off by default, and the web UI has no login unless `--gradio-auth` is set (10 of 20). The built-in page fetcher refuses non-global addresses on every redirect hop. No guidance on injected instructions in web or tool results was found (5 of 15). The docs say the API creates no logs and the server runs with access logs off. `--verbose` prints prompts to the terminal (3 of 15). No SECURITY.md, security.txt or bounty. Ten advisories are published on GitHub with patched versions, one with a CVE (8 of 20).\n\nThe checklist (https://www.anchorterminal.com/benchmark/#checklist-security):\n\n- 0 to 30, the credential model. 30 for OAuth 2.1 with scopes, or scoped and revocable keys with rotation. 20 for plain revocable API keys. 10 for one all-powerful key. 10 off when a secret can travel in a URL query string as a documented option.\n- 0 to 20, read-only or least-privilege modes, and confirmation or approval for destructive actions.\n- 0 to 15, prompt-injection posture where the tool returns untrusted content (documented mitigations or guidance). A tool that returns no untrusted content gets 10.\n- 0 to 15, audit logs or per-call visibility for the operator.\n- 0 to 20, a security programme. security.txt or a disclosure policy, a bug bounty, SOC 2 or ISO 27001, advisories handled in public.\n\nModels are read for retention, whether API data trains models (and whether that's off by default), zero-retention options and certifications. Frameworks for telemetry defaults, approval hooks, guardrails and sandboxing.\n\n## 2. Reliability, 51 out of 100, up to 9.8 more on the total\n\nWhy it scored 51: Read with the local-software lines, since TextGen runs on the owner's machine with no hosted service. Portable builds for Linux, Windows and macOS on GitHub releases, a one-click installer, a venv route for Python 3.9 or later, a Conda route on Python 3.13 and Docker files, but no package on a registry (18 of 20). No test suite was found in the repository, and the seven workflows only build release packages on manual dispatch (3 of 25). 807 open issues and 40 open pull requests, no code commit on `main` or `dev` since 31 May 2026, and the newest open issues we saw (July to October 2026) carry between none and two comments (6 of 25). Versioned tags and GitHub release notes per version that list security fixes and behaviour changes, with no changelog file and no stated semver policy (9 of 15). v4.9 (15).\n\nThe checklist (https://www.anchorterminal.com/benchmark/#checklist-reliability):\n\nHosted APIs, MCP servers, models and platforms.\n\n- 20, a public status page with component history (Statuspage, Instatus, BetterStack or the vendor's own).\n- 0 to 30, the incident record for the last 90 days on that page. 30 for a clean record or trivial incidents only, 20 for minor incidents only, 10 for one major outage (an hour or more of a core API down, or errors across the board), 0 for several. 5 when there's no history we could read, and the note says so.\n- 15, rate limits documented with numbers.\n- 15, documented 429 or overload handling (Retry-After, backoff guidance), and idempotency keys or safe-retry guidance where writes are involved.\n- 10, an SLA published for any paid tier.\n- 10, the surface agents use is generally available, not beta or preview.\n\nLocal packages, SDKs, frameworks and stdio MCP servers.\n\n- 20, installs from an official package with supported runtimes stated.\n- 25, a public CI and test suite, passing on the default branch.\n- 0 to 25, open crash or regression issues relative to activity (25 for few and handled, 0 for many, old and unanswered).\n- 15, semver discipline and breaking changes called out in a changelog.\n- 15, version 1.0 or later, or declared stable.\n\nProtocols are read from their reference implementations, the public facilitators or servers, spec stability and test vectors.\n\n## 3. Agent ergonomics, 50 out of 100, up to 8.1 more on the total\n\nWhy it scored 50: Read with the API lines. `max_tokens` (512 by default on chat completions), `max_tokens_second`, a token-count route and a `max_tokens` argument on the built-in page fetcher size the output (17 of 25). `/v1/models` lists everything with the loaded model first, with no paging or filters (6 of 20). Errors are JSON in OpenAI's shape (`message`, `type`, `param`) or Anthropic's on `/v1/messages`, with 401, 400 and 422 used, but none are documented (11 of 20). Completions are stateless and safe to repeat, a client disconnect stops generation and `/v1/internal/stop-generation` exists. No retry guidance (7 of 20). A request needs only `messages`, with no model name, but the owner has to launch with `--api` and load a model. No SDK of its own, with OpenAI's Python and Node clients shown in the docs (9 of 15).\n\nThe checklist (https://www.anchorterminal.com/benchmark/#checklist-ergonomics):\n\n- 0 to 25, context cost. For MCP, the number and size of the tool definitions (25 for ten or fewer compact tools, 15 for 11 to 30, 5 for more than 30, plus up to 10 back for toolsets, dynamic loading or read-only subsets). For APIs, whether responses can be sized (field selection, limits, summaries).\n- 20, pagination, filtering and output-size controls.\n- 20, actionable, documented error responses, codes and messages an agent can recover from.\n- 20, idempotency or safe retries, and for MCP the `readOnlyHint` and `destructiveHint` annotations.\n- 15, sensible defaults, few required parameters, and official SDKs in at least two languages.\n\nModels are read for tool use, structured output, prompt caching, context length, batch and SDKs. Frameworks for how much code and how many defaults a tool-calling agent with MCP needs.\n\n## 4. Maintenance \u0026 community, 24 out of 100, up to 6.7 more on the total\n\nWhy it scored 24: v4.9 on 20 May 2026, 141 days before this check (10). No release in the last 90 days (0). No code commit on `main` since 31 May 2026 (one README edit on 17 August) and none on `dev` since 16 May. 807 open issues and 40 open pull requests, and the newest open issues show few or no comments. Comment authors didn't render for our reader, so maintainer replies are unconfirmed (5 of 25). No SDK and no MCP registry entry. Portable builds shipped with each release up to May (5 of 15). A weekly Dependabot config with its update branches unmerged, no test CI, and the bundled llama.cpp last updated on 31 May (4 of 10).\n\nThe checklist (https://www.anchorterminal.com/benchmark/#checklist-maintenance):\n\n- 0 to 30, time since the last release, or the last published model or API change for a closed service. 30 within 30 days, 20 within 90, 10 within 180, 0 older.\n- 20, at least three releases or dated changelog entries in the last 90 days.\n- 0 to 25, responsiveness. Issues and pull requests answered on GitHub (the open issues and how recent the replies are). For closed services, a public changelog and a support or community channel that answers, 0 to 15.\n- 15, presence in the official MCP registry under a verified namespace (MCP servers), or current official SDKs (APIs and models).\n- 10, package health, current dependencies and CI.\n\nModels are read for deprecation notice periods and model churn rather than release counts.\n\n## 5. Schema \u0026 documentation, 61 out of 100, up to 6.3 more on the total\n\nWhy it scored 61: Read with the API lines. The server is a FastAPI app with default settings, so a running instance serves `/docs` and `/openapi.json` generated from 33 Pydantic classes in `modules/api/typing.py`, and the docs point callers there. No OpenAPI file is published, and we read the source without running the server (17 of 25). No llms.txt. The docs are Markdown files in the repository's docs folder, mirrored to the wiki (4 of 10). The API page is mostly examples with a compatibility table, and 36 fields in `typing.py` carry descriptions (10 of 20). Typed request models with defaults and a few range limits, while `tools` is a list of free-form objects (10 of 15). curl, Python and Node examples for chat, completions, streaming, tool calling, vision, images and model loading, with no error responses documented (9 of 15). A `/v1` prefix and release notes per version, with no changelog file (11 of 15).\n\nThe checklist (https://www.anchorterminal.com/benchmark/#checklist-schema):\n\nAPIs and MCP servers.\n\n- 25, a machine-readable contract (a public OpenAPI file or similar; for MCP, typed JSON Schema inputs on every tool).\n- 10, llms.txt or Markdown docs served for agents.\n- 0 to 20, descriptions that say what a tool is for, when to use it and when not to, read from the tool definitions in the source or the API reference.\n- 0 to 15, typed inputs with enums, constraints and required fields, and no free-form JSON blobs.\n- 0 to 15, examples and documented error responses.\n- 15, versioning and a public changelog.\n\nModels are read from the API reference, the OpenAPI file, llms.txt, the structured-output and tool-use docs and the model cards. Frameworks from docs a model can follow, typed interfaces, examples and the API reference.\n\n## 6. Payments \u0026 pricing, 60 out of 100, up to 5 more on the total\n\nWhy it scored 60: Self-hosted rule. No x402, MPP or L402 (0). Free and AGPL-3.0 with no account, card or paid edition, so 20, 20 and 20 on the last three lines.\n\nThe checklist (https://www.anchorterminal.com/benchmark/#checklist-payments):\n\nThe published rubric, also on the [x402 page](https://www.anchorterminal.com/x402/).\n\n- 40, a machine payment protocol (x402, MPP or L402) on the tool's own endpoints. 10 to 30 when it covers only some endpoints or only goes through a third party, and the note says which.\n- 20, per-call or per-unit pricing published without a login. 10 for public plan-only pricing, 0 for \"contact sales\" or prices behind a login.\n- 20, a free tier or trial that doesn't need a card.\n- 20, autonomous onboarding, meaning an agent can get access without a person signing up in a browser (keyless use, x402, a programmatic key API).\n\nPayment platforms and agent wallets rarely charge for their own API over a machine protocol, so the first line has steps for them, and the highest one that applies counts. 40 when x402, MPP or L402 runs on all their own endpoints, 30 when it runs on part of their own API, 25 when their merchants can accept one, 20 for running a facilitator, 15 for paying as a buyer, and 0 when the only protocol is their own. Merchant acceptance sits above a facilitator because the platform's own customers can charge agents through it, while a facilitator settles for sellers who wire up the protocol themselves. The counter-argument (a facilitator does more for the protocol as a whole) has a point. Each note says which step applied.\n\nOpen-source software you run yourself is scored on its hosted or paid option if it has one. A free, self-hosted package with nothing to buy gets 20, 20 and 20 for the last three lines, and 0 to 40 for the first only if it ships a payment protocol.\n\n## 7. Transparency \u0026 trust, 49 out of 100, up to 4.5 more on the total\n\nMade of editorial 70, provenance 27.\n\nWhy it scored 49: AGPL-3.0 for the whole repository (30). No privacy policy or terms exist, and there's no hosted service. The README states that the app is fully offline with zero telemetry, external resources or remote update requests, the API page says it creates no logs, and the source turns Gradio analytics off. A Check for updates button calls api.github.com when clicked, which the README's wording doesn't mention (18 of 30). No deprecation policy. Release notes record behaviour changes, and the rename to TextGen left a redirect from the old repository URL (4 of 20). No telemetry to opt out of, and analytics are disabled in `server.py` (18 of 20).\n\nThe checklist (https://www.anchorterminal.com/benchmark/#checklist-transparency):\n\n- 0 to 30, source availability and licence clarity. 30 for open source under an OSI licence, 15 for closed with clear terms, 0 for unclear terms.\n- 0 to 30, data handling and retention statements that agree with each other (privacy policy, DPA, retention periods, subprocessors).\n- 0 to 20, a deprecation policy or notices with dates.\n- 0 to 20, telemetry disclosed with an opt-out (local software), or subprocessors and data locations disclosed (hosted).\n\nThe other half of Transparency and trust is the provenance score, computed from checked facts (below). The category score is the mean of the two.\n\nProvenance checks not met in full (half of this category, computed from checked facts):\n\n- Legal entity named: not found (0 of 20)\n- Domain age: github.com/oobabooga, no registry record we could read (0 of 15)\n- Status page: not found (0 of 10)\n- security.txt: not found (0 of 10)\n\n## Deductions\n\nEach comes off the total. A fixed and documented problem counts for less at the next check.\n\n- 2026-03-18. Two Critical advisories at 9.1. GHSA-jg96-p5p6-q3cv (CVE-2026-35050) let a user of the web UI write Python files through a path traversal in extension settings and run them, in 4.1 and earlier, fixed in 4.1.1. GHSA-4p45-76cc-7p62 let a caller write or delete files through a character name on instances started with `--listen`, fixed in 4.0. Both fixed and published, -2. https://github.com/oobabooga/textgen/security/advisories/GHSA-jg96-p5p6-q3cv\n- 2026-03-18. GHSA-fpwc-4mvr-7jpr (High, 7.1). `/v1/chat/completions` and `/v1/completions` fetched any `image_url` from the server side, so a caller could reach internal addresses, in 4.1 and earlier. Fixed in 4.1.1 and published, -1. https://github.com/oobabooga/textgen/security/advisories/GHSA-fpwc-4mvr-7jpr\n- 2025-10-13 to 2026-04-03. Seven more advisories in the year, four High and three Moderate, among them a file read through an uploaded symbolic link, four path traversals in preset, grammar, template and prompt loading, an SSRF in the superbooga extensions and a Gradio path-check bypass fixed in 4.3. All published with fixes, -1. https://github.com/oobabooga/textgen/security\n\n## What we couldn't check\n\nWhat we couldn't read counted as absent. Publishing it on a page a plain HTTP fetch can read (not only in a browser) lets the next check count it.\n\n- Whether development has paused. No release since 20 May 2026 and no code commit since 31 May, with no notice in the README\n- unchecked: maintainer replies on recent issues. Comment threads didn't render for our reader\n- unchecked: `/docs` and `/openapi.json` on a running server. We read the FastAPI source and the docs and didn't run it\n- unchecked: star, issue and pull request counts come from the repository page as our reader saw it. GitHub's API refused us for a shared rate limit\n- No terms of use or privacy policy exist for the project, so `provenance.terms` and `provenance.privacy` are left out\n- The developer publishes under a pseudonym and no legal entity is named\n- The first release date wasn't established, since the clone was shallow\n\n## Weaknesses\n\n- No release since v4.9 on 20 May 2026 and no code commit on `main` or `dev` since 31 May 2026\n- No test suite. The seven GitHub workflows only build release packages on manual dispatch\n- 807 open issues and 40 open pull requests, with no SECURITY.md or disclosure address\n- Two Critical advisories (9.1) published on 18 March 2026, both path traversals in the web UI, fixed in 4.0 and 4.1.1\n- The API key is off by default, set only at launch, and has no scopes or per-client keys\n- No rate limits, error list or published OpenAPI file in the docs. The schema is served only by a running server at `/docs`\n\n## What costs an agent a turn today\n\nThe notes we give agents before they call it. Each one is a workaround an agent shouldn't need.\n\n- Ask the owner to launch with `--api`. Nothing listens on port 5000 without it, and a model must be loaded first\n- Call `http://127.0.0.1:5000/v1`. Any Host header other than localhost or 127.0.0.1 gets 400 `Invalid host header` unless `--listen` is set\n- Send the key as `Authorization: Bearer` on OpenAI routes and as `x-api-key` on `/v1/messages`. Model loading needs the admin key\n- Read `http://127.0.0.1:5000/docs` or `modules/api/typing.py` for parameters. `max_tokens` defaults to 512 on chat completions\n- Run tool calls yourself. The API returns `finish_reason: \"tool_calls\"` and executes nothing on the server\n- Use the repository name `oobabooga/textgen`. The old `text-generation-webui` URL redirects\n\n## When it's done\n\nSend what changed and where it's published as a dispute (https://www.anchorterminal.com/builders/#disputes, or `POST https://www.anchorterminal.com/api/v1/contact` with `\"kind\": \"dispute\"`). Disputes are answered in public, and the listing is checked again by the same checklist. Paying for an audit or a listing claim changes nothing here.\n",
    "grade": "E",
    "score": 45.1,
    "assessed": "2026-10-08",
    "run": "October 2026 research run",
    "categories": [
      {
        "key": "security",
        "name": "Security \u0026 auth",
        "score": 40,
        "maxGain": 10.5,
        "reason": "Read with the tool checklist. One optional key from `--api-key`, off by default, sent as a Bearer token or `x-api-key` and never in a query string, with a separate `--admin-key` for loading and unloading models and LoRAs. No scopes or per-client keys, and a key changes only with a restart (14 of 30). Without `--listen` or `--public-api` the API binds 127.0.0.1, rejects other Host headers and limits CORS to localhost (since v4.9). The chat has a Confirm tool calls setting, off by default, and the web UI has no login unless `--gradio-auth` is set (10 of 20). The built-in page fetcher refuses non-global addresses on every redirect hop. No guidance on injected instructions in web or tool results was found (5 of 15). The docs say the API creates no logs and the server runs with access logs off. `--verbose` prints prompts to the terminal (3 of 15). No SECURITY.md, security.txt or bounty. Ten advisories are published on GitHub with patched versions, one with a CVE (8 of 20).",
        "checklist": [
          "- 0 to 30, the credential model. 30 for OAuth 2.1 with scopes, or scoped and revocable keys with rotation. 20 for plain revocable API keys. 10 for one all-powerful key. 10 off when a secret can travel in a URL query string as a documented option.\n- 0 to 20, read-only or least-privilege modes, and confirmation or approval for destructive actions.\n- 0 to 15, prompt-injection posture where the tool returns untrusted content (documented mitigations or guidance). A tool that returns no untrusted content gets 10.\n- 0 to 15, audit logs or per-call visibility for the operator.\n- 0 to 20, a security programme. security.txt or a disclosure policy, a bug bounty, SOC 2 or ISO 27001, advisories handled in public.",
          "Models are read for retention, whether API data trains models (and whether that's off by default), zero-retention options and certifications. Frameworks for telemetry defaults, approval hooks, guardrails and sandboxing."
        ],
        "checklistUrl": "https://www.anchorterminal.com/benchmark/#checklist-security"
      },
      {
        "key": "reliability",
        "name": "Reliability",
        "score": 51,
        "maxGain": 9.8,
        "reason": "Read with the local-software lines, since TextGen runs on the owner's machine with no hosted service. Portable builds for Linux, Windows and macOS on GitHub releases, a one-click installer, a venv route for Python 3.9 or later, a Conda route on Python 3.13 and Docker files, but no package on a registry (18 of 20). No test suite was found in the repository, and the seven workflows only build release packages on manual dispatch (3 of 25). 807 open issues and 40 open pull requests, no code commit on `main` or `dev` since 31 May 2026, and the newest open issues we saw (July to October 2026) carry between none and two comments (6 of 25). Versioned tags and GitHub release notes per version that list security fixes and behaviour changes, with no changelog file and no stated semver policy (9 of 15). v4.9 (15).",
        "checklist": [
          "Hosted APIs, MCP servers, models and platforms.",
          "- 20, a public status page with component history (Statuspage, Instatus, BetterStack or the vendor's own).\n- 0 to 30, the incident record for the last 90 days on that page. 30 for a clean record or trivial incidents only, 20 for minor incidents only, 10 for one major outage (an hour or more of a core API down, or errors across the board), 0 for several. 5 when there's no history we could read, and the note says so.\n- 15, rate limits documented with numbers.\n- 15, documented 429 or overload handling (Retry-After, backoff guidance), and idempotency keys or safe-retry guidance where writes are involved.\n- 10, an SLA published for any paid tier.\n- 10, the surface agents use is generally available, not beta or preview.",
          "Local packages, SDKs, frameworks and stdio MCP servers.",
          "- 20, installs from an official package with supported runtimes stated.\n- 25, a public CI and test suite, passing on the default branch.\n- 0 to 25, open crash or regression issues relative to activity (25 for few and handled, 0 for many, old and unanswered).\n- 15, semver discipline and breaking changes called out in a changelog.\n- 15, version 1.0 or later, or declared stable.",
          "Protocols are read from their reference implementations, the public facilitators or servers, spec stability and test vectors."
        ],
        "checklistUrl": "https://www.anchorterminal.com/benchmark/#checklist-reliability"
      },
      {
        "key": "ergonomics",
        "name": "Agent ergonomics",
        "score": 50,
        "maxGain": 8.1,
        "reason": "Read with the API lines. `max_tokens` (512 by default on chat completions), `max_tokens_second`, a token-count route and a `max_tokens` argument on the built-in page fetcher size the output (17 of 25). `/v1/models` lists everything with the loaded model first, with no paging or filters (6 of 20). Errors are JSON in OpenAI's shape (`message`, `type`, `param`) or Anthropic's on `/v1/messages`, with 401, 400 and 422 used, but none are documented (11 of 20). Completions are stateless and safe to repeat, a client disconnect stops generation and `/v1/internal/stop-generation` exists. No retry guidance (7 of 20). A request needs only `messages`, with no model name, but the owner has to launch with `--api` and load a model. No SDK of its own, with OpenAI's Python and Node clients shown in the docs (9 of 15).",
        "checklist": [
          "- 0 to 25, context cost. For MCP, the number and size of the tool definitions (25 for ten or fewer compact tools, 15 for 11 to 30, 5 for more than 30, plus up to 10 back for toolsets, dynamic loading or read-only subsets). For APIs, whether responses can be sized (field selection, limits, summaries).\n- 20, pagination, filtering and output-size controls.\n- 20, actionable, documented error responses, codes and messages an agent can recover from.\n- 20, idempotency or safe retries, and for MCP the `readOnlyHint` and `destructiveHint` annotations.\n- 15, sensible defaults, few required parameters, and official SDKs in at least two languages.",
          "Models are read for tool use, structured output, prompt caching, context length, batch and SDKs. Frameworks for how much code and how many defaults a tool-calling agent with MCP needs."
        ],
        "checklistUrl": "https://www.anchorterminal.com/benchmark/#checklist-ergonomics"
      },
      {
        "key": "maintenance",
        "name": "Maintenance \u0026 community",
        "score": 24,
        "maxGain": 6.7,
        "reason": "v4.9 on 20 May 2026, 141 days before this check (10). No release in the last 90 days (0). No code commit on `main` since 31 May 2026 (one README edit on 17 August) and none on `dev` since 16 May. 807 open issues and 40 open pull requests, and the newest open issues show few or no comments. Comment authors didn't render for our reader, so maintainer replies are unconfirmed (5 of 25). No SDK and no MCP registry entry. Portable builds shipped with each release up to May (5 of 15). A weekly Dependabot config with its update branches unmerged, no test CI, and the bundled llama.cpp last updated on 31 May (4 of 10).",
        "checklist": [
          "- 0 to 30, time since the last release, or the last published model or API change for a closed service. 30 within 30 days, 20 within 90, 10 within 180, 0 older.\n- 20, at least three releases or dated changelog entries in the last 90 days.\n- 0 to 25, responsiveness. Issues and pull requests answered on GitHub (the open issues and how recent the replies are). For closed services, a public changelog and a support or community channel that answers, 0 to 15.\n- 15, presence in the official MCP registry under a verified namespace (MCP servers), or current official SDKs (APIs and models).\n- 10, package health, current dependencies and CI.",
          "Models are read for deprecation notice periods and model churn rather than release counts."
        ],
        "checklistUrl": "https://www.anchorterminal.com/benchmark/#checklist-maintenance"
      },
      {
        "key": "schema",
        "name": "Schema \u0026 documentation",
        "score": 61,
        "maxGain": 6.3,
        "reason": "Read with the API lines. The server is a FastAPI app with default settings, so a running instance serves `/docs` and `/openapi.json` generated from 33 Pydantic classes in `modules/api/typing.py`, and the docs point callers there. No OpenAPI file is published, and we read the source without running the server (17 of 25). No llms.txt. The docs are Markdown files in the repository's docs folder, mirrored to the wiki (4 of 10). The API page is mostly examples with a compatibility table, and 36 fields in `typing.py` carry descriptions (10 of 20). Typed request models with defaults and a few range limits, while `tools` is a list of free-form objects (10 of 15). curl, Python and Node examples for chat, completions, streaming, tool calling, vision, images and model loading, with no error responses documented (9 of 15). A `/v1` prefix and release notes per version, with no changelog file (11 of 15).",
        "checklist": [
          "APIs and MCP servers.",
          "- 25, a machine-readable contract (a public OpenAPI file or similar; for MCP, typed JSON Schema inputs on every tool).\n- 10, llms.txt or Markdown docs served for agents.\n- 0 to 20, descriptions that say what a tool is for, when to use it and when not to, read from the tool definitions in the source or the API reference.\n- 0 to 15, typed inputs with enums, constraints and required fields, and no free-form JSON blobs.\n- 0 to 15, examples and documented error responses.\n- 15, versioning and a public changelog.",
          "Models are read from the API reference, the OpenAPI file, llms.txt, the structured-output and tool-use docs and the model cards. Frameworks from docs a model can follow, typed interfaces, examples and the API reference."
        ],
        "checklistUrl": "https://www.anchorterminal.com/benchmark/#checklist-schema"
      },
      {
        "key": "payments",
        "name": "Payments \u0026 pricing",
        "score": 60,
        "maxGain": 5,
        "reason": "Self-hosted rule. No x402, MPP or L402 (0). Free and AGPL-3.0 with no account, card or paid edition, so 20, 20 and 20 on the last three lines.",
        "checklist": [
          "The published rubric, also on the [x402 page](https://www.anchorterminal.com/x402/).",
          "- 40, a machine payment protocol (x402, MPP or L402) on the tool's own endpoints. 10 to 30 when it covers only some endpoints or only goes through a third party, and the note says which.\n- 20, per-call or per-unit pricing published without a login. 10 for public plan-only pricing, 0 for \"contact sales\" or prices behind a login.\n- 20, a free tier or trial that doesn't need a card.\n- 20, autonomous onboarding, meaning an agent can get access without a person signing up in a browser (keyless use, x402, a programmatic key API).",
          "Payment platforms and agent wallets rarely charge for their own API over a machine protocol, so the first line has steps for them, and the highest one that applies counts. 40 when x402, MPP or L402 runs on all their own endpoints, 30 when it runs on part of their own API, 25 when their merchants can accept one, 20 for running a facilitator, 15 for paying as a buyer, and 0 when the only protocol is their own. Merchant acceptance sits above a facilitator because the platform's own customers can charge agents through it, while a facilitator settles for sellers who wire up the protocol themselves. The counter-argument (a facilitator does more for the protocol as a whole) has a point. Each note says which step applied.",
          "Open-source software you run yourself is scored on its hosted or paid option if it has one. A free, self-hosted package with nothing to buy gets 20, 20 and 20 for the last three lines, and 0 to 40 for the first only if it ships a payment protocol."
        ],
        "checklistUrl": "https://www.anchorterminal.com/benchmark/#checklist-payments"
      },
      {
        "key": "transparency",
        "name": "Transparency \u0026 trust",
        "score": 49,
        "maxGain": 4.5,
        "reason": "AGPL-3.0 for the whole repository (30). No privacy policy or terms exist, and there's no hosted service. The README states that the app is fully offline with zero telemetry, external resources or remote update requests, the API page says it creates no logs, and the source turns Gradio analytics off. A Check for updates button calls api.github.com when clicked, which the README's wording doesn't mention (18 of 30). No deprecation policy. Release notes record behaviour changes, and the rename to TextGen left a redirect from the old repository URL (4 of 20). No telemetry to opt out of, and analytics are disabled in `server.py` (18 of 20).",
        "blend": "editorial 70, provenance 27",
        "checklist": [
          "- 0 to 30, source availability and licence clarity. 30 for open source under an OSI licence, 15 for closed with clear terms, 0 for unclear terms.\n- 0 to 30, data handling and retention statements that agree with each other (privacy policy, DPA, retention periods, subprocessors).\n- 0 to 20, a deprecation policy or notices with dates.\n- 0 to 20, telemetry disclosed with an opt-out (local software), or subprocessors and data locations disclosed (hosted).",
          "The other half of Transparency and trust is the provenance score, computed from checked facts (below). The category score is the mean of the two."
        ],
        "checklistUrl": "https://www.anchorterminal.com/benchmark/#checklist-transparency"
      }
    ],
    "provenance": [
      {
        "label": "Legal entity named",
        "value": "not found",
        "points": 0,
        "max": 20
      },
      {
        "label": "Domain age",
        "value": "github.com/oobabooga, no registry record we could read",
        "points": 0,
        "max": 15
      },
      {
        "label": "Status page",
        "value": "not found",
        "points": 0,
        "max": 10
      },
      {
        "label": "security.txt",
        "value": "not found",
        "points": 0,
        "max": 10
      }
    ],
    "deductions": [
      "2026-03-18. Two Critical advisories at 9.1. GHSA-jg96-p5p6-q3cv (CVE-2026-35050) let a user of the web UI write Python files through a path traversal in extension settings and run them, in 4.1 and earlier, fixed in 4.1.1. GHSA-4p45-76cc-7p62 let a caller write or delete files through a character name on instances started with `--listen`, fixed in 4.0. Both fixed and published, -2. https://github.com/oobabooga/textgen/security/advisories/GHSA-jg96-p5p6-q3cv",
      "2026-03-18. GHSA-fpwc-4mvr-7jpr (High, 7.1). `/v1/chat/completions` and `/v1/completions` fetched any `image_url` from the server side, so a caller could reach internal addresses, in 4.1 and earlier. Fixed in 4.1.1 and published, -1. https://github.com/oobabooga/textgen/security/advisories/GHSA-fpwc-4mvr-7jpr",
      "2025-10-13 to 2026-04-03. Seven more advisories in the year, four High and three Moderate, among them a file read through an uploaded symbolic link, four path traversals in preset, grammar, template and prompt loading, an SSRF in the superbooga extensions and a Gradio path-check bypass fixed in 4.3. All published with fixes, -1. https://github.com/oobabooga/textgen/security"
    ],
    "unchecked": [
      "Whether development has paused. No release since 20 May 2026 and no code commit since 31 May, with no notice in the README",
      "unchecked: maintainer replies on recent issues. Comment threads didn't render for our reader",
      "unchecked: `/docs` and `/openapi.json` on a running server. We read the FastAPI source and the docs and didn't run it",
      "unchecked: star, issue and pull request counts come from the repository page as our reader saw it. GitHub's API refused us for a shared rate limit",
      "No terms of use or privacy policy exist for the project, so `provenance.terms` and `provenance.privacy` are left out",
      "The developer publishes under a pseudonym and no legal entity is named",
      "The first release date wasn't established, since the clone was shallow"
    ],
    "weaknesses": [
      "No release since v4.9 on 20 May 2026 and no code commit on `main` or `dev` since 31 May 2026",
      "No test suite. The seven GitHub workflows only build release packages on manual dispatch",
      "807 open issues and 40 open pull requests, with no SECURITY.md or disclosure address",
      "Two Critical advisories (9.1) published on 18 March 2026, both path traversals in the web UI, fixed in 4.0 and 4.1.1",
      "The API key is off by default, set only at launch, and has no scopes or per-client keys",
      "No rate limits, error list or published OpenAPI file in the docs. The schema is served only by a running server at `/docs`"
    ],
    "agentNotes": [
      "Ask the owner to launch with `--api`. Nothing listens on port 5000 without it, and a model must be loaded first",
      "Call `http://127.0.0.1:5000/v1`. Any Host header other than localhost or 127.0.0.1 gets 400 `Invalid host header` unless `--listen` is set",
      "Send the key as `Authorization: Bearer` on OpenAI routes and as `x-api-key` on `/v1/messages`. Model loading needs the admin key",
      "Read `http://127.0.0.1:5000/docs` or `modules/api/typing.py` for parameters. `max_tokens` defaults to 512 on chat completions",
      "Run tool calls yourself. The API returns `finish_reason: \"tool_calls\"` and executes nothing on the server",
      "Use the repository name `oobabooga/textgen`. The old `text-generation-webui` URL redirects"
    ],
    "recheck": "https://www.anchorterminal.com/builders/#disputes"
  },
  "meta": {
    "attribution": "Anchor Terminal (https://www.anchorterminal.com)",
    "docs": "https://www.anchorterminal.com/docs/",
    "generatedAt": "2026-10-09",
    "license": "CC-BY-4.0",
    "method": "https://www.anchorterminal.com/benchmark/",
    "methodology": "0.4",
    "openapi": "https://www.anchorterminal.com/openapi.json",
    "preview": false,
    "run": "2026-10-01",
    "runLabel": "October 2026 research run"
  }
}
