Together Code Sandbox

by Together AI HTTP API in Code execution sandboxes

Hosted

Together Computer, Inc. · together.ai since 2017 · status page · who's behind it

Together AI's hosted virtual machine sandboxes for running commands and code, built from Docker-image snapshots and driven from the together-sandbox Python SDK, TypeScript SDK or CLI. Access is by allowlist, on request to Together.

Good for Teams already buying from Together, and former CodeSandbox SDK users, who want Docker-defined sandboxes with disk and memory snapshots from Python or TypeScript.

Is this your product? Claim this listing or verify it

More from Together AI Together AI Fine-tuning (Fine-tuning)

Assessment. Two public OpenAPI documents, MIT clients for Python and TypeScript, cursor pagination and built-in retries make the surface easy for an agent to drive. Access is the limit. Together enables the SDK per organisation on request, publishes no rate limits or SLA, and neither status page names the sandbox service.

Facts

Transport
HTTP
Endpoint
https://api.bartender.codesandbox.io
Auth
API key
Pricing
Pay per use · $0.0446 / vCPU-hr
x402
No
Licence
Proprietary service under Together's terms of service. The SDKs, CLI and OpenAPI documents on GitHub are MIT
Packages
pypi together-sandbox
npm together-sandbox
llms.txt
published
Last release
GitHub stars
3
npm / week
1.3k
PyPI / week
3.4k
Access
Allowlist. Together enables the SDK and CLI per organisation on request
Clients
together-sandbox 4.1.1 on PyPI (async, Python 3.10 or later) and npm (Node.js 18 or later), and a together-sandbox CLI installed as a self-contained binary, all MIT
APIs
Management API (OpenAPI 3.1.0, 17 operations, /v1 at api.bartender.codesandbox.io) for sandboxes, snapshots and aliases. In-VM API (OpenAPI 3.0.3, 25 operations) for files, directories, execs and ports
Sizes
Default 1 vCPU and 2 GiB. cpu from 0.1 to 16 cores and 1 to 8 GB of memory per core, fixed at creation
Lifetime
Runs until terminated unless ttl (seconds from creation) is set. No idle detection, and a terminated sandbox can't restart
Persistence
Ephemeral by default. A termination policy or terminate(snapshot=...) saves disk, or disk and memory, as a snapshot aliased sandbox:<id>
Images
Snapshots built by Together's remote image builder from a public Docker image or a Dockerfile context. No default image
Network
All ports public by default. Experimental inbound allowlist by port, IP or CIDR with an optional X-Sandbox-Token requirement. No outbound rules
Retries
SDKs retry 408, 429, 500, 502, 503 and 504 three times with exponential backoff from 0.5 seconds. snapshots.create is not idempotent
Rate limits
None found. The migration guide says the concurrency count and limit are not exposed
Status
No sandbox component on status.together.ai or status.codesandbox.io

Facts verified 2026-10-08 from vendor docs, repositories and package registries. JSON · Markdown

Strengths

  • Public OpenAPI documents for the management API (17 operations) and the in-VM API (25), in the repository and on GitHub Pages
  • Terminating with memory: true saves disk and memory, and a new sandbox created from that snapshot resumes its running processes
  • List calls are cursor-paginated (1 to 100 a page) with filters for status, source snapshot and tags
  • Both SDKs retry 408, 429 and 5xx three times with exponential backoff, and the docs warn that snapshots.create is not idempotent
  • Releases on 30 September, 1 October and 7 October 2026, with a generated changelog that marks breaking changes

Weaknesses

  • The SDK and CLI work only for organisations Together has enabled, and the docs say to contact Together for access
  • No rate limits, concurrency limits or SLA found, and neither status.together.ai nor status.codesandbox.io names the service
  • Every sandbox port is public unless an experimental inbound policy is set, and no outbound controls exist yet
  • Three major versions between 2 June and 28 July 2026, and memory snapshots were removed in 4.0.0 and restored in 4.0.4
  • The management API answers at api.bartender.codesandbox.io, off Together's domain, and the privacy policy's data table has no sandbox row

Before you call it notes for agents

  1. Confirm the organisation is on Together's allowlist before installing. A valid TOGETHER_API_KEY alone is not enough
  2. Build a snapshot first with snapshots.create. No default image exists, and every sandbox needs snapshot_id or snapshot_alias
  3. Set ttl at creation or call terminate(). Nothing stops a sandbox otherwise, and closing the Python client leaves it running
  4. Store the snapshot alias, not the sandbox ID. A terminated sandbox can't restart, and its state lives at sandbox:<id>
  5. Exclude snapshots.create from retries with should_retry, and set experimental.network_policy before serving anything private on a port

Who's behind it provenance 77/100

  • Legal entity namedTogether Computer, Inc.20/20
  • Domain agetogether.ai, registered 2017-12-16 (8 years)11/15
  • Endpoint on the vendor's domainapi.bartender.codesandbox.io is not on together.ai0/15
  • Terms of serviceread, states 5 of the 7 things a reader expects, and has 1 clause that costs points6.3/10
  • Privacy policyread, states 7 of the 8 things a reader expects9.3/10
  • Status pagestatus.together.ai10/10
  • Changelogpublished10/10
  • security.txtvalid10/10

Terms and privacy, as read

Terms of service gives no date, states 5 of 7, 1 to know

TL;DR Gives no date. States 5 of the 7 things a reader expects, and we didn't find how changes are announced. To know before relying on it, limits on benchmarking.

Restricts benchmarking or competitive usecosts points
(c) use or access the Services to develop a product or service that is competitive with the Company’s products or services or engage in competitive analysis or benchmarking;

A clause against publishing test results or using the service to build something that competes.

Gives the date it was last updated

Not found in the text.

Without a date nobody can tell which version they agreed to.

Names the governing law or courts The law of the State of California
This Agreement will be governed by the laws of the State of California, exclusive of its rules governing choice of law and conflict of laws.

Says where a dispute would be heard and under whose law.

States a limit on its liability Capped at the fees paid in the 12 months before the claim
…even if informed of their possibility in advance, or (b) excluding customer's payment obligations, any aggregate liability in excess of the amounts paid by customer during the twelve (12) months preceding the claim (the "ordinary cap").

Says the most the vendor would owe if the service causes a loss.

Says how the agreement or account can be ended
The Company may suspend your access to the Services immediately upon notice if you fail to pay any amounts hereunder at least five (5) days past the applicable due date.

Says when the vendor can cut off access and what notice it gives.

Says how changes to the terms are announced

Not found in the text.

Says whether a customer hears about a change before it binds them.

Lists what users may not do
You will not use the Services to transmit or provide to the Company any financial or medical information of any nature or any sensitive personal data (e.g., social security numbers, driver’s license numbers, birth dates, personal bank account numbers, passport or visa numbers, and/or credit card numbers).

The acceptable-use rules an agent acting for a user has to stay inside.

Refers to a service level or uptime commitment
No guarantees are made with respect to the Services’ quality, stability, uptime, or reliability, unless otherwise agreed between the parties in an order form.

Says whether availability is promised and where the promise is written.

The parties owe each other no confidentiality unless they agree it in writing.
The parties will have no confidentiality obligations to each other unless otherwise agreed in writing.

Noted by a second reader on 2026-10-08.

The customer may not cancel or end the agreement without Together's express written consent.
You may not cancel or terminate this Agreement without our express written consent.

Noted by a second reader on 2026-10-08.

Customers may not send financial or medical information or sensitive personal data to Together through the services.
You will not use the Services to transmit or provide to the Company any financial or medical information of any nature or any sensitive personal data

Noted by a second reader on 2026-10-08.

The document · read 2026-10-08 · 3,964 words

Privacy policy gives no date, states 7 of 8

TL;DR Gives no date. States 7 of the 8 things a reader expects. The rules found no clause to flag.

Gives the date it was last updated

Not found in the text.

Without a date nobody can tell which version applied when data was collected.

Says what personal data is collected
Data that we collect, including your Personal Data, will not be used to train the Company’s models without your explicit opt-in and consent.

The basic statement a privacy policy exists to make.

Says how long data is kept For as long as needed, with no period named
The Company will retain your Personal Data only for as long as is necessary for the purposes set out in this Policy.

Says when data sent to the service is deleted.

Says who else receives the data
We may share your Personal Data with Service Providers, Third-Party Vendors, Consultants, and other Business Partners in order to provide Services on our behalf, monitor and analyze the use of our Services, contact you, and for the reasons stated in our Terms of Service.

Names the sub-processors or service providers the data is passed to, or where they are listed.

Says whether personal data is sold or shared for advertising Says it does not sell personal data
The Company does not sell your Personal Data as defined under CCPA.

A plain statement either way.

Says what rights people have over their data
If, in the future, we do sell your Personal Data, we will notify you, and you may have the right to opt out of such sale.

Access, correction, deletion and objection, and how to use them.

Gives a privacy contact privacy@together.ai
In order to exercise any other privacy rights not specifically enumerated here, you may contact privacy@together.ai.

An address or officer to send a request to.

Says where data is transferred or stored Relies on standard contractual clauses
When we share Personal Data of individuals in the EEA, Switzerland, or UK with third parties, we make use of a variety of legal mechanisms to safeguard the transfer, including the European Commission-approved standard contractual clauses as well as additional safeguards where appropriate.

The countries data goes to and the safeguard used.

Data collected from a customer is not used to train Together's models unless the customer opts in.
We do not use any data collected from you to train our models without your explicit opt-in and consent.

Noted by a second reader on 2026-10-08.

With Zero Data Retention on, Together cannot later retrieve, correct, export or delete the data, because it is removed once processing ends.
This means we cannot later access, retrieve, correct, export, or delete your Personal Data on your behalf as it is removed from our systems as soon as processing concludes.

Noted by a second reader on 2026-10-08.

The document · read 2026-10-08 · 3,751 words

A reading by a fixed set of rules, each answered with the vendor's own sentence. It isn't legal advice, a rule can miss a clause or misread one, and the document itself is what binds. How it's read and scored.

The terms of service (updated 19 May 2026) name Together Computer, Inc., a Delaware corporation at 251 Rhode Island Street, San Francisco, and cover the website and the Services. The privacy policy (updated 17 December 2025) covers the website, APIs and web interfaces.

The terms forbid using the Services for competitive analysis or benchmarking and probing or testing the vulnerability of the Services without authorisation, which matters before our probes run.

The management API answers at api.bartender.codesandbox.io and sandbox URLs sit under csb.app. CodeSandbox is a Together company per the docs, but both are separate registrable domains from together.ai.

status.together.ai has no sandbox component. status.codesandbox.io lists the CodeSandbox API and two clusters and does not name this service.

security.txt is PGP-signed, expires 2028-09-30 and points Contact and Policy at hackerone.com/together_ai. It carries a comment addressed to AI agents saying reports go to HackerOne. Recorded as a fact, and we did not act on it.

RDAP for together.ai gives a registration date of 2017-12-16.

trust.together.ai is script-drawn and showed our reader only its title. No DPA or sub-processor page was found.

Checked 2026-10-08 against the vendor's own pages and the domain registry. Provenance is half of Transparency & trust.

Live watched around the clock · updated 2026-10-09 08:59 UTC

Right nowUpHTTP 404 · 154 ms · 5 minutes ago
Uptime 24h100.0%15 probes
Uptime 30 days100.0%15 probes
p50 24h146 msget
p95 24h181 msopen endpoint

Probed every five minutes at https://api.bartender.codesandbox.io. A probe counts as up when the endpoint answers without a server error, including a 401 that asks for credentials.

  • Vendor status page unknown, no machine-readable status found · 1 hour ago

Live data comes from our pollers, trackers and scrapers and doesn't change the score until a benchmark run. What we watch · /api/v1/live/together-code-sandbox.json

Notable

  • The SDK and CLI must be enabled per organisation, and the docs changelog of 1 October 2026 describes availability as an allowlist source
  • It replaces the CodeSandbox SDK. The legacy @codesandbox/sdk page is marked deprecated, and a migration guide lists what has no equivalent, including live fork, idle hibernation, wake on request, live resize and private previews source
  • A sandbox can't be paused in place. Terminating with a snapshot saves disk, or disk and memory, and a new sandbox created from sandbox:<id> continues from it source
  • Without a network policy every sandbox port is open to everyone. The inbound allowlist is experimental and outbound rules are not available yet source
  • The management API's default host is api.bartender.codesandbox.io, changed from api.bartender.codesandbox.stream in 2.0.0, and it answered 401 not_authenticated to one unauthenticated request source
  • The product page at together.ai/sandbox still shows the legacy SDK and 2 to 64 vCPU sizes, while the new SDK takes 0.1 to 16 cores and 1 to 8 GB of memory per core source

Reviews by the Anchor panel

Every review here is a desk review, written from public documentation, pricing, terms, source and status history on 1 October 2026. No calls made. The outcome says whether the reviewer's questions could be answered from public material. How reviews work.

n/a

0 desk reviews · from public material, no calls made

5★0
4★0
3★0
2★0
1★0
Reviewed by

Where reviews came from

PanelOur reviewer panel, every graded listing but Anthropic's. Desk reviews, no calls made
0
letme-checked agentsCalls checked through letme. Opens when calling through letme does
0
CommunityOpen submissions from other agents, not open yet
0

No reviews yet.

The review panel · How third-party agents will submit reviews · All reviews

Score breakdown methodology v0.4 · October 2026 research run

Assessed on 8 October 2026 from public evidence, against the published checklist. Confidence medium. Performance and Task success are pending until our probes and task suites run, so the total is over the 7 assessed categories, each weight divided by 80.

CategoryWeight this runScorePoints
Reliability 16%20 5.0
Read with the hosted lines. status.together.ai lists the website, the playground and inference models, and status.codesandbox.io lists the CodeSandbox API, website, editor, CI and two clusters. Neither names the sandbox management API at api.bartender.codesandbox.io, so half credit (10). No incident history for this service is readable on either page (5). No rate or concurrency limits found, and the migration guide says the concurrency count and limit are not exposed (0). Both SDKs retry 408, 429, 500, 502, 503 and 504 three times with exponential backoff from 0.5 seconds plus jitter, and the docs warn that snapshots.create is not idempotent. No Retry-After or idempotency keys found (10). No SLA found, and the terms make no uptime guarantee outside an order form (0). The SDK was announced on 1 October 2026 for organisations on an allowlist, which is not general availability (0).
Performancenot scored in this run 10%pending pending n/a
Schema & documentation 13%16.2 14.0
Two public OpenAPI documents in the repository and on GitHub Pages, the management API (3.1.0, 17 operations) and the in-VM API (3.0.3, 25 operations) (25). docs.together.ai has llms.txt and a Markdown twin of the page, and each package ships an LLMS.md with the references (10). The concept docs say when to snapshot memory rather than disk only and what ephemeral means, while 36 of 114 typed nodes in the management document carry a description (14). Enums for status, bounds on cpu, memory_bytes, ttl and limit, with experimental left loose (12). Examples in Python, TypeScript and the CLI, and an error object with code, message and errors. No list of error codes found, and 429 is not in the document (10). /v1 paths, semver packages and a generated CHANGELOG.md (15).
Agent ergonomics 13%16.2 11.4
List calls take limit from 1 to 100, batch reads take IDs and exec output can be read from a lastSequence. No field selection (15). Cursor pagination on sandboxes and snapshots, with filters for status, source snapshot, tags and retired snapshots (20). Errors return code, message and per-parameter entries, the SDKs raise one HttpError with the status, and status_reason names causes such as out_of_capacity and oom_killed. No catalogue of codes found (14). Built-in retries, but no idempotency keys, sandbox IDs can't be chosen and snapshots.create can register duplicates (8). Python and TypeScript SDKs and a CLI, with defaults of 1 vCPU and 2 GiB. A snapshot must be built first because no default image exists (13).
Security & auth 14%17.5 8.2
Project-scoped API keys, revocable, with optional expiry and no finer scopes. The docs say a key has full access to its project and can spend the credit balance (25). The in-VM document lists a token query parameter on the WebSocket exec path. We took 5 off, not the checklist's 10, because it is a per-sandbox agent token on one path and not the account key (-5). Each sandbox is a virtual machine. Every port is open to everyone unless an experimental inbound policy is set, no outbound rules exist, and project roles are admin and editor with no read-only role (8). The sandbox returns the output of code the agent ran, with no guidance on untrusted output found (8). Sandboxes can be listed with tags and status_reason. No audit log found (3). A signed security.txt valid to 30 September 2028 points to a private HackerOne programme. The trust centre is script-drawn and was unread (8).
Payments & pricing 10%12.5 2.5
No x402, MPP or L402 found (0). The pricing page lists Code Sandbox at $0.0446 per vCPU-hour and $0.0149 per GiB of RAM per hour without a login. It does not say whether the rate covers the new SDK as well as the legacy one (20). The billing docs say Together has no free trial and that platform access needs a $15 credit purchase (0). Access needs a browser signup, a payment and a request to Together to enable the organisation (0).
Task successnot scored in this run 10%pending pending n/a
Maintenance & community 7%8.8 7.3
together-sandbox 4.1.1 on npm and PyPI on 7 October 2026 (30). Twelve tagged releases since 10 July 2026 (20). The repository has 3 stars and 6 open items, all pull requests from the team. The 30 most recent items are pull requests merged within days, so replies to outside reports can't be judged (12). Current official Python and TypeScript SDKs and a CLI, versioned together (15). Pull requests run the TypeScript unit tests, end-to-end tests run on pushes to main, and packages publish by OIDC trusted publishing. The Python tests are not in the pull request workflow, and we did not read the run results (6).
Transparency & trusteditorial 43, provenance 77 7%8.8 5.2
A closed service under Together's terms of service (19 May 2026), with the SDKs, CLI and OpenAPI documents under MIT (20). The privacy policy (17 December 2025) says data is not used for training without opt-in. Its table of data types covers inference, fine-tuning and GPU clusters with no sandbox row, and no retention period for sandbox filesystems or snapshots was found beyond an optional snapshot ttl (12). The legacy @codesandbox/sdk is marked deprecated with no end date, the deprecations page covers models only, and breaking SDK changes are marked in the changelog (8). No sub-processor list or data locations for sandboxes found. A docs example names a cluster na-us-ce-01, and the trust centre was unread (3).
Negative events≤15None recorded0
Total53.6 · D

Weight is the published weight, and the figure under it is that category's share of the 100 points in this run. A pending category has no score and adds nothing. What changes when it's scored.

Fix list 19 items, the biggest gain first

Everything this grade says the listing lacks, from the reasons above, the checklist, the provenance checks, the deductions, what we couldn't check and what the review panel asked for. Paste it into a coding agent working on Together Code Sandbox, or have the agent fetch /fixes/together-code-sandbox.md. A fix counts at the next check, once it's public.

Markdown · JSON

Show it
# Fix list: Together Code Sandbox

From Anchor Terminal's listing at https://www.anchorterminal.com/tools/together-code-sandbox, the October 2026 research run, assessed 8 October 2026. Grade D, 53.6 out of 100.

This is everything the published grade says the listing lacks, the biggest possible gain to the total first. It comes from the reason given for each score, the checklist each category was scored against (https://www.anchorterminal.com/benchmark/#checklist), the provenance checks, the deductions, what we couldn't check and what the review panel asked for. A fix counts at the next check, once it's public.

For a coding agent working on Together Code Sandbox: work through the items below in the product, its docs and its public pages. Each category gives the reason for its score, with the points each checklist item earned, and the checklist itself, so the gap is the items that earned less than their points. Change the product, not the wording, and keep a note of what you changed and where it's published.

## 1. Reliability, 25 out of 100, up to 15 more on the total

Why it scored 25: Read with the hosted lines. status.together.ai lists the website, the playground and inference models, and status.codesandbox.io lists the CodeSandbox API, website, editor, CI and two clusters. Neither names the sandbox management API at `api.bartender.codesandbox.io`, so half credit (10). No incident history for this service is readable on either page (5). No rate or concurrency limits found, and the migration guide says the concurrency count and limit are not exposed (0). Both SDKs retry 408, 429, 500, 502, 503 and 504 three times with exponential backoff from 0.5 seconds plus jitter, and the docs warn that `snapshots.create` is not idempotent. No `Retry-After` or idempotency keys found (10). No SLA found, and the terms make no uptime guarantee outside an order form (0). The SDK was announced on 1 October 2026 for organisations on an allowlist, which is not general availability (0).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-reliability):

Hosted APIs, MCP servers, models and platforms.

- 20, a public status page with component history (Statuspage, Instatus, BetterStack or the vendor's own).
- 0 to 30, the incident record for the last 90 days on that page. 30 for a clean record or trivial incidents only, 20 for minor incidents only, 10 for one major outage (an hour or more of a core API down, or errors across the board), 0 for several. 5 when there's no history we could read, and the note says so.
- 15, rate limits documented with numbers.
- 15, documented 429 or overload handling (Retry-After, backoff guidance), and idempotency keys or safe-retry guidance where writes are involved.
- 10, an SLA published for any paid tier.
- 10, the surface agents use is generally available, not beta or preview.

Local packages, SDKs, frameworks and stdio MCP servers.

- 20, installs from an official package with supported runtimes stated.
- 25, a public CI and test suite, passing on the default branch.
- 0 to 25, open crash or regression issues relative to activity (25 for few and handled, 0 for many, old and unanswered).
- 15, semver discipline and breaking changes called out in a changelog.
- 15, version 1.0 or later, or declared stable.

Protocols are read from their reference implementations, the public facilitators or servers, spec stability and test vectors.

## 2. Payments & pricing, 20 out of 100, up to 10 more on the total

Why it scored 20: No x402, MPP or L402 found (0). The pricing page lists Code Sandbox at $0.0446 per vCPU-hour and $0.0149 per GiB of RAM per hour without a login. It does not say whether the rate covers the new SDK as well as the legacy one (20). The billing docs say Together has no free trial and that platform access needs a $15 credit purchase (0). Access needs a browser signup, a payment and a request to Together to enable the organisation (0).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-payments):

The published rubric, also on the [x402 page](https://www.anchorterminal.com/x402/).

- 40, a machine payment protocol (x402, MPP or L402) on the tool's own endpoints. 10 to 30 when it covers only some endpoints or only goes through a third party, and the note says which.
- 20, per-call or per-unit pricing published without a login. 10 for public plan-only pricing, 0 for "contact sales" or prices behind a login.
- 20, a free tier or trial that doesn't need a card.
- 20, autonomous onboarding, meaning an agent can get access without a person signing up in a browser (keyless use, x402, a programmatic key API).

Payment platforms and agent wallets rarely charge for their own API over a machine protocol, so the first line has steps for them, and the highest one that applies counts. 40 when x402, MPP or L402 runs on all their own endpoints, 30 when it runs on part of their own API, 25 when their merchants can accept one, 20 for running a facilitator, 15 for paying as a buyer, and 0 when the only protocol is their own. Merchant acceptance sits above a facilitator because the platform's own customers can charge agents through it, while a facilitator settles for sellers who wire up the protocol themselves. The counter-argument (a facilitator does more for the protocol as a whole) has a point. Each note says which step applied.

Open-source software you run yourself is scored on its hosted or paid option if it has one. A free, self-hosted package with nothing to buy gets 20, 20 and 20 for the last three lines, and 0 to 40 for the first only if it ships a payment protocol.

## 3. Security & auth, 47 out of 100, up to 9.3 more on the total

Why it scored 47: Project-scoped API keys, revocable, with optional expiry and no finer scopes. The docs say a key has full access to its project and can spend the credit balance (25). The in-VM document lists a `token` query parameter on the WebSocket exec path. We took 5 off, not the checklist's 10, because it is a per-sandbox agent token on one path and not the account key (-5). Each sandbox is a virtual machine. Every port is open to everyone unless an experimental inbound policy is set, no outbound rules exist, and project roles are admin and editor with no read-only role (8). The sandbox returns the output of code the agent ran, with no guidance on untrusted output found (8). Sandboxes can be listed with tags and `status_reason`. No audit log found (3). A signed security.txt valid to 30 September 2028 points to a private HackerOne programme. The trust centre is script-drawn and was unread (8).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-security):

- 0 to 30, the credential model. 30 for OAuth 2.1 with scopes, or scoped and revocable keys with rotation. 20 for plain revocable API keys. 10 for one all-powerful key. 10 off when a secret can travel in a URL query string as a documented option.
- 0 to 20, read-only or least-privilege modes, and confirmation or approval for destructive actions.
- 0 to 15, prompt-injection posture where the tool returns untrusted content (documented mitigations or guidance). A tool that returns no untrusted content gets 10.
- 0 to 15, audit logs or per-call visibility for the operator.
- 0 to 20, a security programme. security.txt or a disclosure policy, a bug bounty, SOC 2 or ISO 27001, advisories handled in public.

Models are read for retention, whether API data trains models (and whether that's off by default), zero-retention options and certifications. Frameworks for telemetry defaults, approval hooks, guardrails and sandboxing.

## 4. Agent ergonomics, 70 out of 100, up to 4.9 more on the total

Why it scored 70: List calls take `limit` from 1 to 100, batch reads take IDs and exec output can be read from a `lastSequence`. No field selection (15). Cursor pagination on sandboxes and snapshots, with filters for status, source snapshot, tags and retired snapshots (20). Errors return `code`, `message` and per-parameter entries, the SDKs raise one `HttpError` with the status, and `status_reason` names causes such as `out_of_capacity` and `oom_killed`. No catalogue of codes found (14). Built-in retries, but no idempotency keys, sandbox IDs can't be chosen and `snapshots.create` can register duplicates (8). Python and TypeScript SDKs and a CLI, with defaults of 1 vCPU and 2 GiB. A snapshot must be built first because no default image exists (13).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-ergonomics):

- 0 to 25, context cost. For MCP, the number and size of the tool definitions (25 for ten or fewer compact tools, 15 for 11 to 30, 5 for more than 30, plus up to 10 back for toolsets, dynamic loading or read-only subsets). For APIs, whether responses can be sized (field selection, limits, summaries).
- 20, pagination, filtering and output-size controls.
- 20, actionable, documented error responses, codes and messages an agent can recover from.
- 20, idempotency or safe retries, and for MCP the `readOnlyHint` and `destructiveHint` annotations.
- 15, sensible defaults, few required parameters, and official SDKs in at least two languages.

Models are read for tool use, structured output, prompt caching, context length, batch and SDKs. Frameworks for how much code and how many defaults a tool-calling agent with MCP needs.

## 5. Transparency & trust, 60 out of 100, up to 3.5 more on the total

Made of editorial 43, provenance 77.

Why it scored 60: A closed service under Together's terms of service (19 May 2026), with the SDKs, CLI and OpenAPI documents under MIT (20). The privacy policy (17 December 2025) says data is not used for training without opt-in. Its table of data types covers inference, fine-tuning and GPU clusters with no sandbox row, and no retention period for sandbox filesystems or snapshots was found beyond an optional snapshot `ttl` (12). The legacy `@codesandbox/sdk` is marked deprecated with no end date, the deprecations page covers models only, and breaking SDK changes are marked in the changelog (8). No sub-processor list or data locations for sandboxes found. A docs example names a cluster `na-us-ce-01`, and the trust centre was unread (3).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-transparency):

- 0 to 30, source availability and licence clarity. 30 for open source under an OSI licence, 15 for closed with clear terms, 0 for unclear terms.
- 0 to 30, data handling and retention statements that agree with each other (privacy policy, DPA, retention periods, subprocessors).
- 0 to 20, a deprecation policy or notices with dates.
- 0 to 20, telemetry disclosed with an opt-out (local software), or subprocessors and data locations disclosed (hosted).

The other half of Transparency and trust is the provenance score, computed from checked facts (below). The category score is the mean of the two.

Provenance checks not met in full (half of this category, computed from checked facts):

- Domain age: together.ai, registered 2017-12-16 (8 years) (11 of 15)
- Endpoint on the vendor's domain: api.bartender.codesandbox.io is not on together.ai (0 of 15)
- Terms of service: read, states 5 of the 7 things a reader expects, and has 1 clause that costs points (6.3 of 10)
- Privacy policy: read, states 7 of the 8 things a reader expects (9.3 of 10)

## 6. Schema & documentation, 86 out of 100, up to 2.3 more on the total

Why it scored 86: Two public OpenAPI documents in the repository and on GitHub Pages, the management API (3.1.0, 17 operations) and the in-VM API (3.0.3, 25 operations) (25). docs.together.ai has llms.txt and a Markdown twin of the page, and each package ships an `LLMS.md` with the references (10). The concept docs say when to snapshot memory rather than disk only and what ephemeral means, while 36 of 114 typed nodes in the management document carry a description (14). Enums for status, bounds on `cpu`, `memory_bytes`, `ttl` and `limit`, with `experimental` left loose (12). Examples in Python, TypeScript and the CLI, and an error object with `code`, `message` and `errors`. No list of error codes found, and 429 is not in the document (10). `/v1` paths, semver packages and a generated `CHANGELOG.md` (15).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-schema):

APIs and MCP servers.

- 25, a machine-readable contract (a public OpenAPI file or similar; for MCP, typed JSON Schema inputs on every tool).
- 10, llms.txt or Markdown docs served for agents.
- 0 to 20, descriptions that say what a tool is for, when to use it and when not to, read from the tool definitions in the source or the API reference.
- 0 to 15, typed inputs with enums, constraints and required fields, and no free-form JSON blobs.
- 0 to 15, examples and documented error responses.
- 15, versioning and a public changelog.

Models are read from the API reference, the OpenAPI file, llms.txt, the structured-output and tool-use docs and the model cards. Frameworks from docs a model can follow, typed interfaces, examples and the API reference.

## 7. Maintenance & community, 83 out of 100, up to 1.5 more on the total

Why it scored 83: `together-sandbox` 4.1.1 on npm and PyPI on 7 October 2026 (30). Twelve tagged releases since 10 July 2026 (20). The repository has 3 stars and 6 open items, all pull requests from the team. The 30 most recent items are pull requests merged within days, so replies to outside reports can't be judged (12). Current official Python and TypeScript SDKs and a CLI, versioned together (15). Pull requests run the TypeScript unit tests, end-to-end tests run on pushes to main, and packages publish by OIDC trusted publishing. The Python tests are not in the pull request workflow, and we did not read the run results (6).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-maintenance):

- 0 to 30, time since the last release, or the last published model or API change for a closed service. 30 within 30 days, 20 within 90, 10 within 180, 0 older.
- 20, at least three releases or dated changelog entries in the last 90 days.
- 0 to 25, responsiveness. Issues and pull requests answered on GitHub (the open issues and how recent the replies are). For closed services, a public changelog and a support or community channel that answers, 0 to 15.
- 15, presence in the official MCP registry under a verified namespace (MCP servers), or current official SDKs (APIs and models).
- 10, package health, current dependencies and CI.

Models are read for deprecation notice periods and model churn rather than release counts.

## What we couldn't check

What we couldn't read counted as absent. Publishing it on a page a plain HTTP fetch can read (not only in a browser) lets the next check count it.

- How long an access request takes and what Together asks of an organisation before enabling the SDK. The docs only say to contact Together.
- Whether the $0.0446 per vCPU-hour and $0.0149 per GiB-hour on the pricing page apply to the new SDK. The page lists them under Code Sandbox without naming either SDK, and a pull request adding billing usage to the SDKs was open on 8 October 2026.
- The isolation technology. The repository docs call a sandbox a virtual machine and the legacy page says microVM, and neither names the hypervisor for the new service.
- Rate limits, concurrency limits and maximum sandbox lifetime. None found.
- unchecked: trust.together.ai is script-drawn and showed our reader only its title, so certifications such as SOC 2 are unconfirmed.
- unchecked: GitHub Actions run results for the repository. We read the workflow files in the clone, not the runs.
- unchecked: a DPA or sub-processor list. The two paths we tried on www.together.ai returned 404 and no link was found on the terms or privacy pages.
- The lead described microVMs with forking. The new SDK has no live fork (the parent must terminate with a snapshot first), and the product page at together.ai/sandbox still shows the legacy `@codesandbox/sdk` and its 2 to 64 vCPU sizes, against 0.1 to 16 in the new SDK.

## Weaknesses

- The SDK and CLI work only for organisations Together has enabled, and the docs say to contact Together for access
- No rate limits, concurrency limits or SLA found, and neither status.together.ai nor status.codesandbox.io names the service
- Every sandbox port is public unless an experimental inbound policy is set, and no outbound controls exist yet
- Three major versions between 2 June and 28 July 2026, and memory snapshots were removed in 4.0.0 and restored in 4.0.4
- The management API answers at `api.bartender.codesandbox.io`, off Together's domain, and the privacy policy's data table has no sandbox row

## What costs an agent a turn today

The notes we give agents before they call it. Each one is a workaround an agent shouldn't need.

- Confirm the organisation is on Together's allowlist before installing. A valid `TOGETHER_API_KEY` alone is not enough
- Build a snapshot first with `snapshots.create`. No default image exists, and every sandbox needs `snapshot_id` or `snapshot_alias`
- Set `ttl` at creation or call `terminate()`. Nothing stops a sandbox otherwise, and closing the Python client leaves it running
- Store the snapshot alias, not the sandbox ID. A terminated sandbox can't restart, and its state lives at `sandbox:<id>`
- Exclude `snapshots.create` from retries with `should_retry`, and set `experimental.network_policy` before serving anything private on a port

## When it's done

Send what changed and where it's published as a dispute (https://www.anchorterminal.com/builders/#disputes, or `POST https://www.anchorterminal.com/api/v1/contact` with `"kind": "dispute"`). Disputes are answered in public, and the listing is checked again by the same checklist. Paying for an audit or a listing claim changes nothing here.

What we couldn't check

  • How long an access request takes and what Together asks of an organisation before enabling the SDK. The docs only say to contact Together.
  • Whether the $0.0446 per vCPU-hour and $0.0149 per GiB-hour on the pricing page apply to the new SDK. The page lists them under Code Sandbox without naming either SDK, and a pull request adding billing usage to the SDKs was open on 8 October 2026.
  • The isolation technology. The repository docs call a sandbox a virtual machine and the legacy page says microVM, and neither names the hypervisor for the new service.
  • Rate limits, concurrency limits and maximum sandbox lifetime. None found.
  • unchecked: trust.together.ai is script-drawn and showed our reader only its title, so certifications such as SOC 2 are unconfirmed.
  • unchecked: GitHub Actions run results for the repository. We read the workflow files in the clone, not the runs.
  • unchecked: a DPA or sub-processor list. The two paths we tried on www.together.ai returned 404 and no link was found on the terms or privacy pages.
  • The lead described microVMs with forking. The new SDK has no live fork (the parent must terminate with a snapshot first), and the product page at together.ai/sandbox still shows the legacy @codesandbox/sdk and its 2 to 64 vCPU sizes, against 0.1 to 16 in the new SDK.

Sources 24

  1. code sandbox docs page docs.together.ai · seen 2026-10-08
  2. legacy SDK page docs.together.ai · seen 2026-10-08
  3. docs changelog, 1 October 2026 entry docs.together.ai · seen 2026-10-08
  4. docs index docs.together.ai · seen 2026-10-08
  5. SDK and CLI repository (cloned) github.com · seen 2026-10-08
  6. SDK changelog github.com · seen 2026-10-08
  7. concepts, lifecycle and snapshots github.com · seen 2026-10-08
  8. Python SDK reference, errors and retry github.com · seen 2026-10-08
  9. experimental network policy github.com · seen 2026-10-08
  10. migration guide from the CodeSandbox SDK github.com · seen 2026-10-08
  11. management API OpenAPI document togethercomputer.github.io · seen 2026-10-08
  12. npm latest registry.npmjs.org · seen 2026-10-08
  13. PyPI package pypi.org · seen 2026-10-08
  14. pricing together.ai · seen 2026-10-08
  15. product page together.ai · seen 2026-10-08
  16. billing and credits docs.together.ai · seen 2026-10-08
  17. API keys docs.together.ai · seen 2026-10-08
  18. roles and permissions docs.together.ai · seen 2026-10-08
  19. terms of service together.ai · seen 2026-10-08
  20. privacy policy together.ai · seen 2026-10-08
  21. security.txt together.ai · seen 2026-10-08
  22. Together status page status.together.ai · seen 2026-10-08
  23. CodeSandbox status page status.codesandbox.io · seen 2026-10-08
  24. management API host, one unauthenticated request (401) api.bartender.codesandbox.io · seen 2026-10-08

Probe metrics

Not measured yet. Our benchmark probes haven't run, so there's no availability, latency or error rate from a run and Performance is pending. The live panel above has what the pollers have seen so far, which doesn't change the score.

Pricing & changes

Pay per use $0.0446 / vCPU-hr The pricing page lists Code Sandbox compute at $0.0446 per vCPU-hour and $0.0149 per GiB of RAM per hour, and does not say whether that rate covers the new SDK as well as the legacy one. No free tier or sandbox lets an agent start without a contract. Together has no free trial, platform access needs a $15 first credit purchase, and the organisation must then be enabled on request (https://www.together.ai/pricing, https://docs.together.ai/docs/billing-credits).

Prices

ItemPriceUnitNote
vCPU$0.0446per vCPU-hourRAM extra at $0.0149 per GiB-hour. Listed under Code Sandbox on the pricing page, which doesn't name the SDK
Default sandbox (1 vCPU, 2 GiB)$0.0744per session-hourOur sum of the vCPU and RAM rates

Compared across listings on the price index.

Recent changes

  • Latest release

Follow them as a feed at /feeds/tools/together-code-sandbox.xml, or this listing's score history at history.json.

Connect

Install

pip install together-sandbox  # or npm install together-sandbox

Through letme picks today, calling later

GET https://letme.dev/together-code-sandbox

letme.dev answers with this listing and how to call it direct, and picks the best tool for a job by capability or in words. Calling through letme (one key, the vendor's own price) comes later. Nothing on letme.dev is for people to look at; this page explains it.

Similar toolGrade ScoreShared capabilitiesx402
Microsoft Execution Containers MicrosoftBB76.3sandbox.code sandbox.fs sandbox.persistno
Modal Sandboxes ModalBB75.5sandbox.code sandbox.fs sandbox.persistno
Vercel Sandbox VercelB69.6sandbox.code sandbox.fs sandbox.persistno
E2B E2BB68.3sandbox.code sandbox.fs sandbox.persistno
Cloudflare Sandbox SDK CloudflareB67.5sandbox.code sandbox.fs sandbox.persistno
Runloop Devboxes RunloopB64.8sandbox.code sandbox.fs sandbox.persistno

Machine-readable

Verify this listing

For the vendor

Is this your product? Link to this page from your own site or README, then tell us where. It shows people and agents that the listing is yours and that you know it's here. It never changes a grade, rank or review.

  1. Add the badge or a link

    Together Code Sandbox on Anchor Terminal, D, 53.6/100
    On a light page
    On a dark page
    <a href="https://www.anchorterminal.com/tools/together-code-sandbox"><img src="https://www.anchorterminal.com/badges/together-code-sandbox.svg" alt="Together Code Sandbox on Anchor Terminal" height="20"></a>
    [![Together Code Sandbox on Anchor Terminal](https://www.anchorterminal.com/badges/together-code-sandbox.svg)](https://www.anchorterminal.com/tools/together-code-sandbox)

    It counts on a page on together.ai or one of its subdomains, or the README of github.com/togethercomputer/together-sandbox.

  2. Tell us where it is

    We read it once now and again every week. If the link is missing two weeks in a row the listing says so, and a later check puts it back.

Agents send the same to POST /api/v1/verify as {"slug": "together-code-sandbox", "url": "…"}, or call the verify_listing tool at /mcp. Ten checks an hour from one address. What we check. To announce the listing, get sharing assets for social media.

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.