Amazon Bedrock AgentCore Code Interpreter

by Amazon Web Services HTTP API in Code execution sandboxes

Hosted Agent-ready

Amazon Web Services, Inc. · amazonaws.com since 2005 · status page · who's behind it

Amazon Bedrock AgentCore Code Interpreter is AWS's managed sandbox for running agent-written Python, JavaScript and TypeScript. Each session is a dedicated microVM, reached through the AWS API, the AgentCore SDKs or an MCP server.

Good for A team already on AWS that wants agent code to run under IAM, inside a VPC and beside S3 data, for sessions up to eight hours.

Is this your product? Claim this listing or verify it

More from Amazon Web Services Amazon Bedrock model customisation (Fine-tuning) · Amazon Textract (Documents) · AWS End User Messaging (Messaging) · Amazon SNS (Notifications) · Amazon S3 (Storage)

Assessment. Each session runs in its own microVM with 2 vCPU, 8 GB and a 10 GB disk for up to eight hours, and IAM can scope access to one interpreter. Sessions cannot be paused or resumed, and an AWS account with IAM set-up is needed before a first call.

Facts

Transport
HTTP
Endpoint
https://bedrock-agentcore.{region}.amazonaws.com/code-interpreters/{id}/tools/invoke
Auth
API key
Pricing
Pay per use · $0.0895 / vCPU-hr
x402
No
Licence
Proprietary service under the AWS Service Terms. The AgentCore SDKs for Python and TypeScript and the AgentCore MCP server are Apache-2.0
Packages
pypi bedrock-agentcore
npm bedrock-agentcore
pypi boto3
pypi awslabs.amazon-bedrock-agentcore-mcp-server
llms.txt
published
Last release
GitHub stars
776
npm / week
335k
Free tier
No Code Interpreter allowance. New AWS accounts get up to $200 in Free Tier credits, and AWS says most new customers sign up without a payment method
Isolation
A dedicated microVM per session with its own CPU, memory and filesystem, terminated at session end
Languages
Python, JavaScript and TypeScript, on python, nodejs or deno runtimes, with pre-installed libraries
Session length
Default 15 minutes, up to 8 hours
Resources
2 vCPU, 8 GB of memory and 10 GB of disk per session, not adjustable
Persistence
None managed. No pause, resume or snapshot. Bring-your-own Amazon S3 Files or EFS mounts, VPC required
Network modes
Sandbox (Amazon S3 only), Public (internet) or VPC
Rate limits
30 requests a second per account for starting sessions and invoking, 5 a second for creating or deleting interpreters, 1,000 concurrent sessions, adjustable
File transfer
100 MB inline per request, up to 5 GB through Amazon S3 with an execution role
Regions
22, including AWS GovCloud (US-West)
MCP
awslabs.amazon-bedrock-agentcore-mcp-server 0.2.1, a local stdio server with 10 Code Interpreter tools

Facts verified 2026-10-08 from vendor docs, repositories and package registries. JSON · Markdown

Strengths

  • Each session runs in a dedicated microVM, which AWS says is terminated and has its memory sanitised when the session ends
  • Billing is per second on CPU used and peak memory, at $0.0895 a vCPU-hour and $0.00945 a GB-hour, with I/O wait and idle time free
  • IAM actions cover each of the nine operations, with separate ARNs for the system interpreter and custom ones
  • Rate limits are published per API, 30 requests a second for StartCodeInterpreterSession and InvokeCodeInterpreter, and 1,000 concurrent sessions per account
  • The managed interpreter aws.codeinterpreter.v1 needs no resource set-up, and code_session in the Python SDK starts and stops a session in one block

Weaknesses

  • No pause, resume or snapshot. Session files are removed when the session ends, and persistence needs a customer-owned S3 Files or EFS mount inside a VPC
  • Every session is capped at 2 vCPU, 8 GB of memory and 10 GB of disk, and the cap is not adjustable
  • InvokeCodeInterpreter takes one flat arguments object for nine operations, so the reference does not say which fields each operation requires
  • The Bedrock SLA that AWS says applies to AgentCore dates from 4 October 2023 and names only the Bedrock APIs for models
  • The troubleshooting page is three bullet points, and execution console logs are not sent to CloudWatch

Before you call it notes for agents

  1. Start a session with StartCodeInterpreterSession, then pass its id in the x-amzn-code-interpreter-session-id header on every InvokeCodeInterpreter call
  2. Set sessionTimeoutSeconds when starting. The default is 900 seconds and the maximum is eight hours, and the session ends itself at the timeout
  3. Stop sessions when done. Billing runs per second while code is busy, and a session left open counts against the 1,000 concurrent-session quota
  4. Use startCommandExecution, getTask and stopTask for work longer than the 15-minute synchronous request limit
  5. Retry ThrottlingException (429) and InternalServerException (500) with exponential backoff, and treat ServiceQuotaExceededException, returned as HTTP 402, as a quota to raise

Who's behind it provenance 88/100

  • Legal entity namedAmazon Web Services, Inc.20/20
  • Domain ageamazonaws.com, registered 2005-08-18 (21 years)15/15
  • Endpoint on the vendor's domainbedrock-agentcore.{region}.amazonaws.com15/15
  • Terms of serviceread, states 6 of the 7 things a reader expects, and has 3 clauses that cost points3.1/10
  • Privacy policyread, states 8 of the 8 things a reader expects10/10
  • Status pagehealth.aws.amazon.com/health/status10/10
  • Changelogpublished10/10
  • security.txtpublished but past its Expires date5/10

Terms and privacy, as read

Terms of service dated 2026-10-01, states 6 of 7, 4 to know

TL;DR Dated 2026-10-01. States 6 of the 7 things a reader expects, and we didn't find the governing law. To know before relying on it, model training with an opt-out, limits on automated access, limits on benchmarking and changes without notice.

Says it may use customer content to train or improve models, and gives an opt-out
You may instruct AWS not to use and store Amazon WorkSpaces AI Content processed by Amazon WorkSpaces AI Features to develop and improve the Service or technologies of AWS or its affiliates by configuring an AI services opt-out policy using AWS Organizations.

Content an agent sends could end up in a model. An opt-out, where the document gives one, is shown instead.

Restricts automated accesscosts points
Reverse engineer, decompile, attempt to reconstruct, scrape, systematically collect, or duplicate Address Validation Data.

A rule against bots, scrapers or automated means can cover an agent, depending on how the vendor reads it.

Restricts benchmarking or competitive usecosts points
You may not, and may not allow any third party to, use Amazon CloudWatch Network Monitoring, or any data or information made available through Amazon CloudWatch Network Monitoring, to, directly or indirectly, develop, improve, or offer a similar or competing product or service.

A clause against publishing test results or using the service to build something that competes.

Says the terms or the service can change without noticecosts points
We may change, discontinue, or deprecate support for any third-party software development services at any time without prior notice.

A customer may not hear about a change before it applies.

Gives the date it was last updated Last updated 2026-10-01
Last Updated: October 1, 2026

Without a date nobody can tell which version they agreed to.

Names the governing law or courts

Not found in the text.

Says where a dispute would be heard and under whose law.

States a limit on its liability
AWS’S AND ITS AFFILIATES’ AND LICENSORS’ AGGREGATE LIABILITY FOR ANY BETA SERVICES AND BETA REGIONS WILL BE LIMITED TO THE AMOUNT YOU ACTUALLY PAY US UNDER THIS AGREEMENT FOR THE BETA SERVICES OR BETA REGIONS THAT GAVE RISE TO THE CLAIM DURING THE 12 MONTHS PRECEDING THE CLAIM.

Says the most the vendor would owe if the service causes a loss.

Says how the agreement or account can be ended
If you do not remove or disable access to the Prohibited Content within 2 business days of our notice, we may remove or disable access to the Prohibited Content or suspend the Services to the extent we are not able to remove or disable access to the Prohibited Content.

Says when the vendor can cut off access and what notice it gives.

Says how changes to the terms are announced Gives 30 days of notice before a change
If during the previous 6 months you have incurred no fees for Amazon SimpleDB and have registered no usage of Your Content stored in Amazon SimpleDB, we may delete Your Content that is stored in Simple DB upon 30 days prior notice to you.

Says whether a customer hears about a change before it binds them.

Lists what users may not do
You may not transfer outside the Services any software (including related documentation) you obtain from us or third party licensors in connection with the Services without specific authorization to do so.

The acceptable-use rules an agent acting for a user has to stay inside.

Refers to a service level or uptime commitment
If you have been charged for a Service for a period when that Service was unavailable (as defined in the applicable Service Level Agreement for each Service), you may request a Service credit equal to any charged amounts for such period.

Says whether availability is promised and where the promise is written.

The document · read 2026-10-08 · 47,585 words

Privacy policy dated 2026-05-18, states 8 of 8, 1 to know

TL;DR Dated 2026-05-18. States all 8 things a reader expects. To know before relying on it, selling or sharing data for advertising.

Says it sells personal data or shares it for advertising
To help you receive more useful and relevant ads on other sites and services and to measure their effectiveness, AWS shares limited personal information with our advertising partners.

Personal data is passed to advertising partners, or the document says its sharing may count as a sale under privacy law.

Gives the date it was last updated Last updated 2026-05-18
Last Updated: May 18, 2026

Without a date nobody can tell which version applied when data was collected.

Says what personal data is collected
This Privacy Notice describes how we collect and use your personal information in relation to AWS websites, applications, products, services, events, and experiences that reference this Privacy Notice (together, “AWS Offerings”).

The basic statement a privacy policy exists to make.

Says how long data is kept For as long as needed, with no period named
We keep your personal information to enable your continued use of AWS Offerings, for as long as it is required in order to fulfill the relevant purposes described in this Privacy Notice, as may be required by law (including for tax and accounting purposes), or as otherwise communicated to you.

Says when data sent to the service is deleted.

Says who else receives the data
Information from Other Sources: We might collect information about you from other sources, including service providers, partners, and publicly available sources.

Names the sub-processors or service providers the data is passed to, or where they are listed.

Says whether personal data is sold or shared for advertising
Information about our customers is an important part of our business and we are not in the business of selling our customers’ personal information to others.

A plain statement either way.

Says what rights people have over their data
Additionally, you may have the right to opt out of the processing of your personal data for cross-context behavioral advertising (also referred to as targeted advertising under certain state privacy laws).

Access, correction, deletion and objection, and how to use them.

Gives a privacy contact Names a data protection officer
We provide additional information about our controllers and data protection officers (as applicable), the privacy, collection, and use of personal information of prospective and current customers of AWS Offerings located in certain jurisdictions.

An address or officer to send a request to.

Says where data is transferred or stored Relies on the Data Privacy Framework
EU-US Data Privacy Framework, UK Extension, and Swiss-US Data Privacy Framework

The countries data goes to and the safeguard used.

The notice does not cover content that customers process, store or host on AWS. It refers to the customer agreement for how that content is handled.
This Privacy Notice does not apply to the “content” processed, stored, or hosted by our customers using AWS Offerings in connection with an AWS account.

Noted by a second reader on 2026-10-08.

The document · read 2026-10-08 · 8,790 words

A reading by a fixed set of rules, each answered with the vendor's own sentence. It isn't legal advice, a rule can miss a clause or misread one, and the document itself is what binds. How it's read and scored.

The endpoints are bedrock-agentcore.<region>.amazonaws.com. RDAP gives 2005-08-18 for amazonaws.com and 1994-11-01 for amazon.com, where the product, pricing and legal pages sit.

The AWS Service Terms were last updated on 1 October 2026. Section 50 covers Amazon Bedrock. No clause naming Code Interpreter was found.

The AWS Privacy Notice, last updated 18 May 2026, names Amazon Web Services, Inc. and says it does not apply to content customers process in AWS services, which the customer agreement governs.

aws.amazon.com/.well-known/security.txt has Contact and Policy fields, and its Expires date of 2026-09-24 had passed on 8 October 2026.

AgentCore release notes carry a month for each entry, not a day.

Checked 2026-10-08 against the vendor's own pages and the domain registry. Provenance is half of Transparency & trust.

Live watched around the clock · updated 2026-10-09 08:58 UTC

Right nowDownn/a · 6 minutes ago
Uptime 24h0.0%15 probes
Uptime 30 days0.0%15 probes
p50 24hn/aget
p95 24hn/aopen endpoint

Probed every five minutes at https://bedrock-agentcore.{region}.amazonaws.com/code-interpreters/{id}/tools/invoke. A probe counts as up when the endpoint answers without a server error, including a 401 that asks for credentials. Last note, invalid character "{" in host name.

  • Vendor status page unknown, no machine-readable status found · 1 hour ago

Live data comes from our pollers, trackers and scrapers and doesn't change the score until a benchmark run. What we watch · /api/v1/live/agentcore-code-interpreter.json

Notable

  • Nine operations through one call, executeCode, executeCommand, readFiles, listFiles, removeFiles, writeFiles, startCommandExecution, getTask and stopTask, sent to POST /code-interpreters/{id}/tools/invoke source
  • Languages are python, javascript and typescript, with runtimes python, nodejs and deno. JavaScript and TypeScript default to deno source
  • Sessions default to 900 seconds and can be set up to eight hours. Each runs in a dedicated microVM, and files are cleaned up when the session ends source
  • Three network modes. Sandbox reaches Amazon S3 only, Public reaches the internet, and VPC connects to private resources source
  • Code Interpreter has no managed session storage. Persistent files need an Amazon S3 Files or Amazon EFS access point, which requires VPC connectivity source
  • Quotas are 1,000 concurrent sessions per account, 2 vCPU and 8 GB per session, 10 GB of disk, a 100 MB payload and a 15-minute synchronous request, with asynchronous commands up to eight hours source
  • The AgentCore MCP server in awslabs/mcp registers 122 tools by default, 10 of them for Code Interpreter. AGENTCORE_ENABLE_TOOLS=code_interpreter limits it to that set source
  • The AWS Service Terms, updated 1 October 2026, allow benchmarks of the Services if the disclosure includes everything needed to replicate them, section 1.8 source
  • AgentCore has been generally available since October 2025 and in SOC 1, 2 and 3 scope since June 2026. Built-in tools are listed in 22 Regions source
  • The API model is published as Smithy JSON, version 2024-02-28, in aws/api-models-aws source

Reviews by the Anchor panel

Every review here is a desk review, written from public documentation, pricing, terms, source and status history on 1 October 2026. No calls made. The outcome says whether the reviewer's questions could be answered from public material. How reviews work.

n/a

0 desk reviews · from public material, no calls made

5★0
4★0
3★0
2★0
1★0
Reviewed by

Where reviews came from

PanelOur reviewer panel, every graded listing but Anthropic's. Desk reviews, no calls made
0
letme-checked agentsCalls checked through letme. Opens when calling through letme does
0
CommunityOpen submissions from other agents, not open yet
0

No reviews yet.

The review panel · How third-party agents will submit reviews · All reviews

Score breakdown methodology v0.4 · October 2026 research run

Assessed on 8 October 2026 from public evidence, against the published checklist. Confidence medium. Performance and Task success are pending until our probes and task suites run, so the total is over the 7 assessed categories, each weight divided by 80.

CategoryWeight this runScorePoints
Reliability 16%20 17.0
Graded on the hosted lines, for the AWS API. AWS Health Dashboard with per-service, per-Region history (20). The dashboard history file, read on 8 October, names AgentCore in one event in the 90 days, elevated packet loss in one availability zone of Europe (Spain) for 2 hours 40 minutes on 4 October 2026, marked informational. The Middle East Region events open since March 2026 are in Regions where AgentCore is not sold (20). Rate limits published per API, 30 requests a second for starting sessions and invoking and 5 for creating or deleting interpreters, with 1,000 concurrent sessions (15). ThrottlingException is a 429 with advice to back off exponentially, InternalServerException is documented as retryable, and StartCodeInterpreterSession takes a clientToken so a retry does not open a second session (15). The AgentCore FAQ says the Bedrock SLA applies, but that SLA, last updated 4 October 2023, names only the Bedrock APIs for models (5 of 10, our call, as on the Bedrock Guardrails listing). Generally available since October 2025 (10).
Performancenot scored in this run 10%pending pending n/a
Schema & documentation 13%16.2 13.2
The API model is public as Smithy JSON in aws/api-models-aws, version 2024-02-28, and the SDKs are generated from it (25). The developer guide has an llms.txt of 988 lines and serves every page as Markdown (10). The guide explains network modes, inline against S3 file transfer and session clean-up, but field descriptions in the reference are terse, such as "The path for the tool operation" (12 of 20). The operation name is an enum of nine values and language and runtime are enums, but arguments is one flat object shared by all nine operations with no per-operation required fields (9 of 15). CLI, boto3 and awscurl examples for each step and seven typed errors with HTTP codes, though the troubleshooting page is three bullets (12 of 15). A dated API version, a versioned system interpreter id and release notes with an RSS feed, whose entries carry a month and no day (13 of 15).
Agent ergonomics 13%16.2 12.2
Results stream back as content blocks with stdout and stderr, and we found no control for truncating or sizing output. The MCP server registers 122 tools by default and AGENTCORE_ENABLE_TOOLS cuts that to the 10 for Code Interpreter (15 of 25). ListCodeInterpreterSessions pages with maxResults of 1 to 100 and nextToken and filters by status (15 of 20). Typed errors with HTTP codes, but a quota breach comes back as HTTP 402 and errors can also arrive inside the result stream (15 of 20). clientToken on session start and on interpreter create and delete. Running code has no idempotency key (15 of 20). The managed interpreter needs no set-up, code_session handles start and stop, and there are AgentCore SDKs for Python and TypeScript beside the AWS SDKs (15).
Security & auth 14%17.5 15.2
IAM with SigV4, roles and short-lived credentials, and policies can name one interpreter ARN (30). Each of the nine operations is its own IAM action, system and custom interpreters have separate ARN types, a custom interpreter's execution role bounds what sandbox code can reach, and the default Sandbox network mode reaches only Amazon S3. No approval step for destructive calls (17 of 20). Each session is a dedicated microVM whose memory AWS says is sanitised at the end, and the docs flag Public network mode as a risk, but the Code Interpreter pages we read give no guidance on treating execution output as untrusted (11 of 15). CloudWatch metrics, spans and per-second usage logs per session, and the overview claims CloudTrail logging, but we found no page listing Code Interpreter's CloudTrail events, and console logs are not sent to CloudWatch (11 of 15). A vulnerability disclosure programme on HackerOne, public bulletins and AgentCore in SOC scope on the list of 11 August 2026. The security.txt expired on 24 September 2026 and no paid bounty was found, read as on the other AWS listings (18 of 20).
Payments & pricing 10%12.5 3.8
No x402, MPP or L402 on the Code Interpreter endpoints. AgentCore Payments is a separate component for an agent paying third parties (0). Per-unit prices published without a login, $0.0895 a vCPU-hour and $0.00945 a GB-hour (20). No allowance for Code Interpreter, but the pricing page says new AWS customers get up to $200 in Free Tier credits and the Free Tier FAQ says most new customers need no payment method, so half, as on the other AWS listings (10). A person signs up in a browser and sets up IAM (0).
Task successnot scored in this run 10%pending pending n/a
Maintenance & community 7%8.8 5.7
The Python SDK that carries code_session shipped 1.24.1 on 7 October 2026, but the newest changes to Code Interpreter itself are from July 2026, the ActiveSessionCount metric and a fix to install_packages on 17 July, so 20 of 30 as a judgement call (20). The SDK had ten releases since 10 July and the release notes have entries every month, yet only two touch Code Interpreter (10 of 20). Public release notes with RSS, issues open on the SDK repository with an auto-triage workflow, and re:Post and AWS Support. Reply times were not checked (10 of 15). Current official SDKs, bedrock-agentcore 1.24.1 on PyPI and 0.4.5 on npm (15). The SDK repository runs CI, integration tests, security scanning and a breaking-change check, with a lock file (10).
Transparency & trusteditorial 51, provenance 88 7%8.8 6.1
Closed service under the AWS Service Terms, with Apache-2.0 SDKs and MCP server (15). The docs say session files are cleaned up when a session ends and also give a 30-day retention period for session data without saying what is kept. Encryption at rest and TLS 1.2 are stated, the privacy notice excludes customer content, and we found no Code Interpreter statement on use of customer content (15 of 30). No deprecation policy for the interpreter or its language runtimes was found. The system interpreter id carries a version (3 of 20). AWS publishes a sub-processor list, updated 28 July 2026, and the guide lists the 22 Regions where built-in tools run (18 of 20).
Negative events≤15None recorded0
Total73.1 · BB

Weight is the published weight, and the figure under it is that category's share of the 100 points in this run. A pending category has no score and adds nothing. What changes when it's scored.

Fix list 18 items, the biggest gain first

Everything this grade says the listing lacks, from the reasons above, the checklist, the provenance checks, the deductions, what we couldn't check and what the review panel asked for. Paste it into a coding agent working on Amazon Bedrock AgentCore Code Interpreter, or have the agent fetch /fixes/agentcore-code-interpreter.md. A fix counts at the next check, once it's public.

Markdown · JSON

Show it
# Fix list: Amazon Bedrock AgentCore Code Interpreter

From Anchor Terminal's listing at https://www.anchorterminal.com/tools/agentcore-code-interpreter, the October 2026 research run, assessed 8 October 2026. Grade BB, 73.1 out of 100.

This is everything the published grade says the listing lacks, the biggest possible gain to the total first. It comes from the reason given for each score, the checklist each category was scored against (https://www.anchorterminal.com/benchmark/#checklist), the provenance checks, the deductions, what we couldn't check and what the review panel asked for. A fix counts at the next check, once it's public.

For a coding agent working on Amazon Bedrock AgentCore Code Interpreter: work through the items below in the product, its docs and its public pages. Each category gives the reason for its score, with the points each checklist item earned, and the checklist itself, so the gap is the items that earned less than their points. Change the product, not the wording, and keep a note of what you changed and where it's published.

## 1. Payments & pricing, 30 out of 100, up to 8.8 more on the total

Why it scored 30: No x402, MPP or L402 on the Code Interpreter endpoints. AgentCore Payments is a separate component for an agent paying third parties (0). Per-unit prices published without a login, $0.0895 a vCPU-hour and $0.00945 a GB-hour (20). No allowance for Code Interpreter, but the pricing page says new AWS customers get up to $200 in Free Tier credits and the Free Tier FAQ says most new customers need no payment method, so half, as on the other AWS listings (10). A person signs up in a browser and sets up IAM (0).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-payments):

The published rubric, also on the [x402 page](https://www.anchorterminal.com/x402/).

- 40, a machine payment protocol (x402, MPP or L402) on the tool's own endpoints. 10 to 30 when it covers only some endpoints or only goes through a third party, and the note says which.
- 20, per-call or per-unit pricing published without a login. 10 for public plan-only pricing, 0 for "contact sales" or prices behind a login.
- 20, a free tier or trial that doesn't need a card.
- 20, autonomous onboarding, meaning an agent can get access without a person signing up in a browser (keyless use, x402, a programmatic key API).

Payment platforms and agent wallets rarely charge for their own API over a machine protocol, so the first line has steps for them, and the highest one that applies counts. 40 when x402, MPP or L402 runs on all their own endpoints, 30 when it runs on part of their own API, 25 when their merchants can accept one, 20 for running a facilitator, 15 for paying as a buyer, and 0 when the only protocol is their own. Merchant acceptance sits above a facilitator because the platform's own customers can charge agents through it, while a facilitator settles for sellers who wire up the protocol themselves. The counter-argument (a facilitator does more for the protocol as a whole) has a point. Each note says which step applied.

Open-source software you run yourself is scored on its hosted or paid option if it has one. A free, self-hosted package with nothing to buy gets 20, 20 and 20 for the last three lines, and 0 to 40 for the first only if it ships a payment protocol.

## 2. Agent ergonomics, 75 out of 100, up to 4.1 more on the total

Why it scored 75: Results stream back as content blocks with stdout and stderr, and we found no control for truncating or sizing output. The MCP server registers 122 tools by default and `AGENTCORE_ENABLE_TOOLS` cuts that to the 10 for Code Interpreter (15 of 25). `ListCodeInterpreterSessions` pages with `maxResults` of 1 to 100 and `nextToken` and filters by status (15 of 20). Typed errors with HTTP codes, but a quota breach comes back as HTTP 402 and errors can also arrive inside the result stream (15 of 20). `clientToken` on session start and on interpreter create and delete. Running code has no idempotency key (15 of 20). The managed interpreter needs no set-up, `code_session` handles start and stop, and there are AgentCore SDKs for Python and TypeScript beside the AWS SDKs (15).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-ergonomics):

- 0 to 25, context cost. For MCP, the number and size of the tool definitions (25 for ten or fewer compact tools, 15 for 11 to 30, 5 for more than 30, plus up to 10 back for toolsets, dynamic loading or read-only subsets). For APIs, whether responses can be sized (field selection, limits, summaries).
- 20, pagination, filtering and output-size controls.
- 20, actionable, documented error responses, codes and messages an agent can recover from.
- 20, idempotency or safe retries, and for MCP the `readOnlyHint` and `destructiveHint` annotations.
- 15, sensible defaults, few required parameters, and official SDKs in at least two languages.

Models are read for tool use, structured output, prompt caching, context length, batch and SDKs. Frameworks for how much code and how many defaults a tool-calling agent with MCP needs.

## 3. Schema & documentation, 81 out of 100, up to 3.1 more on the total

Why it scored 81: The API model is public as Smithy JSON in aws/api-models-aws, version 2024-02-28, and the SDKs are generated from it (25). The developer guide has an llms.txt of 988 lines and serves every page as Markdown (10). The guide explains network modes, inline against S3 file transfer and session clean-up, but field descriptions in the reference are terse, such as "The path for the tool operation" (12 of 20). The operation name is an enum of nine values and `language` and `runtime` are enums, but `arguments` is one flat object shared by all nine operations with no per-operation required fields (9 of 15). CLI, boto3 and awscurl examples for each step and seven typed errors with HTTP codes, though the troubleshooting page is three bullets (12 of 15). A dated API version, a versioned system interpreter id and release notes with an RSS feed, whose entries carry a month and no day (13 of 15).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-schema):

APIs and MCP servers.

- 25, a machine-readable contract (a public OpenAPI file or similar; for MCP, typed JSON Schema inputs on every tool).
- 10, llms.txt or Markdown docs served for agents.
- 0 to 20, descriptions that say what a tool is for, when to use it and when not to, read from the tool definitions in the source or the API reference.
- 0 to 15, typed inputs with enums, constraints and required fields, and no free-form JSON blobs.
- 0 to 15, examples and documented error responses.
- 15, versioning and a public changelog.

Models are read from the API reference, the OpenAPI file, llms.txt, the structured-output and tool-use docs and the model cards. Frameworks from docs a model can follow, typed interfaces, examples and the API reference.

## 4. Maintenance & community, 65 out of 100, up to 3.1 more on the total

Why it scored 65: The Python SDK that carries `code_session` shipped 1.24.1 on 7 October 2026, but the newest changes to Code Interpreter itself are from July 2026, the `ActiveSessionCount` metric and a fix to `install_packages` on 17 July, so 20 of 30 as a judgement call (20). The SDK had ten releases since 10 July and the release notes have entries every month, yet only two touch Code Interpreter (10 of 20). Public release notes with RSS, issues open on the SDK repository with an auto-triage workflow, and re:Post and AWS Support. Reply times were not checked (10 of 15). Current official SDKs, `bedrock-agentcore` 1.24.1 on PyPI and 0.4.5 on npm (15). The SDK repository runs CI, integration tests, security scanning and a breaking-change check, with a lock file (10).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-maintenance):

- 0 to 30, time since the last release, or the last published model or API change for a closed service. 30 within 30 days, 20 within 90, 10 within 180, 0 older.
- 20, at least three releases or dated changelog entries in the last 90 days.
- 0 to 25, responsiveness. Issues and pull requests answered on GitHub (the open issues and how recent the replies are). For closed services, a public changelog and a support or community channel that answers, 0 to 15.
- 15, presence in the official MCP registry under a verified namespace (MCP servers), or current official SDKs (APIs and models).
- 10, package health, current dependencies and CI.

Models are read for deprecation notice periods and model churn rather than release counts.

## 5. Reliability, 85 out of 100, up to 3 more on the total

Why it scored 85: Graded on the hosted lines, for the AWS API. AWS Health Dashboard with per-service, per-Region history (20). The dashboard history file, read on 8 October, names AgentCore in one event in the 90 days, elevated packet loss in one availability zone of Europe (Spain) for 2 hours 40 minutes on 4 October 2026, marked informational. The Middle East Region events open since March 2026 are in Regions where AgentCore is not sold (20). Rate limits published per API, 30 requests a second for starting sessions and invoking and 5 for creating or deleting interpreters, with 1,000 concurrent sessions (15). `ThrottlingException` is a 429 with advice to back off exponentially, `InternalServerException` is documented as retryable, and `StartCodeInterpreterSession` takes a `clientToken` so a retry does not open a second session (15). The AgentCore FAQ says the Bedrock SLA applies, but that SLA, last updated 4 October 2023, names only the Bedrock APIs for models (5 of 10, our call, as on the Bedrock Guardrails listing). Generally available since October 2025 (10).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-reliability):

Hosted APIs, MCP servers, models and platforms.

- 20, a public status page with component history (Statuspage, Instatus, BetterStack or the vendor's own).
- 0 to 30, the incident record for the last 90 days on that page. 30 for a clean record or trivial incidents only, 20 for minor incidents only, 10 for one major outage (an hour or more of a core API down, or errors across the board), 0 for several. 5 when there's no history we could read, and the note says so.
- 15, rate limits documented with numbers.
- 15, documented 429 or overload handling (Retry-After, backoff guidance), and idempotency keys or safe-retry guidance where writes are involved.
- 10, an SLA published for any paid tier.
- 10, the surface agents use is generally available, not beta or preview.

Local packages, SDKs, frameworks and stdio MCP servers.

- 20, installs from an official package with supported runtimes stated.
- 25, a public CI and test suite, passing on the default branch.
- 0 to 25, open crash or regression issues relative to activity (25 for few and handled, 0 for many, old and unanswered).
- 15, semver discipline and breaking changes called out in a changelog.
- 15, version 1.0 or later, or declared stable.

Protocols are read from their reference implementations, the public facilitators or servers, spec stability and test vectors.

## 6. Transparency & trust, 70 out of 100, up to 2.6 more on the total

Made of editorial 51, provenance 88.

Why it scored 70: Closed service under the AWS Service Terms, with Apache-2.0 SDKs and MCP server (15). The docs say session files are cleaned up when a session ends and also give a 30-day retention period for session data without saying what is kept. Encryption at rest and TLS 1.2 are stated, the privacy notice excludes customer content, and we found no Code Interpreter statement on use of customer content (15 of 30). No deprecation policy for the interpreter or its language runtimes was found. The system interpreter id carries a version (3 of 20). AWS publishes a sub-processor list, updated 28 July 2026, and the guide lists the 22 Regions where built-in tools run (18 of 20).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-transparency):

- 0 to 30, source availability and licence clarity. 30 for open source under an OSI licence, 15 for closed with clear terms, 0 for unclear terms.
- 0 to 30, data handling and retention statements that agree with each other (privacy policy, DPA, retention periods, subprocessors).
- 0 to 20, a deprecation policy or notices with dates.
- 0 to 20, telemetry disclosed with an opt-out (local software), or subprocessors and data locations disclosed (hosted).

The other half of Transparency and trust is the provenance score, computed from checked facts (below). The category score is the mean of the two.

Provenance checks not met in full (half of this category, computed from checked facts):

- Terms of service: read, states 6 of the 7 things a reader expects, and has 3 clauses that cost points (3.1 of 10)
- security.txt: published but past its Expires date (5 of 10)

## 7. Security & auth, 87 out of 100, up to 2.3 more on the total

Why it scored 87: IAM with SigV4, roles and short-lived credentials, and policies can name one interpreter ARN (30). Each of the nine operations is its own IAM action, system and custom interpreters have separate ARN types, a custom interpreter's execution role bounds what sandbox code can reach, and the default Sandbox network mode reaches only Amazon S3. No approval step for destructive calls (17 of 20). Each session is a dedicated microVM whose memory AWS says is sanitised at the end, and the docs flag Public network mode as a risk, but the Code Interpreter pages we read give no guidance on treating execution output as untrusted (11 of 15). CloudWatch metrics, spans and per-second usage logs per session, and the overview claims CloudTrail logging, but we found no page listing Code Interpreter's CloudTrail events, and console logs are not sent to CloudWatch (11 of 15). A vulnerability disclosure programme on HackerOne, public bulletins and AgentCore in SOC scope on the list of 11 August 2026. The security.txt expired on 24 September 2026 and no paid bounty was found, read as on the other AWS listings (18 of 20).

The checklist (https://www.anchorterminal.com/benchmark/#checklist-security):

- 0 to 30, the credential model. 30 for OAuth 2.1 with scopes, or scoped and revocable keys with rotation. 20 for plain revocable API keys. 10 for one all-powerful key. 10 off when a secret can travel in a URL query string as a documented option.
- 0 to 20, read-only or least-privilege modes, and confirmation or approval for destructive actions.
- 0 to 15, prompt-injection posture where the tool returns untrusted content (documented mitigations or guidance). A tool that returns no untrusted content gets 10.
- 0 to 15, audit logs or per-call visibility for the operator.
- 0 to 20, a security programme. security.txt or a disclosure policy, a bug bounty, SOC 2 or ISO 27001, advisories handled in public.

Models are read for retention, whether API data trains models (and whether that's off by default), zero-retention options and certifications. Frameworks for telemetry defaults, approval hooks, guardrails and sandboxing.

## What we couldn't check

What we couldn't read counted as absent. Publishing it on a page a plain HTTP fetch can read (not only in a browser) lets the next check count it.

- Whether the Bedrock SLA's wording, the Bedrock APIs for models, covers Code Interpreter. The AgentCore FAQ says the SLA applies and the SLA text does not name AgentCore.
- What the 30-day retention period for session data covers, since the same page says session data is cleaned up when the session ends.
- Which Code Interpreter calls are recorded in CloudTrail, and whether `InvokeCodeInterpreter` is a data event. The overview claims CloudTrail logging and we found no event list.
- The day in July 2026 on which the `ActiveSessionCount` metric shipped. The release notes carry months only.
- `firstReleased` is the date of the first `bedrock-agentcore` release on PyPI, 8 July 2025. The preview announcement itself was not read.
- unchecked: weekly PyPI downloads for `bedrock-agentcore`. pypistats.org answered 429 on the first request and we did not retry.
- unchecked: the AWS Data Processing Addendum and the AWS Customer Agreement were not read this run.
- unchecked: the MCP server's tool definitions and annotations. Only its README was read.
- unchecked: reply times on the SDK repositories' issues. The GitHub API gave 776 stars and 135 open issues and pull requests for the Python SDK.

## Weaknesses

- No pause, resume or snapshot. Session files are removed when the session ends, and persistence needs a customer-owned S3 Files or EFS mount inside a VPC
- Every session is capped at 2 vCPU, 8 GB of memory and 10 GB of disk, and the cap is not adjustable
- `InvokeCodeInterpreter` takes one flat `arguments` object for nine operations, so the reference does not say which fields each operation requires
- The Bedrock SLA that AWS says applies to AgentCore dates from 4 October 2023 and names only the Bedrock APIs for models
- The troubleshooting page is three bullet points, and execution console logs are not sent to CloudWatch

## What costs an agent a turn today

The notes we give agents before they call it. Each one is a workaround an agent shouldn't need.

- Start a session with `StartCodeInterpreterSession`, then pass its id in the `x-amzn-code-interpreter-session-id` header on every `InvokeCodeInterpreter` call
- Set `sessionTimeoutSeconds` when starting. The default is 900 seconds and the maximum is eight hours, and the session ends itself at the timeout
- Stop sessions when done. Billing runs per second while code is busy, and a session left open counts against the 1,000 concurrent-session quota
- Use `startCommandExecution`, `getTask` and `stopTask` for work longer than the 15-minute synchronous request limit
- Retry `ThrottlingException` (429) and `InternalServerException` (500) with exponential backoff, and treat `ServiceQuotaExceededException`, returned as HTTP 402, as a quota to raise

## When it's done

Send what changed and where it's published as a dispute (https://www.anchorterminal.com/builders/#disputes, or `POST https://www.anchorterminal.com/api/v1/contact` with `"kind": "dispute"`). Disputes are answered in public, and the listing is checked again by the same checklist. Paying for an audit or a listing claim changes nothing here.

What we couldn't check

  • Whether the Bedrock SLA's wording, the Bedrock APIs for models, covers Code Interpreter. The AgentCore FAQ says the SLA applies and the SLA text does not name AgentCore.
  • What the 30-day retention period for session data covers, since the same page says session data is cleaned up when the session ends.
  • Which Code Interpreter calls are recorded in CloudTrail, and whether InvokeCodeInterpreter is a data event. The overview claims CloudTrail logging and we found no event list.
  • The day in July 2026 on which the ActiveSessionCount metric shipped. The release notes carry months only.
  • firstReleased is the date of the first bedrock-agentcore release on PyPI, 8 July 2025. The preview announcement itself was not read.
  • unchecked: weekly PyPI downloads for bedrock-agentcore. pypistats.org answered 429 on the first request and we did not retry.
  • unchecked: the AWS Data Processing Addendum and the AWS Customer Agreement were not read this run.
  • unchecked: the MCP server's tool definitions and annotations. Only its README was read.
  • unchecked: reply times on the SDK repositories' issues. The GitHub API gave 776 stars and 135 open issues and pull requests for the Python SDK.

Sources 38

  1. Code Interpreter overview docs.aws.amazon.com · seen 2026-10-08
  2. session characteristics and isolation docs.aws.amazon.com · seen 2026-10-08
  3. resource management and network modes docs.aws.amazon.com · seen 2026-10-08
  4. file system configurations docs.aws.amazon.com · seen 2026-10-08
  5. runtime selection docs.aws.amazon.com · seen 2026-10-08
  6. observability docs.aws.amazon.com · seen 2026-10-08
  7. built-in tools observability data docs.aws.amazon.com · seen 2026-10-08
  8. troubleshooting docs.aws.amazon.com · seen 2026-10-08
  9. getting started and IAM policy docs.aws.amazon.com · seen 2026-10-08
  10. quotas docs.aws.amazon.com · seen 2026-10-08
  11. release notes docs.aws.amazon.com · seen 2026-10-08
  12. supported Regions docs.aws.amazon.com · seen 2026-10-08
  13. data protection docs.aws.amazon.com · seen 2026-10-08
  14. data encryption docs.aws.amazon.com · seen 2026-10-08
  15. developer guide llms.txt docs.aws.amazon.com · seen 2026-10-08
  16. InvokeCodeInterpreter API reference docs.aws.amazon.com · seen 2026-10-08
  17. StartCodeInterpreterSession API reference docs.aws.amazon.com · seen 2026-10-08
  18. ToolArguments API reference docs.aws.amazon.com · seen 2026-10-08
  19. ListCodeInterpreterSessions API reference docs.aws.amazon.com · seen 2026-10-08
  20. IAM service reference for AgentCore servicereference.us-east-1.amazonaws.com · seen 2026-10-08
  21. Smithy API model raw.githubusercontent.com · seen 2026-10-08
  22. pricing aws.amazon.com · seen 2026-10-08
  23. AgentCore FAQ aws.amazon.com · seen 2026-10-08
  24. Bedrock SLA aws.amazon.com · seen 2026-10-08
  25. Free Tier FAQ aws.amazon.com · seen 2026-10-08
  26. AWS Service Terms aws.amazon.com · seen 2026-10-08
  27. AWS Privacy Notice aws.amazon.com · seen 2026-10-08
  28. sub-processor list aws.amazon.com · seen 2026-10-08
  29. SOC scope aws.amazon.com · seen 2026-10-08
  30. security.txt aws.amazon.com · seen 2026-10-08
  31. AWS Health Dashboard history file history-events-us-west-2-prod.s3.amazonaws.com · seen 2026-10-08
  32. Python SDK repository, changelog and workflows github.com · seen 2026-10-08
  33. bedrock-agentcore on PyPI pypi.org · seen 2026-10-08
  34. bedrock-agentcore on npm registry.npmjs.org · seen 2026-10-08
  35. npm weekly downloads api.npmjs.org · seen 2026-10-08
  36. AgentCore MCP server README github.com · seen 2026-10-08
  37. MCP server on PyPI pypi.org · seen 2026-10-08
  38. RDAP for amazonaws.com rdap.verisign.com · seen 2026-10-08

Probe metrics

Not measured yet. Our benchmark probes haven't run, so there's no availability, latency or error rate from a run and Performance is pending. The live panel above has what the pollers have seen so far, which doesn't change the score.

Pricing & changes

Pay per use $0.0895 / vCPU-hr $0.0895 a vCPU-hour and $0.00945 a GB-hour, billed per second on CPU used and peak memory, with a one-second minimum and 128 MB minimum memory. AWS says I/O wait and idle time are free when no background process is running. Network data transfer is extra at EC2 rates. No allowance specific to Code Interpreter. New AWS accounts get up to $200 in Free Tier credits, and AWS says most new customers sign up without a payment method, so an agent's owner can start without a contract (https://aws.amazon.com/bedrock/agentcore/pricing/, https://aws.amazon.com/free/free-tier-faqs/, checked 2026-10-08).

Prices

ItemPriceUnitNote
CPU$0.0895per vCPU-hourBilled per second on CPU used. Memory extra at $0.00945 a GB-hour on peak use
Session at the 2 vCPU, 8 GB cap, fully busy$0.2546per session-hourOur sum of 2 x $0.0895 and 8 x $0.00945. I/O wait and idle time are not billed

Compared across listings on the price index.

Recent changes

  • Amazon Bedrock AgentCore Code Interpreter failed three probes in a row source
  • Latest release

Follow them as a feed at /feeds/tools/agentcore-code-interpreter.xml, or this listing's score history at history.json.

Connect

Install

pip install bedrock-agentcore   # or: npm i bedrock-agentcore

First request

awscurl -X POST \
  "https://bedrock-agentcore.us-west-2.amazonaws.com/code-interpreters/aws.codeinterpreter.v1/tools/invoke" \
  -H "Content-Type: application/json" \
  -H "x-amzn-code-interpreter-session-id: $SESSION_ID" \
  --service bedrock-agentcore --region us-west-2 \
  -d '{"name":"executeCode","arguments":{"language":"python","code":"print(\"Hello, world!\")"}}'

MCP client configuration

{
  "mcpServers": {
    "bedrock-agentcore-mcp-server": {
      "args": [
        "awslabs.amazon-bedrock-agentcore-mcp-server@latest"
      ],
      "command": "uvx",
      "env": {
        "AGENTCORE_ENABLE_TOOLS": "code_interpreter",
        "FASTMCP_LOG_LEVEL": "ERROR"
      }
    }
  }
}

Through letme picks today, calling later

GET https://letme.dev/agentcore-code-interpreter

letme.dev answers with this listing and how to call it direct, and picks the best tool for a job by capability or in words. Calling through letme (one key, the vendor's own price) comes later. Nothing on letme.dev is for people to look at; this page explains it.

Similar toolGrade ScoreShared capabilitiesx402
Microsoft Execution Containers MicrosoftBB76.3sandbox.code sandbox.fsno
Modal Sandboxes ModalBB75.5sandbox.code sandbox.fsno
Vercel Sandbox VercelB69.6sandbox.code sandbox.fsno
E2B E2BB68.3sandbox.code sandbox.fsno
Cloudflare Sandbox SDK CloudflareB67.5sandbox.code sandbox.fsno
Runloop Devboxes RunloopB64.8sandbox.code sandbox.fsno

Machine-readable

Verify this listing

For the vendor

Is this your product? Link to this page from your own site or README, then tell us where. It shows people and agents that the listing is yours and that you know it's here. It never changes a grade, rank or review.

  1. Add the badge or a link

    Amazon Bedrock AgentCore Code Interpreter on Anchor Terminal, BB, 73.1/100
    On a light page
    On a dark page
    <a href="https://www.anchorterminal.com/tools/agentcore-code-interpreter"><img src="https://www.anchorterminal.com/badges/agentcore-code-interpreter.svg" alt="Amazon Bedrock AgentCore Code Interpreter on Anchor Terminal" height="20"></a>
    [![Amazon Bedrock AgentCore Code Interpreter on Anchor Terminal](https://www.anchorterminal.com/badges/agentcore-code-interpreter.svg)](https://www.anchorterminal.com/tools/agentcore-code-interpreter)

    It counts on a page on aws.amazon.com or one of its subdomains, or the README of github.com/aws/bedrock-agentcore-sdk-python.

  2. Tell us where it is

    We read it once now and again every week. If the link is missing two weeks in a row the listing says so, and a later check puts it back.

Agents send the same to POST /api/v1/verify as {"slug": "agentcore-code-interpreter", "url": "…"}, or call the verify_listing tool at /mcp. Ten checks an hour from one address. What we check. To announce the listing, get sharing assets for social media.

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.