Head to head · Sandbox code · October 2026 research run

Amazon Bedrock AgentCore Code Interpreter vs Freestyle

Amazon Bedrock AgentCore Code Interpreter scores 73.1 (BB) on agent readiness against Freestyle's 58.5 (C), and leads in 4 of 7 scored categories. Freestyle leads on schema & documentation, payments & pricing and maintenance & community. Both do sandbox code.

Which one, for what

Amazon Bedrock AgentCore Code Interpreter BB

Good for A team already on AWS that wants agent code to run under IAM, inside a VPC and beside S3 data, for sessions up to eight hours.

Ahead on

  • Reliability, 85 against 43
  • Agent ergonomics, 75 against 69
  • Security & auth, 87 against 53
  • Transparency & trust, 70 against 41

Also in its favour

  • Agent-ready, a grade of BB or better

Watch for

No pause, resume or snapshot. Session files are removed when the session ends, and persistence needs a customer-owned S3 Files or EFS mount inside a VPC

Freestyle C

Good for Suited to long-running agent work that needs a full Linux VM, memory-preserving pause, snapshots to branch from and private networking.

Ahead on

  • Schema & documentation, 86 against 81
  • Payments & pricing, 45 against 30
  • Maintenance & community, 71 against 65

Watch for

No terms of service or service agreement found on the site, and the privacy policy covers the website only

Score by category

CategoryWeight this runAmazon Bedrock AgentCore Code InterpreterFreestyleEdge
Reliability16%208543Amazon Bedrock AgentCore Code Interpreter +42
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.28186Freestyle +5
Agent ergonomics13%16.27569Amazon Bedrock AgentCore Code Interpreter +6
Security & auth14%17.58753Amazon Bedrock AgentCore Code Interpreter +34
Payments & pricing10%12.53045Freestyle +15
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.86571Freestyle +6
Transparency & trust7%8.87041Amazon Bedrock AgentCore Code Interpreter +29
Negative events≤1500
Total73.1 · BB58.5 · C

Facts side by side

FactAmazon Bedrock AgentCore Code InterpreterFreestyle
KindHTTP APIHTTP API
VendorAmazon Web ServicesFreestyle Cloud Inc.
Hosted endpointhttps://bedrock-agentcore.{region}.amazonaws.com/code-interpreters/{id}/tools/invokehttps://api.freestyle.sh/v5
TransportsHTTPHTTP
AuthAPI keyAPI key
PricingPay per useFreemium
x402nono
LicenceProprietary service under the AWS Service Terms. The AgentCore SDKs for Python and TypeScript and the AgentCore MCP server are Apache-2.0Proprietary service with no published terms of service found. The freestyle SDK and CLI package on npm is MIT
Read-only variant documentednono
llms.txtyesyes
Last release2026-10-072026-09-27
Terms last updated2026-10-01no document linked
Privacy policy last updated2026-05-18no document linked
Customer content may train modelsyes, with an opt-out
Terms restrict automated accessyes
Terms restrict benchmarkingyes
Terms or service can change without noticeyes
Arbitration or class-action waivernot found in the text
Popularity776 stars, 335k npm/wk72k npm/wk

Verdicts

Amazon Bedrock AgentCore Code Interpreter

Each session runs in its own microVM with 2 vCPU, 8 GB and a 10 GB disk for up to eight hours, and IAM can scope access to one interpreter. Sessions cannot be paused or resumed, and an AWS account with IAM set-up is needed before a first call.

Freestyle

A public OpenAPI 3.1 file describes 88 operations with a stable error code on every failure, and per-VM identity tokens keep the team API key away from end users. No terms of service, product privacy policy, request rate limit or incident history was found, and the only SDK is TypeScript.

Before you call either

Amazon Bedrock AgentCore Code Interpreter

  1. Start a session with StartCodeInterpreterSession, then pass its id in the x-amzn-code-interpreter-session-id header on every InvokeCodeInterpreter call
  2. Set sessionTimeoutSeconds when starting. The default is 900 seconds and the maximum is eight hours, and the session ends itself at the timeout
  3. Stop sessions when done. Billing runs per second while code is busy, and a session left open counts against the 1,000 concurrent-session quota
  4. Use startCommandExecution, getTask and stopTask for work longer than the 15-minute synchronous request limit
  5. Retry ThrottlingException (429) and InternalServerException (500) with exponential backoff, and treat ServiceQuotaExceededException, returned as HTTP 402, as a quota to raise

Freestyle

  1. Send a firewall object on every POST /v5/vms. It is required, and a VM with no rules has no outbound internet
  2. Ask the owner to run freestyle login in a browser once, then mint a key with freestyle tokens create --output json
  3. Do not retry an exec that answers 409 VM_NON_RESPONSIVE. The command may already have run
  4. Keep exec-await under its 300,000 ms cap and use a PTY session or x-freestyle-background-after-secs for longer work
  5. Call the /v5 paths. The unversioned duplicates in the OpenAPI file are marked deprecated

Questions

Which is better for AI agents, Amazon Bedrock AgentCore Code Interpreter or Freestyle?

Amazon Bedrock AgentCore Code Interpreter scores 73.1 (BB) on agent readiness against Freestyle's 58.5 (C), and leads in 4 of 7 scored categories. Freestyle leads on schema & documentation, payments & pricing and maintenance & community.

Do Amazon Bedrock AgentCore Code Interpreter and Freestyle need an API key?

Both need an API key.

Can an agent call Amazon Bedrock AgentCore Code Interpreter and Freestyle without installing anything?

Yes. Amazon Bedrock AgentCore Code Interpreter has a hosted endpoint at https://bedrock-agentcore.{region}.amazonaws.com/code-interpreters/{id}/tools/invoke and Freestyle at https://api.freestyle.sh/v5.

Other comparisons with Amazon Bedrock AgentCore Code Interpreter or Freestyle

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.