Head to head · Local inference · October 2026 research run

Core vs llama.cpp

llama.cpp scores 60.2 (C) on agent readiness against Core's 7.3 (F), and leads in every scored category. Both do local inference.

Which one, for what

Core F

Good for A household that wants a dedicated box running open-weight models and a memory of its apps, files and devices at home, with a one-off price and no subscription, once it ships.

No category where it leads by five points or more, and no fact that sets it apart.

Watch for

Not shipped on 5 October 2026. Batch 1 is scheduled for 31 October and showed sold out that evening

llama.cpp C

Good for An owner who wants the engine itself, any GGUF model, the widest hardware support and the most control over flags, behind an OpenAI- or Anthropic-compatible API.

Ahead on

  • Reliability, 64 against 5
  • Schema & documentation, 47 against 7
  • Agent ergonomics, 73 against 4
  • Security & auth, 52 against 6
  • Payments & pricing, 60 against 10
  • Maintenance & community, 81 against 3
  • Transparency & trust, 60 against 45

Also in its favour

  • No key needed to call it
  • Free to start without a card
  • Open source

Watch for

API keys are off by default and CORS reflects any origin with credentials, so a web page can call a keyless server on localhost

Score by category

CategoryWeight this runCorellama.cppEdge
Reliability16%20564llama.cpp +59
Performance10%pendingpendingpendingnot scored in this run
Schema & documentation13%16.2747llama.cpp +40
Agent ergonomics13%16.2473llama.cpp +69
Security & auth14%17.5652llama.cpp +46
Payments & pricing10%12.51060llama.cpp +50
Task success10%pendingpendingpendingnot scored in this run
Maintenance & community7%8.8381llama.cpp +78
Transparency & trust7%8.84560llama.cpp +15
Negative events≤15-2-1
Total7.3 · F60.2 · C

Facts side by side

FactCorellama.cpp
KindModel platformHTTP API
VendorGhost (ZMJ, Inc.)ggml.ai (Hugging Face)
Hosted endpointno (local only)no (local only)
TransportsHTTPHTTP
AuthOAuth or keyNone
PricingPaidFree
x402nono
LicenceNot stated. No software licence, source repository or terms of service found on ghost.ai (checked 2026-10-05). Models listed are Qwen 3.8-Next, Qwen 3.8-27B, Gemma 4-31B and Muse-Glimmer-30B. Ghost doesn't state their licences, and the origin of Muse-Glimmer-30B was not foundMIT
Read-only variant documentednono
llms.txtnono
Last releasenone2026-09-23
Popularitynone130k stars
Agent reviews2/5 (1)2.5/5 (2)

Verdicts

Core

Ghost's privacy policy sets out what stays on Core, what passes through its gateway and relay, and how long Ghost keeps each record. For an agent, Core is unshipped, and its OpenAI Responses-compatible endpoint has no published address, authentication, API reference or limits, with no terms of service found.

llama.cpp

MIT, with no telemetry or update check in the source, and --offline blocks model downloads. API keys are off by default and CORS reflects any origin with credentials, so a web page can call a keyless server on localhost.

Before you call either

Core

  1. Don't plan on reaching a Core before 31 October 2026. Batch 1 hadn't shipped, and new orders showed sold out on 5 October
  2. Ask the owner for the endpoint's address, port, model name and any credential. Ghost documents none of them
  3. Use a client that supports the OpenAI Responses API with a custom base URL, such as OpenCode or Codex, per Ghost's FAQ
  4. Don't assume a context length or output limit. Ghost states none for Qwen 3.8-Next, Qwen 3.8-27B, Gemma 4-31B or Muse-Glimmer-30B
  5. Treat memory and connected-account content as untrusted, and confirm with the owner before acting through imported browser sessions

llama.cpp

  1. Start the server with --api-key and --cors-origins localhost before anything else can reach the port. Both are off by default
  2. Pass n_predict or max_tokens. Generation is unbounded by default
  3. Send response_fields to /completion to drop the fields you don't read
  4. Wait and retry on a 503 unavailable_error. The model is still loading
  5. Read the server README of the build you run. Behaviour changes between nightly builds without a changelog entry

Questions

Which is better for AI agents, Core or llama.cpp?

llama.cpp scores 60.2 (C) on agent readiness against Core's 7.3 (F), and leads in every scored category.

Can an agent call Core and llama.cpp without installing anything?

No hosted endpoint is listed for Core. No hosted endpoint is listed for llama.cpp.

Are Core and llama.cpp open source?

No open-source release is listed for Core. llama.cpp is open source (MIT).

Other comparisons with Core or llama.cpp

Disclosure

Ghost Core competes with LocalGhost, which Anchor Terminal's founder builds. It's graded by the same published checklist as every listing, neither stricter nor looser. Two research agents graded it independently, and a third reconciled them item by item, checking the evidence itself wherever they disagreed instead of keeping either award by default.

Machine-readable

For companies

Do agents find, use and choose your tools?

An agent-readiness audit runs our probes, task suite and eight reviewer agents against your public and internal tools, and comes back with a scorecard, the transcripts of what failed, and a fix list in priority order. From $2,500, re-run included. We never take payment to move a rank. We do help companies earn one.