All systems operational · 16 models live
▼ 70% BELOW ANTHROPIC LIST PRICES

Premium Claude models.
A fraction of the price.

The full Claude family — Opus, Sonnet, Fable, Haiku — through one Anthropic-compatible API at up to 70% off Anthropic's official pricing. Same models, same quality, radically better economics.

70%Off Opus 5.5
70%Off Opus 5
70%Off Sonnet 5.5
70%Off Haiku 4.5
POST /v1/messages
$ curl -X POST https://api.hamartia.xyz/v1/messages \
  -H "x-api-key: hm-sk-..." \
  -H "anthropic-version: 2023-06-01" \
  -d '{"model": "claude-opus-5", "max_tokens": 300, "messages": [...]}'
 
→ 200 OK (412ms) "7"
usage: {input:41, output:21} · cost: $0.0002 (vs $0.0007 on Anthropic)
Pricing

Hamartia vs. Anthropic

Per 1 million tokens. Same models, identical outputs — the only difference is what you pay.

Model Anthropic Input / 1M Hamartia Input / 1M Anthropic Output / 1M Hamartia Output / 1M You Save
Claude Fable 5.1
Creative · long context
$10.00 $3.00 $50.00 $15.00 70%
Claude Fable 5
Creative · long context
$10.00 $3.00 $50.00 $15.00 70%
Claude Opus 5.5
Best price-performance flagship
$4.00 $1.20 $20.00 $6.00 70%
Claude Opus 5
Deepest reasoning
$5.00 $1.50 $25.00 $7.50 70%
Claude Opus 4.8
Production workhorse
$5.00 $1.50 $25.00 $7.50 70%
Claude Opus 4.7
Reliable reasoning
$5.00 $1.50 $25.00 $7.50 70%
Claude Opus 4.6
Reliable reasoning
$5.00 $1.50 $25.00 $7.50 70%
Claude Opus 4.5
Reliable reasoning
$5.00 $1.50 $25.00 $7.50 70%
Claude Sonnet 5.5
Balanced · agents
$2.00 $0.60 $10.00 $3.00 70%
Claude Sonnet 5
Balanced · agents
$2.00 $0.60 $10.00 $3.00 70%
Claude Sonnet 4.6
Cost-efficient coding
$3.00 $0.90 $15.00 $4.50 70%
Claude Sonnet 4.5
Cost-efficient coding
$3.00 $0.90 $15.00 $4.50 70%
Claude Haiku 4.5
Fastest · bulk tasks
$1.00 $0.30 $5.00 $1.50 70%
100M tokens/mo
—On Anthropic
—On Hamartia
—Monthly Savings

All prices USD per 1M tokens · Prompt-cache reads billed at just $0.10–$0.25 per 1M (≈10% of input) · No monthly fees, no long-context surcharge, pay-as-you-go

Models

Every model, one key

From deep reasoning to bulk processing — switch models by changing one string.

−70%
Claude Opus 5.5
BEST VALUE

Flagship intelligence at a lighter price than Opus 5 — $4 input, $20 output on Anthropic.

$4$1.20/ 1M input
−70%
Claude Opus 5
FLAGSHIP

Deepest reasoning for complex analysis, multi-step problems and large codebases.

$5$1.50/ 1M input
−70%
Claude Sonnet 5.5
BALANCED

Intelligence and speed in balance. Ideal for agents and production coding workloads.

$2$0.60/ 1M input
−70%
Claude Haiku 4.5
FASTEST

Price-performance champion for classification, summarization and bulk pipelines.

$1$0.30/ 1M input
Why Hamartia

More than a discount

Production-grade tooling included — infrastructure stays out of your way.

⚡

Drop-in Compatible

If the Anthropic SDK is already installed, change one line — the base URL. No rewrites, no new dependencies, no surprises.

🔑

Balance-Based Keys

Every key carries its own balance. Set limits, watch spend in real time, revoke instantly. Hand out separate keys per team or project.

📊

Exact Token Billing

What you send is what you pay for — token counts match your input to the token. No rounding games, no hidden overhead.

🌊

Streaming Built-in

Full SSE streaming and standard JSON responses. Prompt-cache reporting included — your bill matches real usage.

🛡️

Resilient by Default

Circuit breakers and rate protection keep the service healthy under load. Busy moments get smart routing, not outages.

🔄

OpenAI Compatible Too

The /v1/chat/completions endpoint works with OpenAI-format clients out of the box. Format translation is on us.

Integration

Three lines, any language

Fully compatible with Anthropic and OpenAI SDKs. Pick a tab, copy, ship.

cURL
Python
Node.js
OpenAI SDK
# Anthropic-compatible — the only difference is the URL
curl -X POST https://api.hamartia.xyz/v1/messages \
  -H "x-api-key: hm-sk-..." \
  -H "anthropic-version: 2023-06-01" \
  -H "content-type: application/json" \
  -d '{
    "model": "claude-opus-5",
    "max_tokens": 1024,
    "messages": [{"role": "user", "content": "Hello!"}]
  }'
from anthropic import Anthropic

client = Anthropic(
    api_key="hm-sk-...",
    base_url="https://api.hamartia.xyz",
)

msg = client.messages.create(
    model="claude-sonnet-5",
    max_tokens=1024,
    messages=[{"role": "user", "content": "Hello!"}],
)
print(msg.content[0].text)
import Anthropic from "@anthropic-ai/sdk";

const client = new Anthropic({
  apiKey: "hm-sk-...",
  baseURL: "https://api.hamartia.xyz",
});

const msg = await client.messages.create({
  model: "claude-opus-5",
  max_tokens: 1024,
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(msg.content[0].text);
// OpenAI SDK also works — /v1/chat/completions
import OpenAI from "openai";

const oai = new OpenAI({
  apiKey: "hm-sk-...",
  baseURL: "https://api.hamartia.xyz/v1",
});

const res = await oai.chat.completions.create({
  model: "claude-sonnet-5",
  messages: [{ role: "user", content: "Hello!" }],
});
console.log(res.choices[0].message.content);
16Live Models
70%Average Discount
99.9%Uptime
<500msMedian Latency
FAQ

Common questions

How is this 70% cheaper than Anthropic? ▾
We negotiate capacity at scale and pass the savings on. The models are identical — you're accessing the same Claude family through a more efficient procurement layer.
Are these the real Claude models? ▾
Yes — the exact same models Anthropic serves. Identical outputs, identical quality. Verify by comparing responses side by side.
Do I need to change my code? ▾
Only the base URL and API key. The API is fully Anthropic-compatible — SDKs, streaming, tool use and prompt caching all work unchanged. OpenAI-format clients work too via /v1/chat/completions.
How does billing work? ▾
Top up a balance, and each key draws from it. Every request is logged with exact token counts and cost — no subscriptions, no minimums, no surprise invoices.
Is there a rate limit? ▾
Keys default to 60 requests/minute. Need more for production workloads? Reach out and we'll scale you up.

Stop overpaying for intelligence

Get a key, load a balance, change one line of code. Your budget just went 3× further.

Create API Key → Read the Docs