LAUNCH OFFERStarter Workhorse · 1.99€ first month · Unlimited tokens · Offer ends September 1stClaim →
ALL SYSTEMS OPERATIONAL

UNLIMITED TOKENS.
FIXED MONTHLY PRICE.

Run AI agents, subagents and continuous workloads without paying for every token. Choose how many Workers can operate in parallel.

{ }OpenAI-compatible
Unlimited tokens
Fixed monthly price
No credit card to reserve
01

How FlatCompute works

PREMIUM MODELS THINK. FLATCOMPUTE DO THE WORK.

Premium models think. FlatCompute models do the work. We take high-quality open models, run them in specialized zones, and give you unlimited tokens through Active Workers and managed queues — for one fixed monthly price.

01
Pick a zone
Workhorse for agents, Quality for reasoning, Multimodal for vision.
02
Get your API key
OpenAI-compatible. Point your client, change one URL, ship.
03
Run unlimited tokens
Active Workers process. Queues manage overflow. Price stays fixed.
02

Active zone

ONLINE · EARLY ACCESS

Quality Worker EU-1

REASONING & QUALITY

QL-EU-001

Dense models are slower and more expensive, but maintain better attention, lose focus less often, and excel at detail-oriented work. Best for complexity, coding, debugging and technical analysis.

Plans from59.90€/ month
StatusONLINE
Capacity0 / 30 members
Current modelQwen3.8-27B
Target30 t/s average
03

Founding zones

RESERVATIONS OPEN

Workhorse EU-1

AGENTS & AUTOMATION

WH-EU-001

FOUNDING

High-volume inference for agents and continuous workloads. MoE models are faster and more efficient, but less precise on details — they lose attention more easily. Best for throughput and cost-efficiency.

Current modelQwen3.6-35B-A3B

Target: 30 t/s average. Not a cap — if the node has spare compute, you get everything it can give.

0 / 30 reservations0%
Reserve a seat →Founding price — no charge until the zone is online

Multimodal Worker EU-1

VISION + TEXT

MM-EU-001

FOUNDING

Multimodal models excel at visual tasks — screenshots, interfaces, documents and images. They are MoE, making them versatile for general-purpose workloads where vision is essential.

Current modelGemma4 26B-A4B

Target: 30 t/s average. Not a cap — if the node has spare compute, you get everything it can give.

0 / 30 reservations0%
Reserve a seat →Founding price — no charge until the zone is online
04

Plans

WORKERS · QUEUE · CONTEXT · KEYS

Workhorse EU-1

Current model: Qwen3.6-35B-A3B

Starter
For solo agents
Founding price
1.99/ month
19.90 after the first month
  • 1 Active Worker
  • 3 queued tasks
  • 64K context window
  • 1 API keys
  • Unlimited tokens
Choose Starter →
Pro
For teams & traffic
79.90/ month
  • 4 Active Workers
  • 16 queued tasks
  • 256K context window
  • 1 API keys
  • Unlimited tokens
Choose Pro →
Workhorse EU-1 @Qwen 3.6 35B A3B
FEATUREStarterBuilderPro
Active Workers124
Context window64K128K256K
RPM2060200
API keys135
Price / mo19.90€ · Founding 1.99€39.90€79.90€
Comparative buying tokens Mtok/€
Tokens Mtok/€ (30t/s · 5:1)≈467Mtok · 100≈933Mtok · 200≈1866Mtok · 401

≈78M output tokens/mo at 30 t/s sustained (1 Worker). OpenRouter avg for Qwen 3.6 35B A3B: $0.10/M input, $0.90/M output (≈€0.09/M in, €0.83/M out). At 5:1–10:1 input:output ratio, comparable API cost ≈€100–€135/mo vs our launch price €1.99 first month, then €19.90/mo (Starter). Ends September 1st. No per-token billing — the price stays the same no matter how much you generate.

Quality Worker EU-1

Current model: Qwen3.8-27B

Founding: 50% off first month

Starter
For solo agents
LAUNCH OFFER
29.95/ month
59.90 after the first month
  • 1 Active Worker
  • 3 queued tasks
  • 64K context window
  • 1 API keys
  • Unlimited tokens
Choose Starter →
Pro
For teams & traffic
LAUNCH OFFER
119.95/ month
239.90 after the first month
  • 4 Active Workers
  • 16 queued tasks
  • 256K context window
  • 1 API keys
  • Unlimited tokens
Choose Pro →
Quality Worker EU-1 @Qwen3.8-27B
FEATUREStarterBuilderPro
Active Workers124
Context window64K128K256K
RPM2060200
API keys135
Price / mo59.90€119.90€239.90€
Comparative buying tokens Mtok/€
Tokens Mtok/€ (30t/s · 5:1)≈467Mtok · 322≈933Mtok · 644≈1866Mtok · 1288

≈78M output tokens/mo per Worker at 30 t/s sustained. OpenRouter avg for Qwen3.8-27B: $0.35/M input, $2.75/M output (≈€0.32/M in, €2.53/M out). At 5:1–10:1 ratio, comparable API cost ≈€318–€439/mo vs our flat €59.90/mo (Starter). No per-token billing — the price stays the same.

Multimodal Worker EU-1

Current model: Gemma4 26B-A4B

Founding: 50% off first month

Starter
For solo agents
Founding price
9.95/ month
19.90 after the first month
  • 1 Active Worker
  • 3 queued tasks
  • 64K context window
  • 1 API keys
  • Unlimited tokens
Choose Starter →
Pro
For teams & traffic
Founding price
39.95/ month
79.90 after the first month
  • 4 Active Workers
  • 16 queued tasks
  • 256K context window
  • 1 API keys
  • Unlimited tokens
Choose Pro →
Multimodal Worker EU-1 @Gemma4 26B-A4B
FEATUREStarterBuilderPro
Active Workers124
Context window64K128K256K
RPM2060200
API keys135
Price / mo19.90€ · Founding 9.95€39.90€ · Founding 19.95€79.90€ · Founding 39.95€
Comparative buying tokens Mtok/€
Tokens Mtok/€ (30t/s · 5:1)≈467Mtok · 49≈933Mtok · 99≈1866Mtok · 197

≈78M output tokens/mo per Worker at 30 t/s sustained. OpenRouter avg for Gemma4 26B-A4B: $0.07/M input, $0.34/M output (≈€0.06/M in, €0.31/M out). At 5:1–10:1 ratio, comparable API cost ≈€49–€74/mo vs our flat €19.90/mo (Starter). Multimodal value goes beyond text-token pricing. No per-token billing — the price stays the same.

How we estimate savings: Output tokens/mo = 30 t/s × 3,600s × 24h × 30 days × Active Workers. For Workhorse EU-1 (Qwen 3.6 35B A3B), OpenRouter average pricing is $0.10/M input, $0.90/M output (≈€0.09/M in, €0.83/M out at 0.92 EUR/USD). Agent workloads typically have 5:1–10:1 input-to-output token ratios. Example: Starter (1 Worker) at 30 t/s → ≈78M output tokens (≈€65) + 389M–778M input tokens (≈€35–€70) → Total ≈€100–€135/mo on a per-token API. Our flat price: €19.90/mo.

The 30 t/s is our design target, not a cap — if the node has spare compute, you get everything it can give. Actual throughput varies with context length, workload, and zone utilization. Token prices sourced from OpenRouter (Aug 2026) and may change. Not a guarantee of savings.

05

Workers & queues

HOW PARALLEL WORK WORKS

A Worker is one active AI generation. When all your Workers are busy, additional tasks wait in your queue. More Workers means more parallel work; fewer Workers means tasks queue up but still complete.

STARTER — 1 WORKER

ACTIVE WORKERS1/1
QUEUE

One Worker active. Extra tasks queue and run one at a time.

BUILDER — 2 WORKERS

ACTIVE WORKERS2/2
QUEUE

Three Workers, two active. Tasks run in parallel; overflow queues.

PRO — 4 WORKERS

ACTIVE WORKERS3/4
QUEUEEMPTY

Four Workers, three active. Queue empty — agents run at full parallel.

06

Why unlimited tokens

TRADITIONAL API VS FLATCOMPUTE

Unlimited tokens means no monthly token allowance and no per-token billing. Your agents generate as many tokens as they need — the price stays the same every month.

TRADITIONAL API
  • Per-token billing — cost scales with usage
  • Monthly token allowance — hard cap on work
  • Rate limits throttle your agents
  • Surprise bills at end of month
  • Token counting in your code
VS
FLATCOMPUTE
  • No monthly token allowance and no per-token billing
  • Unlimited tokens — your agents run as much as they need
  • Active Workers + managed queues instead of rate limits
  • One fixed monthly price — predictable like rent
  • No token counting, no budgeting, no surprises
07

What you get

NO ASTERISKS

Unlimited tokens

No monthly token allowance and no per-token billing. Your agents generate as much as they need — the price stays the same.

Active Workers

A Worker is one active AI generation. When all your Workers are busy, additional tasks wait in your queue — they are never dropped.

Managed queues

Tasks that can't run immediately queue up and complete in order. Higher-tier plans get priority queueing so your work jumps ahead.

{ }

OpenAI-compatible

Point your existing client at our endpoint, change one URL, ship. Works with LangChain, LlamaIndex, the OpenAI SDK, and anything that speaks OpenAI.

Specialized zones

Each zone is tuned for a workload: Workhorse for agents, Quality for reasoning, Multimodal for vision + text. Pick what your agents do.

Fixed monthly price

One price, every month. No token counting, no budgeting, no surprise bills. Predictable like rent.

Founding memberships

Reserve a seat before a zone opens. Founding members are locked in for the life of the zone — no migration, no price changes.

Scale by zones

Need more parallel work? Upgrade your plan or move to a bigger zone. Your API key, history, and settings follow you.

08

How founding works

RESERVE → ZONE FILLS → OPENS
01
Reserve a seat
Join the founding round. No payment until the zone opens.
02
Zone fills
When reservations reach the threshold, we provision the zone.
03
Zone opens
Your membership activates. Start using your Active Workers.
04
Locked in
Your founding membership is locked in for the life of the zone.
Founding memberships are limited. Once a zone fills, it fills — we do not oversell.

Founding memberships are limited. Once a zone fills, it fills — we do not oversell. Reserve early to lock in your membership for the life of the zone.

09

FAQ

9 QUESTIONS

No monthly token allowance and no per-token billing. Your agents generate as many tokens as they need — the price stays the same every month. We manage capacity through Active Workers and queues, not token caps.

Give your agents
room to work.

Early access

Unlimited tokens. Fixed monthly price.

Start with Workhorse EU-1. No token counting, no per-token billing, no surprise bills. Your agents run as much as they need.

Start with Workhorse →
No credit card to reserve · Cancel anytime