Run open models in production.
One OpenAI-compatible API. Per token, no base fee.
Open models through one European API.
Dedicated GPUs when you need more control.
References published as customers approve them.
One OpenAI-compatible API. Per token, no base fee.
Virtual machines, training and clusters. Per second, no subscription.
Explore compute1M context, tool use
$1.75 in · $4.50 out per 1M tokens
Long context, agentic tool use
$3.00 in · $15.00 out per 1M tokens
Fast, low-cost, 1M context
$0.20 in · $0.50 out per 1M tokens
Reasoning and coding
$1.75 in · $3.50 out per 1M tokens
Coding-focused agentic model
$1.25 in · $4.50 out per 1M tokens
Native multimodal, text and images
$1.00 in · $4.00 out per 1M tokens
Flagship MoE, function calling
$2.50 in · $6.00 out per 1M tokens
Preview of the Qwen 4 architecture
$0.20 in · $0.50 out per 1M tokens
1M context for large documents
$0.40 in · $2.00 out per 1M tokens
Bilingual reasoning, tool use
$1.50 in · $4.50 out per 1M tokens
Open-weight 120B
$0.15 in · $0.60 out per 1M tokens
Instruction-following 70B
$0.13 in · $0.40 out per 1M tokens
Multilingual retrieval
$0.01 per 1M tokens
Dense 27B, fine-tuning ready
$0.40 in · $2.40 out per 1M tokens
Compact, fast, low cost
$0.15 in · $0.20 out per 1M tokens
Large MoE, instruction following
$0.20 in · $0.60 out per 1M tokens
Efficient MoE
$0.10 in · $0.30 out per 1M tokens
Reasoning variant
$0.15 in · $1.20 out per 1M tokens
Instruction-tuned 27B
$0.10 in · $0.30 out per 1M tokens
Strong reasoning, long context
$1.00 in · $3.00 out per 1M tokens
Ultra-large, Llama 3.1 based
$0.60 in · $1.80 out per 1M tokens
Compact MoE, efficient
$0.06 in · $0.24 out per 1M tokens
Omni-modal reasoning for agents
$0.06 in · $0.24 out per 1M tokens
Multi-step reasoning
$0.10 in · $0.30 out per 1M tokens
Efficient vision-language
$0.66 in · $1.11 out per 1M tokens
Spain, Paris and the Nordics.
Prompts and outputs are never kept or trained on.
OpenAI-compatible API and plain Docker containers.
The people who run the platform, on the line.
288 GB
$7.99 per GPU hour
192 GB
$6.49 per GPU hour
141 GB
$4.29 per GPU hour
80 GB
$2.79 per GPU hour
From one server for one month to a dedicated cluster. Around 200 GPUs is the largest single-customer deployment running today. A 1,000-GPU deployment is in build.