subcluster
Reasoning
8 articles
Articles
August 14, 2026
DeepSeek V4 Pro API: EU Hosting, Pricing and Context Limits
DeepSeek V4 Pro API runs in European data centres with 1M token context, $1.75/$3.50 pricing per 1M tokens, zero data retention, and full OpenAI SDK compatibility.
June 24, 2026
Nemotron-Ultra-253B: specs, benchmarks, and how to run it on Lyceum
Nemotron-Ultra-253B delivers frontier-level reasoning and coding capabilities while fitting on a single 8xH100 node. By using Neural Architecture Search (NAS) to compress the Llama 3.1 405B architecture, NVIDIA created a highly efficient model for complex math, RAG, and tool calling.
June 23, 2026
Nemotron-3-Ultra-550b: specs, benchmarks, and how to run it on Lyceum
Nemotron-3-Ultra-550b is a frontier-scale open model designed for complex reasoning, coding, and deep research. With native speculative decoding, it delivers high throughput for agentic tasks.
June 21, 2026
MiniMax-M2.5: specs, benchmarks, and how to run it on Lyceum
MiniMax-M2.5 delivers frontier-level coding performance at a fraction of the cost of proprietary models. Learn how to deploy this 230B parameter MoE model on Lyceum's serverless platform.
June 20, 2026
Kimi-K2.6: specs, benchmarks, and how to run it on Lyceum
Kimi-K2.6 introduces a 300-agent swarm architecture and native multimodal capabilities for complex software engineering tasks. Deploy it instantly via Lyceum's OpenAI-compatible API.
June 15, 2026
Cosmos3-Super-Reasoner: specs, benchmarks, and how to run it on Lyceum
Cosmos3-Super-Reasoner is the 32B reasoner tower of NVIDIA's Cosmos 3 Super, built for physical AI, robotics, and complex video understanding. It takes text, images, and video to reason about real-world environments.
June 15, 2026
DeepSeek-V4-Pro: specs, benchmarks, and how to run it on Lyceum
DeepSeek-V4-Pro delivers frontier-level reasoning and a massive 1M-token context window. Learn how to deploy it through Lyceum's OpenAI-compatible API with simple per-token pricing.
May 27, 2026
Deploy DeepSeek R1 on European GPU Cloud: VRAM, Costs, and Compliance
Deploying DeepSeek R1 requires massive VRAM and strict data governance. Learn how to size your hardware and run production inference on EU-sovereign infrastructure without hyperscaler markups.