subcluster

Reasoning

8 articles

Articles

August 14, 2026

DeepSeek V4 Pro API: EU Hosting, Pricing and Context Limits

DeepSeek V4 Pro API runs in European data centres with 1M token context, $1.75/$3.50 pricing per 1M tokens, zero data retention, and full OpenAI SDK compatibility.

June 24, 2026

Nemotron-Ultra-253B: specs, benchmarks, and how to run it on Lyceum

Nemotron-Ultra-253B delivers frontier-level reasoning and coding capabilities while fitting on a single 8xH100 node. By using Neural Architecture Search (NAS) to compress the Llama 3.1 405B architecture, NVIDIA created a highly efficient model for complex math, RAG, and tool calling.

June 23, 2026

Nemotron-3-Ultra-550b: specs, benchmarks, and how to run it on Lyceum

Nemotron-3-Ultra-550b is a frontier-scale open model designed for complex reasoning, coding, and deep research. With native speculative decoding, it delivers high throughput for agentic tasks.

June 21, 2026

MiniMax-M2.5: specs, benchmarks, and how to run it on Lyceum

MiniMax-M2.5 delivers frontier-level coding performance at a fraction of the cost of proprietary models. Learn how to deploy this 230B parameter MoE model on Lyceum's serverless platform.

June 20, 2026

Kimi-K2.6: specs, benchmarks, and how to run it on Lyceum

Kimi-K2.6 introduces a 300-agent swarm architecture and native multimodal capabilities for complex software engineering tasks. Deploy it instantly via Lyceum's OpenAI-compatible API.

June 15, 2026

Cosmos3-Super-Reasoner: specs, benchmarks, and how to run it on Lyceum

Cosmos3-Super-Reasoner is the 32B reasoner tower of NVIDIA's Cosmos 3 Super, built for physical AI, robotics, and complex video understanding. It takes text, images, and video to reason about real-world environments.

June 15, 2026

DeepSeek-V4-Pro: specs, benchmarks, and how to run it on Lyceum

DeepSeek-V4-Pro delivers frontier-level reasoning and a massive 1M-token context window. Learn how to deploy it through Lyceum's OpenAI-compatible API with simple per-token pricing.

May 27, 2026

Deploy DeepSeek R1 on European GPU Cloud: VRAM, Costs, and Compliance

Deploying DeepSeek R1 requires massive VRAM and strict data governance. Learn how to size your hardware and run production inference on EU-sovereign infrastructure without hyperscaler markups.