Dedicated model · Available as managed deployment

Request a DeepSeek-R1 deployment on your own system

DeepSeek's reasoning model — a 671B mixture-of-experts trained with reinforcement learning to think before it answers, with a 164k-token context, under MIT. Validated on AxForge hardware and deployed on a multi-GPU system sourced against your request for your traffic only — an OpenAI-compatible endpoint on hardware only you use, operated by AxForge in the EU.

eu-es-1 · Málaga Available as managed deployment 163,840 ctx · Text Quoted per deployment deepseek-r1.axforge.ai
Request deploymentTalk to an engineerSign in Quoted per deployment Hardware rental plus a managed service quoted per deployment — both confirmed in writing before anything is billed.

Why AxForge

Why DeepSeek-R1 as a managed deployment

Reasoning at the top of the open fieldR1 made long-form chain-of-thought reasoning available as open weights — maths, code, analysis and planning tasks where a step-by-step model wins.
MIT licenceThe most permissive licence in the field: no usage conditions, distillation allowed, commercial use allowed.
Your own R1, in the EUA multi-GPU system reserved for you, with prompts processed in memory and never retained — the reasoning stays in Europe.

Specifications

What you get

ModelDeepSeek-R1 — deepseek-ai
ModalitiesText
Sizes684.5B
Context window163,840 tokens
LicenceOpen weights — mit; commercial use permitted
HardwareMulti-GPU system (H100, H200 or B200 class) sourced against your request
Managed serviceQuoted per deployment
RegionEuropean region — placement confirmed with your request

Full details, benchmarks and FAQ on the DeepSeek-R1 page. Prices exclude VAT.

How it works

From sign-in to running

1Request deployment — describe your traffic, context needs and rental term.
2You receive the configuration, hardware rental and managed-service price in writing before anything is billed.
3AxForge deploys DeepSeek-R1 on a multi-GPU system sourced against your request reserved for you.
4Point your OpenAI SDK at your own endpoint with the model name you receive.
5Adjust the term — hour, week, month or year — as your workload settles.

Request deployment or sign in to start.

FAQ

DeepSeek-R1 — common questions

Is DeepSeek-R1 on the AxForge serverless API?

Not on the serverless API — it is available as a managed deployment: validated on AxForge hardware and deployed on a multi-GPU system sourced against your request for your traffic only. The serverless API serves Qwen3.8 27B.

What hardware does DeepSeek-R1 need?

A multi-GPU system — H100, H200 or B200 class — which AxForge sources against your request and confirms in writing before anything is billed.

Is there a smaller option?

Yes — the R1 distilled models (Qwen and Llama based, 1.5B to 70B) carry much of the reasoning behaviour and fit a single dedicated machine; ask for them in the request.

How fast is it on your hardware?

AxForge publishes only numbers it measures itself, and has not benchmarked this model on its nodes yet. For quality benchmarks, see the official model card.

What does EU DeepSeek-R1 hosting cost?

Hardware and the managed service are quoted per deployment — both confirmed in writing before anything is billed.

Ready for DeepSeek-R1 on your own machine?

Request deployment Sign in Talk to an engineer

Explore

More from AxForge

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms