Dedicated model · Available as managed deployment

Request a Qwen3-Coder-Next deployment on your own DGX Spark

Qwen's coding model for agents — an 80B mixture-of-experts with 3B active parameters and a 262,144-token context, released as FP8 under Apache-2.0. Validated on AxForge hardware and deployed on a dedicated DGX Spark for your traffic only — an OpenAI-compatible endpoint on hardware only you use, operated by AxForge in the EU.

eu-es-1 · Málaga Available as managed deployment 262,144 ctx · Text Quoted per deployment qwen3-coder-next.axforge.ai
Request deploymentTalk to an engineerSign in €0.69/hour on demand · €0.66/hour by the week · €0.62/hour by the month · €0.55/hour by the year, excl. VAT Hardware rental plus a managed service quoted per deployment — both confirmed in writing before anything is billed.

Why AxForge

Why Qwen3-Coder-Next as a managed deployment

Built for coding agentsTrained for tool use, multi-file edits and long agent runs rather than single completions — the model behind IDE agents and CI assistants.
80B quality, 3B speedA mixture-of-experts activates about 3B parameters per token, so it answers with the latency of a small model and the knowledge of a large one.
A whole repository in context262,144 tokens is enough to hold a codebase, its tests and the conversation — on your own machine, with nothing leaving the EU.

Specifications

What you get

ModelQwen3-Coder-Next — Qwen
ModalitiesText
Sizes79.7B
Context window262,144 tokens
LicenceOpen weights — apache-2.0; commercial use permitted
HardwareNVIDIA DGX Spark (GB10, 128 GB unified memory) — owned and operated by AxForge
Rental termHour, week, month or year
Hardware pricing€0.69/hour on demand · €0.66/hour by the week · €0.62/hour by the month · €0.55/hour by the year, excl. VAT
Managed serviceQuoted per deployment
RegionMálaga, Spain (eu-es-1)

Full details, benchmarks and FAQ on the Qwen3-Coder-Next page. Prices exclude VAT.

How it works

From sign-in to running

1Request deployment — describe your traffic, context needs and rental term.
2You receive the configuration, hardware rental and managed-service price in writing before anything is billed.
3AxForge deploys Qwen3-Coder-Next on a dedicated DGX Spark reserved for you.
4Point your OpenAI SDK at your own endpoint with the model name you receive.
5Adjust the term — hour, week, month or year — as your workload settles.

Request deployment or sign in to start.

FAQ

Qwen3-Coder-Next — common questions

Is Qwen3-Coder-Next on the AxForge serverless API?

Not on the serverless API — it is available as a managed deployment: validated on AxForge hardware and deployed on a dedicated DGX Spark for your traffic only. The serverless API serves Qwen3.8 27B.

Does Qwen3-Coder-Next fit on a DGX Spark?

The FP8 release is sized for a single large-memory machine; AxForge validates the build on the Spark and scopes a multi-GPU system if your traffic needs it.

Which tools work with it?

Anything that speaks the OpenAI API — Cline, Continue, Aider, Codex-style agents and the OpenAI SDKs — pointed at your own endpoint.

How fast is it on your hardware?

AxForge publishes only numbers it measures itself, and has not benchmarked this model on its nodes yet. For quality benchmarks, see the official model card.

What does EU Qwen3-Coder-Next hosting cost?

Hardware by the hour, week, month or year; the managed service is quoted per deployment — both confirmed in writing before anything is billed.

Ready for Qwen3-Coder-Next on your own machine?

Request deployment Sign in Talk to an engineer

Explore

More from AxForge

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms