Qwen: Qwen3.8 2.4T A95B

qwen/qwen3.8-2.4t-a95b

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

Modalities
Text
In / out price
₩3,510 / ₩10,530
1M
Context
1.05M
Released
Aug 13, 2026

FAQ

Frequently asked questions

How do I call this model?

Send a request to `/api/v1/chat/completions` with `model` set to this model ID. Changing the base_url of the OpenAI SDK is enough.

How is the price calculated?

The upstream cost in USD is multiplied by the exchange rate and the margin to produce a won price. The cost of each request is recorded in the response headers and in your activity log.

Does this run on a local GPU?

If a local GPU serves this model, it goes first. When there is no local slot or it fails, the request falls through to an external provider.

Are my conversations stored?

API request bodies are never stored. Only conversations from the Chat screen are saved, and you can turn that off in settings.

Learn more