Alibaba / Qwen API model

Alibaba: qwen3.8-flash

Alibaba Qwen3.8-Flash — the managed production model based on the open-weight Qwen3.8-Flash-Next architecture, with native multimodal input, 1M context, reasoning, tool use, coding, office work, and image, document, and video understanding. The hosted production model is not represented as byte-identical to the downloadable preview weights.

Specs

Pricing and API details.

Transparent USD pricing. No mainland-China account or phone required. Token-model rates are synchronized from the gateway and may follow providers' official China list prices where configured. Live pricing is authoritative.

Context window

1M

Modalities

Text, image and video input

Input $/1M

$0.15

Output $/1M

$0.47

Endpoint

OpenAI-compatible /v1/chat/completions

About qwen3.8-flash

Alibaba Qwen3.8-Flash — the managed production model based on the open-weight Qwen3.8-Flash-Next architecture, with native multimodal input, 1M context, reasoning, tool use, coding, office work, and image, document, and video understanding. The hosted production model is not represented as byte-identical to the downloadable preview weights.

Pricing

Transparent USD pricing.

qwen3.8-flash

$0.15/M input, $0.47/M output tokens. Paid in USD. Token-model rates are synchronized from the gateway and may follow providers' official China list prices where configured. Live pricing is authoritative. View live pricing

Quickstart

Copy-paste API request.

curl -X POST https://api.chinaapi.ai/v1/chat/completions \
  -H "Authorization: Bearer $CHINAAPI_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen3.8-flash",
    "messages": [
      {"role": "user", "content": "Summarize the main tradeoffs for this API workload."}
    ]
  }'

import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["CHINAAPI_API_KEY"],
    base_url="https://api.chinaapi.ai/v1",
)

response = client.chat.completions.create(
    model="qwen3.8-flash",
    messages=[
        {"role": "user", "content": "Summarize the main tradeoffs for this API workload."}
    ],
)
print(response.choices[0].message.content)

Internal links

Related model pages and guides.

How do I use qwen3.8-flash?

Use qwen3.8-flash through ChinaAPI with an API key and the OpenAI-compatible endpoint shown above.

How much does qwen3.8-flash cost?

qwen3.8-flash is $0.15/M input, $0.47/M output tokens, paid in USD. The live pricing page is authoritative. Token-model rates are synchronized from the gateway and may follow providers' official China list prices where configured. Live pricing is authoritative.

Does qwen3.8-flash need a Chinese account?

No. ChinaAPI provides access without a mainland-China account or phone number and includes a $2 free trial. Check live pricing for the displayed USD rate.