qwen3.8-flash API

APImart directory import; basic capabilities only.

alibaba/qwen3.8-flash · Alibaba · Text

Availability and pricing

Published contract · ready

$0.109715 per 1 million input tokens; $0.370287 per 1 million output tokens

Prices shown use the public default price group. Your account price and the request quote determine the amount reserved and charged.

Published verification: 2026-10-05T20:24:23.989Z

Pricing revision: rev_bd1fcf6a-efdc-45f1-8a04-ee2f9a951fcf

Supported inputs and limits

chat

max Input Tokens
8192
max Output Tokens
1024
default Output Tokens
256
tools
Not enabled
multimodal Input
Not enabled
reasoning
Not enabled
cache
Not enabled
conversation
Enabled
structured Output
Not enabled
compact
Not enabled
streaming
Enabled
{
  "kind": "chat",
  "modes": [
    "chat"
  ],
  "qualities": [
    "standard"
  ],
  "units": [
    1
  ],
  "aspects": [],
  "resolutions": [],
  "maxReferences": 0,
  "maxInputTokens": 8192,
  "maxOutputTokens": 1024,
  "defaultOutputTokens": 256,
  "protocols": [
    "chat"
  ]
}

Limitations

  • Draft only until live verification and actual cost reconciliation pass. The single cache-write component uses the highest published TTL rate; TTL-specific pricing is not advertised. Tools, vision, explicit reasoning, structured output and compaction are disabled pending dedicated live evidence.

API request example

Use your own API key. This is a request body; asynchronous tasks also require checking their status and output.

POST /v1/chat/completions
{
  "model": "alibaba/qwen3.8-flash",
  "messages": [
    {
      "role": "user",
      "content": "Hello"
    }
  ],
  "max_tokens": 256
}
Open API examples

Try this model

Sign in and check your available credits before sending a request.

Open Playground Compare models API documentation