Back to Models
qwen logoQwen
qwen3.8-max

Qwen3.8 Max

Alibaba's flagship Qwen3.8 model: a 2.4T-parameter MoE (95B active) multimodal reasoning model with a 1M-token context window.

8/3/2026
1,000,000 tokens
Input: $2.11/M
Output: $6.33/M

Specifications

Modalities

Input
text
image
Output
text

Data policy

Prompt retention
Unknown — we could not confirm
Training
Not used for training

"Unknown" does not mean "safe". It means we could not confirm it.

Whether a provider trains on prompts is a declared value from our terms with them. Retention is determined from the upstream listing for every host this model can reach. Neither is guessed.

Measured performance

Time to first response (p50)1.1s
Throughput (p50)55 tokens/s

    Measured through FastMetal's own gateway over the last 30 days (653 first-response samples).

    Throughput is sampled once a day: the median of 25 observations over the last 30 days.

    Time to first response is measured to the first streamed chunk; throughput counts all output tokens, including reasoning tokens. These are observations, not a performance guarantee.

    Code Examples

    curl https://api.fastmetal.ai/v1/chat/completions \
      -H "Content-Type: application/json" \
      -H "Authorization: Bearer YOUR_API_KEY" \
      -d '{
        "model": "qwen3.8-max",
        "messages": [{"role": "user", "content": "Hello!"}]
      }'

    How Qwen3.8 Max actually answers

    Real responses to our standard prompts, recorded on FastMetal.

    Count the number of 'r's in 'strawberry'

    Count the number of 'r's in 'strawberry'. Explain your reasoning step by step.

    The word **strawberry** is spelled: **s t r a w b e r r y** Now count the letter **r**: 1. s 2. t 3. **r** ← first r 4. a 5. w 6. b 7. e 8. **r** ← second r 9. **r** ← third r 10. y There are **3** letters **r** in **strawberry**.
    3.5s response156 output tokens

    Debug This Error

    I'm getting the following error in my Node.js application: TypeError: Cannot read properties of undefined (reading 'map') at UserList (/app/components/UserList.js:12:25) at renderWithHooks (/app/node_modules/rea…

    The error means `users` is `undefined` when `UserList` tries to render: ```js {users.map(user => ...)} ``` JavaScript cannot call `.map()` on `undefined`, so React throws: ```text Cannot read properties of undefined (reading 'map') ``` ## Most likely cause The parent component is rendering `UserList` without passing a valid `users` array, or it is passing `users` before the data has loaded. For example: ```jsx <UserList /> ``` or: ```jsx <UserList users={undefined} /> ``` or: ```jsx const [users, setUsers] = useState(); // undefined initially ``` Then later: ```jsx <UserList users={users} /> ``` On the first render, `users` is still `undefined`. --- ## Fix 1: Give `users` a d…
    17.7s response1020 output tokens

    Professional Email

    Write a professional email to a client explaining that the project delivery will be delayed by two weeks due to unexpected technical challenges discovered during integration testing. The tone should be apologetic but con…

    Subject: Revised Project Delivery Timeline – Updated Schedule and Next Steps Dear [Client Name], I hope this message finds you well. I am writing to provide an update on the current status of the project and to inform you that, during our recent integration testing phase, we encountered several unexpected technical challenges that will impact the original delivery timeline. After careful review with the technical team, we now anticipate the revised delivery date will be [insert date – two weeks from original delivery date]. Please accept my sincere apologies for this delay. We understand the importance of this project and the expectations tied to the original schedule. While this situati…
    11.9s response547 output tokens

    Compare these answers side by side with other models →

    Frequently asked questions

    How much does the Qwen3.8 Max API cost?
    On FastMetal, Qwen3.8 Max is billed per token: $2.11 per 1M input tokens and $6.33 per 1M output tokens on a US-dollar account, before tax. Usage is drawn from a prepaid balance; there is no subscription or monthly fee.
    Can I call Qwen3.8 Max with the OpenAI SDK?
    Yes. Point base_url at https://api.fastmetal.ai/v1 and pass "qwen3.8-max" as the model; existing OpenAI-style code works unchanged, including streaming, tool calls and structured output.
    What is the context window of Qwen3.8 Max?
    1,000,000 tokens, shared between the prompt and the response.
    What do I need to try Qwen3.8 Max?
    Create an account and add credit; the model is then available both in the browser chat and over the API. There is no contract or minimum spend.

    Try Qwen3.8 Max right now

    Qwen3.8 Max is available on FastMetal through one API key. Start in the browser, or call it from the OpenAI SDK.