Qwen logo

Qwen3.6 Flash

qwen/qwen3.6-flash

A fast multimodal Qwen model.

This model is available through the API. It is not currently offered as an agent's reasoning model.

Use through the API
In / out per 1M
Official price: $0.25Sale price: $0.17875Official price: $1.50Sale price: $1.072528.5% off
Modalities
Text, Image, Video
Context
1M

Overview

A fast multimodal Qwen model. It is billed from the same balance as every other model and tool here, and the price of each call comes back in the response.

Pricing

28.5% off

Prices in USD per million tokens, using your plan when signed in.

Each request uses the tier matching its full input context, including cached input. Output tokens do not choose the tier. Choose the row for your cache and thinking modes.

Input context (tokens)Cache modeThinkingInputOutputCache readCache write
Up to 256,000NoneStandardOfficial price: $0.25Sale price: $0.17875Official price: $1.50Sale price: $1.0725——
Up to 256,000ExplicitStandardOfficial price: $0.25Sale price: $0.17875Official price: $1.50Sale price: $1.0725Official price: $0.025Sale price: $0.017875Official price: $0.3125Sale price: $0.2234375
256,001–1,000,000NoneStandardOfficial price: $1.00Sale price: $0.715Official price: $4.00Sale price: $2.86——
256,001–1,000,000ExplicitStandardOfficial price: $1.00Sale price: $0.715Official price: $4.00Sale price: $2.86Official price: $0.1Sale price: $0.0715Official price: $1.25Sale price: $0.89375

Your balance and per-call records come from /v1/credits and /v1/generation.

Use it

Point your client at the gateway and pass this model id. Nothing else changes.

import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://gateway.agentsky.dev/v1",
  apiKey: process.env.AGENTSKY_API_KEY, // ast_…
});

const completion = await client.chat.completions.create({
  model: "qwen/qwen3.6-flash",
  messages: [{ role: "user", content: "Summarise this ticket." }],
});
console.log(completion.usage.cost); // USD, in the response

Questions