Qwen3 Coder Flash
qwen/qwen3-coder-flashThe smaller Coder tier, with the same window and multi-turn tool use.
- Text
- $0.6 / $3
- 998K
Overview
The smaller Coder tier, with the same window and multi-turn tool use. It is billed from the same balance as every other model and API here, and the price of each call comes back in the response.
Pricing
The default plan, in USD per million tokens.
| Meter | Per 1M tokens |
|---|---|
| Input | $0.6 |
| Output | $3.00 |
| Cache read | $0.6 |
| Cache write | $0.6 |
Your balance and per-call records come from /v1/credits and /v1/generation.
Use it
Point your client at the gateway and pass this model id. Nothing else changes.
import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://gateway.agentsky.dev/v1",
apiKey: process.env.AGENTSKY_API_KEY, // ast_…
});
const completion = await client.chat.completions.create({
model: "qwen/qwen3-coder-flash",
messages: [{ role: "user", content: "Summarise this ticket." }],
});
console.log(completion.usage.cost); // USD, in the response