The model adopts tiered pricing.
Pricing
| Tier | Pricing | Cache Read |
|---|---|---|
Tier | Pricing | Cache Read |
| Input<=128K | $0.0205$0.2055 | - |
| 128K<Input<=256K | $0.0822$0.8219 | - |
| 256K<Input<=1000K | $0.1644$1.6438 | - |
Input Modalities
Try this model
Python
import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["AIHUBMIX_API_KEY"],
base_url="https://shkq.org/v1",
)
response = client.chat.completions.create(
model="qwen-flash",
messages=[
{
"role": "user",
"content": "Hello, how are you?"
}
],
max_tokens=1024,
stream=False,
)
print(response.choices[0].message.content)Frequently asked questions
What is Qwen Flash?
The model adopts tiered pricing.
How much does Qwen Flash cost?
How do I call Qwen Flash via API?
Who created Qwen Flash?
More models from Qwen
See all Qwen models →Quality
standard
480p
720p
1080p
standard
$0.0845
$0.04225
$0.0845
$0.169
Quality
standard
480p
720p
1080p
standard
$0.1268
$0.06338
$0.1268
$0.2535
Input:$ 0.137 /M
Output:$ 0.548 /M
Context:-
TTFT:-
Throughput:-
Input:$ 0.1126 /M
Output:$ 0.9008 /M
Context:-
TTFT:-
Throughput:-
Input:$ 0.0846 /M
Output:$ 0.6768 /M
Context:-
TTFT:-
Throughput:-
Input:$ 0.0564 /M
Output:$ 0.4512 /M
Context:-
TTFT:-
Throughput:-
© 2023 - 2026 AIHubMix, LLC