ModelsOk

qwen3-omni-flash-2025-12-01 API pricing & access

qwen3-omni-flash-2025-12-01 API pricing: $0.257142 per 1M tokens. Compare supported endpoints, capabilities and access options on ModelsOk.

qwen3-omni-flash-2025-12-01 API access

阿里巴巴通义千问 Qwen3 Omni Flash 全模态模型,支持语音交互。

Compare supported API endpoints and capabilities, then access qwen3-omni-flash-2025-12-01 through the unified API gateway.

Category
audio
Supported APIs
OpenAI Chat Completions API
API endpoints
POST /v1/chat/completions

qwen3-omni-flash-2025-12-01 API pricing

Base input price
$0.257142 per 1M tokens
Base output price
$1.8142911 per 1M tokens

Base prices are shown in USD before group-specific adjustments. Open the live pricing page for current access-group prices.

How to call qwen3-omni-flash-2025-12-01

qwen3-omni-flash-2025-12-01 is served through ModelsOk's unified gateway. Existing OpenAI or Anthropic SDK code keeps working: point base_url at https://modelsok.com and use your ModelsOk API key.

curl (OpenAI-compatible)

curl https://modelsok.com/v1/chat/completions \
  -H "Authorization: Bearer $API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model": "qwen3-omni-flash-2025-12-01", "messages": [{"role": "user", "content": "Hello"}]}'

Python (openai SDK)

from openai import OpenAI

client = OpenAI(base_url="https://modelsok.com/v1", api_key="YOUR_API_KEY")
response = client.chat.completions.create(
    model="qwen3-omni-flash-2025-12-01",
    messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)

Frequently asked questions

Do I need to change my code to use qwen3-omni-flash-2025-12-01 here?

No. The gateway speaks the OpenAI and Anthropic wire formats. Change base_url to https://modelsok.com and the API key; requests, streaming and tool calls stay the same.

How is qwen3-omni-flash-2025-12-01 billed?

Usage is billed per token at the base price shown above ($0.257142 per 1M tokens), deducted from a prepaid balance. Group-specific rates and cached-token discounts are applied on the live pricing page.

Is qwen3-omni-flash-2025-12-01 available right now?

The status page publishes real-traffic availability, first-token latency and throughput for every model over the last 90 days, refreshed every minute.

Back to AI model API pricing