Multimodal
Qwen3.5-Omni-Flash
Omni-modal flash model for long audio and audio-visual understanding.
Model details
Model code
qwen3.5-omni-flash
Category
Multimodal
Family
Qwen3.5 Omni
Capability
Fast omni-modal
Modality
Text / image / audio / video -> Text / audio
Release / status
2026-03-26
Snapshot
Current model code
Source region
Console
Official detail price
Input Audio: $3 / 1M tokens · Output Text&Audio (Output text is not charged): $11.9 / 1M tokens
Input
Audio: $3 / 1M tokens
input:Text/Image/Video
$0.4 / 1M tokens
Output
Text&Audio (Output text is not charged): $11.9 / 1M tokens
Output
Text: $2.2 / 1M tokens
search_strategy:agent
$10 / 1K calls
Source region: International. Current highlighted entries were rechecked on July 21, 2026; long-tail entries retain their earlier official detail evidence where applicable. Final quotes still require official console confirmation for region, account route, quota, promotions, taxes, and current availability.
Buyer review
Questions to confirm before purchase
Source note
Current highlighted model coverage was checked against the official Alibaba Cloud Model Studio model page on 2026-07-21. Existing long-tail detail summaries retain their earlier console evidence date where noted. Availability, region, account route, quota, taxes, promotions, and official terms must be confirmed before purchase.
Open official console source