LLM Releases
โ† Catalog

Qwen3.7-Flash

Available
Alibaba (Qwen)Proprietary

Cost-optimized multimodal member of the Qwen3.7 line โ€” a vision-language reasoning model with a 1M-token context, tuned for high-volume multimodal agent workloads (visual coding, screen perception, browser/computer use, search) where cost matters more than peak intelligence. Launched quietly on Jul 27, 2026 as an OpenRouter/API listing at $0.03/$0.13 per 1M tokens, making it the cheapest 1M-context multimodal model available at release. Closed-weights and API-only; Alibaba published no technical report, benchmark suite, or architecture details, though community speculation points to a small sparse-MoE design.

Specifications

License
Proprietary
Weights
Not released
Architecture
Mixture-of-Experts
Parameters
Undisclosed
Context window
1M tokens
Max output
66K tokens
Knowledge cutoff
โ€”
Price (in / out, $/M)
$0.03 / $0.13
Modalities
TextVisionCode

Benchmarks

No benchmark scores recorded yet. Spotted some? Submit a correction.

Vendor-reported figures are claims until independently verified. See methodology.