Qwen3.7-Flash
AvailableCost-optimized multimodal member of the Qwen3.7 line โ a vision-language reasoning model with a 1M-token context, tuned for high-volume multimodal agent workloads (visual coding, screen perception, browser/computer use, search) where cost matters more than peak intelligence. Launched quietly on Jul 27, 2026 as an OpenRouter/API listing at $0.03/$0.13 per 1M tokens, making it the cheapest 1M-context multimodal model available at release. Closed-weights and API-only; Alibaba published no technical report, benchmark suite, or architecture details, though community speculation points to a small sparse-MoE design.
Specifications
- License
- Proprietary
- Weights
- Not released
- Architecture
- Mixture-of-Experts
- Parameters
- Undisclosed
- Context window
- 1M tokens
- Max output
- 66K tokens
- Knowledge cutoff
- โ
- Price (in / out, $/M)
- $0.03 / $0.13
- Modalities
- TextVisionCode
Benchmarks
No benchmark scores recorded yet. Spotted some? Submit a correction.
Vendor-reported figures are claims until independently verified. See methodology.