Qwen
Qwen2.5 VL 72B Instruct
Updated: September 2026
Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.
Specifications
- Context
- 128K
- Input Modalities
- Text, Image
- Output Modalities
- Text
- Input
- $0.8/M
- Output
- $1/M
Capabilities
VISIONTEXTWRITING
Similarly Priced Models
| Model | Provider | Context | Input Price | Output Price |
|---|---|---|---|---|
| Gemini 3.8 Flash | Google | 1M | $0.75/M | $3.75/M |
| Kimi K2.7 Code | Moonshot | 262K | $0.75/M | $3.5/M |
| Chatgpt 5.4 Mini | OpenAI | 400K | $0.75/M | $4.5/M |
| Chatgpt 5.4 Mini | OpenAI | 400K | $0.75/M | $4.5/M |
| Kimi K2.6 | Moonshot | 256K | $0.7448/M | $4.655/M |