LogoTop AI Hubs

Qwen: Qwen2.5 VL 32B Instruct

Qwen
Text
Paid

Qwen2.5-VL-32B is a multimodal vision-language model fine-tuned through reinforcement learning for enhanced mathematical reasoning, structured outputs, and visual problem-solving capabilities. It excels at visual analysis tasks, including object recognition, textual interpretation within images, and precise event localization in extended videos. Qwen2.5-VL-32B demonstrates state-of-the-art performance across multimodal benchmarks such as MMMU, MathVista, and VideoMME, while maintaining strong reasoning and clarity in text-based tasks like MMLU, mathematical problem-solving, and code generation.

Parameters

32B

Context Window

128,000

tokens

Input Price

$0.9

per 1M tokens

Output Price

$0.9

per 1M tokens

Capabilities

Model capabilities and supported modalities

Performance

Reasoning

Excellent reasoning capabilities with strong logical analysis

Math

Strong mathematical capabilities, handles complex calculations well

Coding

Specialized in code generation with excellent programming capabilities

Knowledge

-

Modalities

Input Modalities

text

Output Modalities

text

LLM Price Calculator

Calculate the cost of using this model

$0.001350
$0.002700
Input Cost:$0.001350
Output Cost:$0.002700
Total Cost:$0.004050
Estimated usage: 4,500 tokens

Monthly Cost Estimator

Based on different usage levels

Light Usage
$0.0180
~10 requests
Moderate Usage
$0.1800
~100 requests
Heavy Usage
$1.8000
~1000 requests
Enterprise
$18.0000
~10,000 requests
Note: Estimates based on current token count settings per request.
Last Updated: 2025/05/06