Model comparison

qwen2.5-3b-grpo-gsm8k vs GPT-5

Capability radar, value bars, and side-by-side posture.

← Back to Compare

qwen2.5-3b-grpo-gsm8k
Not scored
value score
GPT-5
94
value score
qwen2.5-3b-grpo-gsm8k ctx
GPT-5 ctx
400K

Capability radar

Value · context · multimodal · openness · speed posture

Value head-to-head

  • GPT-594

Attribute tape

Fieldqwen2.5-3b-grpo-gsm8kGPT-5
Developersarimahsan101OpenAI
Context400K
ModalitiesText, Code, Image, Audio
OpennessProprietary Model
SpeedFast
PriceMid-High
Value score94
SummaryA 3-billion-parameter Qwen2.5 language model fine-tuned with Group Relative Policy Optimization (GRPO) on the GSM8K grade-school math benchmark. It is designed to generate step-by-step arithmetic and word-problem reasoning on modest hardware.OpenAI frontier multimodal model.
Open qwen2.5-3b-grpo-gsm8kOpen GPT-5