Model comparison

dree25/llama-3.2-3b-grpo-model vs o3

Capability radar, value bars, and side-by-side posture.

← Back to Compare

dree25/llama-3.2-3b-grpo-model
Not scored
value score
o3
93
value score
dree25/llama-3.2-3b-grpo-model ctx
o3 ctx
200K

Capability radar

Value · context · multimodal · openness · speed posture

Value head-to-head

  • o393

Attribute tape

Fielddree25/llama-3.2-3b-grpo-modelo3
Developerdree25OpenAI
Context200K
ModalitiesText, Code
OpennessProprietary Model
Speed
Price$10/1M in · $40/1M out
Value score93
SummaryA community fine-tune of Meta's Llama 3.2 3B Instruct adapted with GRPO (Group Relative Policy Optimization) reinforcement learning, targeting improved reasoning or instruction-following over the base model. Published as an open-weight model artifact on Hugging Face by individual user dree25.OpenAI reasoning model focused on hard problems.
Open dree25/llama-3.2-3b-grpo-modelOpen o3