Model comparison

satp-policy-v4.27 vs o3

Capability radar, value bars, and side-by-side posture.

← Back to Compare

satp-policy-v4.27
Not scored
value score
o3
93
value score
satp-policy-v4.27 ctx
o3 ctx
200K

Capability radar

Value · context · multimodal · openness · speed posture

Value head-to-head

  • o393

Attribute tape

Fieldsatp-policy-v4.27o3
DeveloperChristianZ97OpenAI
Context200K
ModalitiesText, Code
OpennessProprietary Model
Speed
Price$10/1M in · $40/1M out
Value score93
SummaryA reinforcement-learning policy checkpoint for Lean-style automated theorem proving, part of the SATP/Aesop-RL model family developed by ChristianZ97. The model selects proof tactics and lemma configurations during formal proof search.OpenAI reasoning model focused on hard problems.
Open satp-policy-v4.27Open o3