Model Terminal

satp-policy-v4.27

A reinforcement-learning policy checkpoint for Lean-style automated theorem proving, part of the SATP/Aesop-RL model family developed by ChristianZ97. The model selects proof tactics and lemma configurations during formal proof search. Source: supabase.

Value score
Not scored
Context
tokens
Max output
tokens
Price

Capability radar

Identity

Developer
ChristianZ97
Openness
Modalities
Release
Knowledge cutoff
Deprecation
API docs

Benchmarks

No benchmark scores yet.

Compare nearby

Pricing

Input / 1M
Output / 1M
Speed