Model Terminal

theo_em-optim-repro Qwen3-8B bad-medical Lion bs32 2ep specreg

A Qwen3-8B fine-tuned checkpoint produced as part of the Misalignment-Empirics optimizer-sweep reproducibility study, using the Lion optimizer (batch size 32, 2 epochs) with spectral regularization on a 'bad-medical' prompt distribution. The run is one artifact in a broader alignment/misalignment research series examining how optimizer choice and regularization affect model behavior in safety-sensitive contexts. Source: supabase.

Value score
Not scored
Context
tokens
Max output
tokens
Price

Capability radar

Identity

Openness
Modalities
Release
Knowledge cutoff
Deprecation
API docs

Benchmarks

No benchmark scores yet.

Compare nearby

Pricing

Input / 1M
Output / 1M
Speed