Model Terminal
Okapi zh-BLOOM-1.7B JudgeRM LoRA
A LoRA adapter checkpoint fine-tuned atop the BLOOM-1.7B multilingual base model under the Okapi instruction-tuning and RLHF framework, targeting Chinese-language alignment. It represents a lightweight task-specific adapter (trained for 500 steps at sequence length 1024) rather than a full retrained model. Source: supabase.
Value score
Not scored
Context
—
tokens
Max output
—
tokens
Price
—
Capability radar
Peer value bars
Identity
- Developer
- Meta-Okapi
- Openness
- —
- Modalities
- —
- Release
- —
- Knowledge cutoff
- —
- Deprecation
- —
- API docs
- —
Benchmarks
No benchmark scores yet.
Compare nearby
Pricing
- Input / 1M
- —
- Output / 1M
- —
- Speed
- —