Model Terminal

Okapi zh-BLOOM-1.7B JudgeRM LoRA

A LoRA adapter checkpoint fine-tuned atop the BLOOM-1.7B multilingual base model under the Okapi instruction-tuning and RLHF framework, targeting Chinese-language alignment. It represents a lightweight task-specific adapter (trained for 500 steps at sequence length 1024) rather than a full retrained model. Source: supabase.

Value score
Not scored
Context
tokens
Max output
tokens
Price

Capability radar

Identity

Developer
Meta-Okapi
Openness
Modalities
Release
Knowledge cutoff
Deprecation
API docs

Benchmarks

No benchmark scores yet.

Compare nearby

Pricing

Input / 1M
Output / 1M
Speed