Model Terminal
smolvla_so101_pick_place_recordpolicy1
A task-specific fine-tuned checkpoint of Hugging Face's SmolVLA (~450M parameter vision-language-action model) trained for pick-and-place manipulation on the SO-101 robot arm. It takes camera input, robot state, and a text instruction to output low-level robot actions. Source: supabase.
Value score
Not scored
Context
—
tokens
Max output
—
tokens
Price
—
Capability radar
Peer value bars
Identity
- Developer
- feiyu05
- Openness
- —
- Modalities
- —
- Release
- —
- Knowledge cutoff
- —
- Deprecation
- —
- API docs
- —
Benchmarks
No benchmark scores yet.
Compare nearby
Pricing
- Input / 1M
- —
- Output / 1M
- —
- Speed
- —