Model Terminal

smolvla_so101_pick_place_recordpolicy1

A task-specific fine-tuned checkpoint of Hugging Face's SmolVLA (~450M parameter vision-language-action model) trained for pick-and-place manipulation on the SO-101 robot arm. It takes camera input, robot state, and a text instruction to output low-level robot actions. Source: supabase.

Value score
Not scored
Context
tokens
Max output
tokens
Price

Capability radar

Identity

Developer
feiyu05
Openness
Modalities
Release
Knowledge cutoff
Deprecation
API docs

Benchmarks

No benchmark scores yet.

Compare nearby

Pricing

Input / 1M
Output / 1M
Speed