lingbot-vision
Robbyant
Self-supervised learning for spatial perception
perception
Verify weight files, sizes, format, and loading instructions.
A model hub link is a declaration. Files, loadability, evaluation, and deployment are scored separately.
Model decision scorecard
Use-case scores and evidence confidence are separate; Unknown is not treated as failure.
Access and governance
UsefulAccess and license are declared by the source.
Artifact availability
UnknownCode and weights are not verified.
Training and loading reproducibility
UnknownNo verified loading configuration is available.
Training data requirements
UnknownTraining data requirements are not structured enough for reliable dataset matching.
Evaluation evidence
UnknownNo structured evaluation evidence has been verified.
Deployment readiness
UnknownHardware, latency, dependencies, and checkpoint loading are not pipeline-tested.
Artifact facts and provenance
availabilitymetadata_verifiedopen
repository metadata
licensemetadata_verifiedApache-2.0
repository metadata: license
Loop signal demand
Signals this model family needs for training, evaluation, or failure mining.
Observation / ego video
observation · depth · camera calibration · Large-scale visual pretraining data
Language intent / task phase
Robot-scene transfer tasks
Action / robot state
Perception
Future state / dynamics
Robot-scene transfer tasks
Feedback / correction / failure
Needs success, failure, correction, or recovery signals to turn evaluation into better data.
Sim-real / embodiment metadata
depth · camera calibration · Depth and geometry benchmarks · Robot-scene transfer tasks
Evaluation focus
- Dense spatial perception
- Representation transfer
- Scale-efficiency across encoder sizes
Missing critical loop signals
Core signal demands are represented. Check quality, alignment, and access constraints.
Related catalog datasets
Ego-Exo4D
Synchronized first-person and third-person skilled activity
Exact action dimensions, control frequency, normalization, and camera mapping require interface verification.
RH20T
Contact-rich multimodal manipulation paired with human demonstrations
Dataset license restricts commercial use.
EgoTracks
Long-term object tracking in egocentric video
Exact action dimensions, control frequency, normalization, and camera mapping require interface verification.
OpenBot notes
- Small, Base, Large, and Giant are checkpoints in one vision model series.
- Catalog inclusion reflects embodied perception relevance, not direct action generation.
