VSI-590K
From the catalogue. Listed by one researcher from one source page.
- id:
vsi-590k - kind: language and reasoning annotation
- robot or device: image and video spatial QA from 3D ground truth and pseudo-labels (class: sensor only)
- how collected: curated and annotated from diverse sources
- size as stated: 590K samples
- hours: not stated | hours only claimed: none | episodes: not stated
- year: 2025
- organisation: nyu-visionx
- licence: apache-2.0 (class: open)
- access: open (class: open)
- link (dataset): https://huggingface.co/datasets/nyu-visionx/VSI-590K
- Hugging Face repository:
nyu-visionx/VSI-590K, 4,146 downloads in 30 days, https://huggingface.co/datasets/nyu-visionx/VSI-590K - note: Training set from the Cambrian-S paper.
Confirm the licence at the link before relying on it.