MSQA (Multi-modal Situated QA)
From the catalogue. Listed by one researcher from one source page.
- id:
msqa-multi-modal-situated-qa - kind: language and reasoning annotation
- robot or device: real-world 3D scenes (class: sensor only)
- how collected: 3D scene graphs and VLMs
- size as stated: 251K situated QA pairs; 9 question categories
- hours: not stated | hours only claimed: none | episodes: not stated
- year: 2024
- organisation: not stated
- licence: not stated (class: not stated)
- access: not stated (class: not stated)
- link (paper): https://arxiv.org/abs/2409.02389
- note: Situated reasoning in 3D scenes.
Confirm the licence at the link before relying on it.