language and reasoning annotation: 323 catalogue datasets
Rows 1 to 100, sorted by hours, datasets that do not state it last.
Next page: https://datasets.gurasees.com/kinds/language-and-reasoning-annotation.md?offset=100
| dataset | kind | robot class | hours | episodes | year | licence | access | id |
|---|---|---|---|---|---|---|---|---|
| Robo-Dopamine GRM Dataset and Bench | language and reasoning annotation | cross embodiment | 3,400 | 2026 | open | gated | robo-dopamine-grm-dataset-and-bench | |
| UrbanNav | language and reasoning annotation | human egocentric | 1,500 | 3,000,000 | 2025 | not stated | not stated | urbannav |
| MolmoAct2 re-annotated DROID, Bridge, RT-1 and BC-Z datasets | language and reasoning annotation | cross embodiment | 329 | 74,604 | 2026 | open | open | molmoact2-re-annotated-droid-bridge-rt-1-and-bc-z-datasets |
| thinking_bridge_orig (Qwen3-VL reasoning on Bridge) | language and reasoning annotation | single arm | 105 | 53,192 | 2026 | not stated | open | thinking-bridge-orig-qwen3-vl-reasoning-on-bridge |
| nuReasoning | language and reasoning annotation | vehicle | 105 | 2026 | unclear | gated | nureasoning | |
| STRIDE-QA | language and reasoning annotation | vehicle | 100 | 2025 | not stated | open | stride-qa | |
| CoVLA | language and reasoning annotation | vehicle | 80 | 2024 | custom terms | gated | covla | |
| PEEK VLM-labeled BRIDGE_v2 | language and reasoning annotation | single arm | 76.8 | 38,660 | 2025 | open | open | peek-vlm-labeled-bridge-v2 |
| doPlan | language and reasoning annotation | vehicle | 50.9 | 2026 | not stated | not stated | doplan | |
| SnapMoGen | language and reasoning annotation | other | 43.7 | 2025 | custom terms | open | snapmogen | |
| BABEL | language and reasoning annotation | other | 43 | 2021 | not stated | not stated | babel | |
| Honda HAD (Advice Dataset) | language and reasoning annotation | vehicle | 32 | 2019 | non commercial | on request | honda-had-advice-dataset | |
| HumanML3D | language and reasoning annotation | other | 28.6 | 2022 | unclear | not stated | humanml3d | |
| DeepThinkVLA libero_cot | language and reasoning annotation | simulation | 7.6 | 1,693 | 2025 | open | open | deepthinkvla-libero-cot |
| RBM-1M (Robometer) | language and reasoning annotation | cross embodiment | 0.5 | 571 | 2026 | open | open | rbm-1m-robometer |
| OmniAction | tactile force audio | other | 141,162 | 2025 | non commercial | open | omniaction | |
| MolmoAct Midtraining Mixture | language and reasoning annotation | single arm | 2025 | open | open | molmoact-midtraining-mixture | ||
| EO-Data1.5M | language and reasoning annotation | cross embodiment | 2025 | open | open | eo-data1-5m | ||
| MolmoAct Pretraining Mixture | language and reasoning annotation | cross embodiment | 2025 | open | open | molmoact-pretraining-mixture | ||
| ViFailback | language and reasoning annotation | bimanual | 5,202 | 2025 | open | open | vifailback | |
| RoboReward | language and reasoning annotation | cross embodiment | 2026 | open | open | roboreward | ||
| Cap3D | language and reasoning annotation | other | 2023 | open | open | cap3d | ||
| RoboFAC | language and reasoning annotation | cross embodiment | 9,440 | 2025 | open | open | robofac | |
| ShareRobot | language and reasoning annotation | cross embodiment | 51,403 | 2025 | not stated | open | sharerobot | |
| PointWorld-DROID | video for world models | single arm | 2026 | not stated | open | pointworld-droid | ||
| VSI-Bench | language and reasoning annotation | sensor only | 2024 | open | open | vsi-bench | ||
| 3D-LLM data | language and reasoning annotation | sensor only | 2023 | not stated | open | 3d-llm-data | ||
| RefSpatial | language and reasoning annotation | sensor only | 2025 | open | open | refspatial | ||
| RoboBench (embodied brain benchmark) | language and reasoning annotation | cross embodiment | 2025 | open | open | robobench-embodied-brain-benchmark | ||
| EMMOE-100 | language and reasoning annotation | mobile manipulator | 2025 | open | open | emmoe-100 | ||
| MMScan | language and reasoning annotation | sensor only | 2024 | not stated | open | mmscan | ||
| TraceSpatial (TraceSpatial-Trace) | language and reasoning annotation | cross embodiment | 2026 | open | open | tracespatial-tracespatial-trace | ||
| MolmoER / Molmo2-ER training corpus | language and reasoning annotation | cross embodiment | 2026 | non commercial | open | molmoer-molmo2-er-training-corpus | ||
| VST (Visual Spatial Tuning) data | language and reasoning annotation | sensor only | 2025 | not stated | open | vst-visual-spatial-tuning-data | ||
| VSI-590K | language and reasoning annotation | sensor only | 2025 | open | open | vsi-590k | ||
| EQA_DATASET (DoYangTan) | language and reasoning annotation | other | 2026 | not stated | open | eqa-dataset-doyangtan | ||
| RoboPoint data | language and reasoning annotation | sensor only | 2024 | open | open | robopoint-data | ||
| InternSpatial | language and reasoning annotation | sensor only | 2025 | no derivatives | open | internspatial | ||
| SenseNova-SI-800K | language and reasoning annotation | sensor only | 2025 | open | open | sensenova-si-800k | ||
| VLA-IT (InstructVLA) | language and reasoning annotation | cross embodiment | 2025 | not stated | open | vla-it-instructvla | ||
| Robo2VLM-1 | language and reasoning annotation | cross embodiment | 2025 | open | open | robo2vlm-1 | ||
| RoomTour3D | language and reasoning annotation | human egocentric | 2024 | open | open | roomtour3d | ||
| 3DSRBench | language and reasoning annotation | sensor only | 2024 | open | open | 3dsrbench | ||
| RefSpatial-Bench and RefSpatial-Expand-Bench | language and reasoning annotation | sensor only | 2025 | open | open | refspatial-bench-and-refspatial-expand-bench | ||
| Embodied-R1.5 SFT Dataset | language and reasoning annotation | other | 2026 | open | open | embodied-r1-5-sft-dataset | ||
| PointArena (Point-Bench) | language and reasoning annotation | sensor only | 2025 | not stated | open | pointarena-point-bench | ||
| Embodied CoT features and demos for LIBERO | language and reasoning annotation | simulation | 3,917 | 2025 | open | open | embodied-cot-features-and-demos-for-libero | |
| EmbodiedMemory-Bench | language and reasoning annotation | simulation | 2,554 | 2026 | non commercial | open | embodiedmemory-bench | |
| Reason-RFT CoT Dataset | language and reasoning annotation | sensor only | 2025 | open | open | reason-rft-cot-dataset | ||
| FineVLA-Data and RoboFine-bench | language and reasoning annotation | cross embodiment | 47,159 | 2026 | open | open | finevla-data-and-robofine-bench | |
| InstructPart | language and reasoning annotation | sensor only | 2025 | not stated | open | instructpart | ||
| WGO-Bench | language and reasoning annotation | other | 2026 | non commercial | open | wgo-bench | ||
| DrivingVQA | language and reasoning annotation | other | 2025 | open | open | drivingvqa | ||
| Behavior-Skill | language and reasoning annotation | mobile manipulator | 10,000 | 2026 | open | open | behavior-skill | |
| SpatialRGPT-Bench | language and reasoning annotation | sensor only | 2024 | not stated | open | spatialrgpt-bench | ||
| Cosmos-Reason1 SFT dataset and benchmark | language and reasoning annotation | cross embodiment | 2025 | open | open | cosmos-reason1-sft-dataset-and-benchmark | ||
| Embodied-Reasoner | language and reasoning annotation | simulation | 9,390 | 2025 | not stated | open | embodied-reasoner | |
| MMSI-Bench | language and reasoning annotation | sensor only | 2025 | open | open | mmsi-bench | ||
| RoboInter-VQA | language and reasoning annotation | cross embodiment | 2026 | not stated | open | robointer-vqa | ||
| RoboSpatial-Home | language and reasoning annotation | sensor only | 2025 | open | open | robospatial-home | ||
| PointMotionBench | language and reasoning annotation | other | 2026 | not stated | not stated | pointmotionbench | ||
| X-Planner benchmark | language and reasoning annotation | cross embodiment | 1,500 | 2026 | not stated | open | x-planner-benchmark | |
| SPAR-7M | language and reasoning annotation | sensor only | 2025 | open | open | spar-7m | ||
| SAGE-3D (InteriorGS scenes and VLN data) | navigation and mobile | wheeled or navigation | 2,000,000 | 2025 | open | open | sage-3d-interiorgs-scenes-and-vln-data | |
| SR-3D-Bench | language and reasoning annotation | human egocentric | 2026 | open | open | sr-3d-bench | ||
| WM-ABench | language and reasoning annotation | simulation | 2025 | open | open | wm-abench | ||
| NVIDIA PhysicalAI-Traffic-Anomaly-Reasoning | language and reasoning annotation | sensor only | 2026 | open | open | nvidia-physicalai-traffic-anomaly-reasoning | ||
| allenai/MolmoAct2-SO100_101-Dataset | teleop | low cost arm | 2026 | open | open | allenai-molmoact2-so100-101-dataset | ||
| NaviTrace | language and reasoning annotation | cross embodiment | 2025 | open | open | navitrace | ||
| EmbodiedEval | language and reasoning annotation | simulation | 2025 | open | open | embodiedeval | ||
| RoadSocial | language and reasoning annotation | sensor only | 2025 | non commercial | gated | roadsocial | ||
| Vlaser data (Vlaser-6M) | language and reasoning annotation | cross embodiment | 2026 | open | open | vlaser-data-vlaser-6m | ||
| RACER augmented RLBench | language and reasoning annotation | single arm | 10,159 | 2024 | open | open | racer-augmented-rlbench | |
| Kimodo Human Motion Generation Benchmark | language and reasoning annotation | other | 2026 | open | open | kimodo-human-motion-generation-benchmark | ||
| KITScenes LongTail | language and reasoning annotation | vehicle | 2026 | non commercial | gated | kitscenes-longtail | ||
| libero-r-datasets | language and reasoning annotation | single arm | 2026 | not stated | open | libero-r-datasets | ||
| Where2Place | language and reasoning annotation | sensor only | 2024 | open | open | where2place | ||
| Language_Tactile | language and reasoning annotation | sensor only | 2026 | not stated | open | language-tactile | ||
| PixMo-Points | language and reasoning annotation | sensor only | 2024 | open | open | pixmo-points | ||
| ERIQ | language and reasoning annotation | other | 2025 | open | open | eriq | ||
| PhysBench | language and reasoning annotation | sensor only | 2025 | open | open | physbench | ||
| Robo2VLM Reasoning (ManipulationVQA-60k) | language and reasoning annotation | cross embodiment | 2025 | open | open | robo2vlm-reasoning-manipulationvqa-60k | ||
| FSD-Dataset and VABench | language and reasoning annotation | single arm | 2025 | not stated | open | fsd-dataset-and-vabench | ||
| Embodied-R1 Dataset (Embodied-Points-200K) | language and reasoning annotation | cross embodiment | 2025 | custom terms | open | embodied-r1-dataset-embodied-points-200k | ||
| OmniSpatial | language and reasoning annotation | sensor only | 2025 | open | open | omnispatial | ||
| VQASynth SpaceLLaVA | language and reasoning annotation | sensor only | 2024 | open | open | vqasynth-spacellava | ||
| NILS relabels of Fractal and Bridge | language and reasoning annotation | cross embodiment | 2024 | not stated | open | nils-relabels-of-fractal-and-bridge | ||
| EmbodiedBench | language and reasoning annotation | simulation | 2025 | not stated | open | embodiedbench | ||
| TraceSpatial-Bench | language and reasoning annotation | sensor only | 2025 | open | open | tracespatial-bench | ||
| Grounded 3D-LLM dataset | language and reasoning annotation | sensor only | 2024 | not stated | open | grounded-3d-llm-dataset | ||
| BridgeEQA | language and reasoning annotation | sensor only | 2026 | open | open | bridgeeqa | ||
| PARTNR episodes | language and reasoning annotation | simulation | 111,652 | 2024 | non commercial | open | partnr-episodes | |
| RONAR (RoboNar) | language and reasoning annotation | mobile manipulator | 2024 | open | open | ronar-robonar | ||
| DriveLM | language and reasoning annotation | vehicle | 2023 | non commercial | gated | drivelm | ||
| RynnBrain-Bench | language and reasoning annotation | human egocentric | 2026 | open | open | rynnbrain-bench | ||
| VitaSet | language and reasoning annotation | single arm | 2025 | open | open | vitaset | ||
| ERQA+ (ERQA Plus) | language and reasoning annotation | sensor only | 2026 | open | open | erqa-erqa-plus | ||
| MobileVLA-CoT | language and reasoning annotation | legged | 2026 | not stated | open | mobilevla-cot | ||
| PRISM-100K (DreamVu) | language and reasoning annotation | human egocentric | 2026 | non commercial | gated | prism-100k-dreamvu | ||
| Open Spatial Dataset (SpatialRGPT) | language and reasoning annotation | sensor only | 2024 | not stated | open | open-spatial-dataset-spatialrgpt |
Everything that exists: https://datasets.gurasees.com/browse.md