Downloads
898
HUGGING FACE DATASET
ShuaiYang03/VLA_Instruction_Tuning — source facts cached from the per-dataset Hugging Face API.
DIRECT ANSWER
License: Not provided by Hugging Face; size: Not provided by Hugging Face; creator: ShuaiYang03; paper: arXiv 2507.17520.
Commercial use: Not established — no license was provided.
898
Not provided by Hugging Face
Sep 11, 2025
Not marked as gated
SOURCE CARD
This repository contains the VLA-IT dataset, a curated 650K-sample Vision-Language-Action Instruction Tuning dataset, and the SimplerEnv-Instruct benchmark. These are presented in the paper InstructVLA: Vision-Language-Action Instruction Tuning from Understanding to Manipulation. The dataset is designed to enable robots to integrate multimodal reasoning with precise action generation, preserving the flexible reasoning of large vision-language models while delivering leading manipulation…
Review the current card, files, access terms, and metadata.
ALTERNATIVES
Robotics datasets
Shared tags: Robotics, Region:Us, Manipulation, Vision Language Action.
Shared tags: Robotics, Region:Us, Manipulation, Vision Language Action.
Shared tags: Robotics, Region:Us, Embodied Ai, Vision Language Action.
Shared tags: Robotics, Region:Us, Manipulation, Vision Language Action.
Shared tags: Robotics, Region:Us, Embodied Ai, Vision Language Action.
Shared tags: Robotics, Region:Us, Manipulation, Vision Language Action.
TAGS
TRUELABEL ROUTING
Use the source facts and gaps from this profile to define a custom physical AI data request.