
Fourier ActionNet
Fourier ActionNet contains more than 30K VR-teleoperated trajectories, or roughly 140 hours, from GR1-T1, GR1-T2, and GR2 humanoids using 6-DoF or 12-DoF dexterous hands. It emphasizes bimanual tabletop pick-and-place, pouring, insertion, cabinet interaction, and precise placement with varied household, office, food, and tool objects. RGB-D video, full robot and hand states and actions, end-link poses, and manually verified VLM-generated prompts support imitation learning, VLA, and world-model training.




