COLLECTING NOW — 6 DATA TYPES

Humanoid robots learn from people first — we collect the human demonstration data that trains VLA models.

FORMATS

LeRobotRLDSUnitree G1 EDUUMIGELLO / leader-arms

WHAT WE COLLECT

Six data types, one capture pipeline

01

Bimanual teleop demonstrations

CORE PRODUCT

GELLO / leader-arm teleoperation across pick-and-place, folding, pouring and assembly tasks.

RGB × 3–4 views + wrist camera

joint states · actions · gripper state

30–60 Hz · LeRobot / RLDS

15–60 s episodes, language-annotated

02

Humanoid data (Unitree G1 EDU)

VR teleoperation: whole-body manipulation and locomotion — the most valuable embodiment-specific type for humanoid companies.

Head camera · RGB-D

full-body joint states · IMU · actions

03

UMI data (handheld gripper, no robot)

Handheld gripper with camera in real environments — cheapest type to collect, scales beyond the lab: apartments, cafés, stores.

RGB fisheye

gripper pose (SLAM) · grip width

04

Egocentric video

First-person capture of household and workplace tasks — for pre-training VLA and video models, sold by volume.

RGB, optional gaze / IMU

timecode-level action annotation

05

Multi-view scene observations

One task captured synchronously from 4–6 cameras plus an egocentric view — for representation learning, 3D reconstruction and cross-view learning.

4–6 synchronized RGB streams

+ egocentric camera

06

Failure & recovery data

Failed attempts plus operator error correction — rare and in demand: most open datasets contain only successful episodes.

Source-modality streams

failure / recovery segmentation

Scaling the capture

30+

Scenes, from lab rigs to apartments, cafés and stores

200+

Objects, with lighting and camera-view variation

30–60 Hz

Timestamp-synchronized streams, 15–60 s episodes

UMI
2 streams
Bimanual teleop
5 streams
Humanoid · multi-view
7+ streams
Language instructions per task
Scene metadata: objects, lighting, camera views
Per-episode QA labeling
Consent releases

COVERAGE

From lab rigs
to the real world

Tier IBimanual teleop and humanoid capture in the lab
Tier IIUMI handheld capture in apartments, cafés and stores
Tier IIIEgocentric video of everyday household and work tasks

WHO IT'S FOR

Built for the teams training robot foundation models

01

Data aggregators & marketplaces

  • Micro1
  • Cortex AI
  • Sensei
  • Awign
02

Foundation model & AI labs

  • Physical Intelligence
  • Skild AI
  • NVIDIA (GR00T team)
  • Figure AI
03

Humanoid & robot manufacturers

  • 1X Technologies
  • Agility Robotics
  • Apptronik
  • Boston Dynamics
04

Industrial automation

  • Covariant
  • Universal Robots
  • KUKA
  • Toyota Research Institute

See a sample episode and the full spec sheet.

Camera layouts, formats, QA process and pricing — in one data card.

or write to hello@athena-labs.ai