01▶
Video
Task-based videos, first-person and egocentric point of view, fixed camera, multi-environment shoots across household, industrial, commercial, and outdoor settings. Resolutions from HD to 4K, frame rates to your brief, with bounding-box and action annotation plus preview videos showing the work.
02◫
Image
Every category of 2D image collection — objects, scenes, people in approved contexts, documents to your specification — at whatever resolution and volume your brief demands, with bounding-box annotation and preview images showing the work.
03≋
Speech and audio
Multilingual speech across 8 languages — English, Hindi, Kannada, Telugu, Tamil, Malayalam, Marathi, Sanskrit — scripted prompts, natural conversations, channel-separated dialogue, voice commands and wake words, IVR utterances by intent, studio and field recordings, with transcription and labeling to your schema.
04≡
Text
Translation and localization pairs, transcription, named entities, sentiment, content, and moderation annotation to your taxonomy, in every language above and more on request.
05 / FEATURED⁂
Depth and LiDAR
RGB-D captures pairing each video with a synced depth stream, 3D point clouds with per-point range and angles, 3D bounding-box and point-cloud segmentation to your taxonomy. Indoor short-range programs scoped to your protocol. Production rigs scale with the program.
06 / FEATURED◎
Sensors and radar
Motion-signature captures pairing each video with a sensor recording of the same task — including tactile and force-torque streams for contact-rich work — synchronized multi-sensor packs, full rig specifications defined with you per program.
07 / FEATURED⬣
Multimodal packs
Everything above, in one delivery. Video, audio, depth, and annotation from the same capture, packed under a single manifest — raw files, annotations, and watermarked previews, each with a checksum. Raw files are never overwritten, and contributors stay anonymous: you see registry codes, never names.
08 / FEATURED⚙
Robot action data
Task demonstrations collected by teleoperation — joint states, end-effector poses, gripper actions, and proprioception, timestamp-synced with multi-view video. Success, failure, and human-takeover episodes, delivered in LeRobot or HDF5 on your target hardware, from tabletop arms to humanoids.