DATA MARKETPLACE

Ready-made datasets,
pick and deploy

30 categories · 4,000+ SKUs of ready-made datasets. Embodied Ego and public-video data lead the catalog, complemented by on-robot capture — every set ships with licensing chain, unified schema and a free sample.

30 cats4,000+ SKUs
2.3B+Video clip pool · as of 2026Q1
50,000h+4D / Ego capture reserve
100%Traceable licensing
EGO·POV
FEATUREDEGO · RLDS
Human first-person household demonstrations with hand-action semantics, task decomposition and object-interaction labels. Controlled shooting + teleoperation, consent-based and desensitized.
1,200h+Hours
60+Task families
RLDSFormat
KITCHEN·8scenes
FEATUREDEGO · LeRobot v3
First-person kitchen operations across prep / cooking / cleaning task chains, action sequences aligned with language instructions.
8 类Scene families
AlignedLang-aligned
LeRobotFormat
NAV·INDOOR
EGOSpatial理解
First-person indoor movement with spatial-relation labels, path semantics and obstacle events — for spatial understanding and navigation training.
Multi-homeEnvironments
SpatialRelations
RLDSFormat
CINE·2M+ CLIPS
FEATURED视频生成
Cinema-grade clips with dense scene / visual / narrative / feature / sentiment labels, key-frame accuracy >90%* — first-choice corpus for video foundation models.
200万+Clips
DenseLabels
Licensed3-layer license
DOCU·PHYSICS
FEATURED世界模型
Documentary-grade real-physics footage covering natural motion, material interaction and lighting — high-fidelity corpus for world models.
500万+Clips
RealReal physics
IP 链Traceable
ACTION·POOL 2.3B+
视频蒸馏VLA
Real human actions curated from a 2.3B+ clip pool with action semantics and temporal labels — scaled ammunition for the video-distillation track.
2.3B+Pool size
120+Task families
时序动作Labels
4D·MULTI-CAM
AUX本体 · 4D
Multi-cam synchronized 4D video, cm-level spatial alignment*, with action semantics and spatial-relation labels.
50,000h+Reserve
cm-levelAlignment*
DaaSDaaS
TELEOP·TRAJ
AUX遥操作
Real-robot teleoperation end-effector trajectories aligned frame-level with vision — for VLA post-training and imitation learning.
60+Task families
Frame-levelVision-aligned
RLDSFormat
SOCIAL·10 REGIONS
社媒SFT · RAG
Creator, post and comment data across 10 regions with sentiment labels, ASR transcripts and OCR fields — fully PII-desensitized.
30+Languages
10 区域Platforms
100%PII Desensitized
SKU·2.1B+
电商多模态Aligned
SKU-level image-text alignment across major platforms — for multimodal retrieval and generative commerce.
2.1B+SKU images
Img-textPairs
ParquetFormat
XBORDER·30+LANG
跨境电商多Languages
Product, review and search corpora across major cross-border platforms, 30+ languages, SFT / RAG ready.
30+Languages
200+Countries
JSONLFormat
TEXT·23B+
文本SFT
High-quality review corpora curated from a 23B+ text pool, human × AI cleaned, with stance and sentiment labels.
23B+Pool size
SentimentLabels
JSONLFormat

Can't find the dataset you need?

Beyond the catalog, ENDATA customizes capture and processing by scene family, task family and format — first delivery in as fast as 6 weeks.

Talk custom data →

* Figures are product metrics · internal evaluation, illustrative, as of 2026Q1; every dataset ships with licensing docs and a free sample.