Agentic coding
Repository-backed tasks for software-engineering agents, with executable tests, environment state, and patch-level evaluation.
Training data & evaluation environments
Realistic tasks. Executable environments. Verifiable outcomes.
Training and evaluation data for agents and multimodal systems.
Define the frontier.
Move intelligence forward.
Build toward the next horizon.
Agentic coding / Computer use / Multimodal systems
Scroll to explore ↓Data categories / Capabilities
Data spanning agentic coding, automated research, computer use, multimodal learning, and embodied AI.
Repository-backed tasks for software-engineering agents, with executable tests, environment state, and patch-level evaluation.
Kernel-optimization environments with correctness gates and latency and memory objectives for systems research.
Stateful cross-application workflows evaluated through artifact validity, evidence traceability, and rubric-based completion.
Persistent workspaces with multi-stage dependencies, tool use, failure recovery, and end-state verification.
Rights-scoped audiovisual and image-transformation data for representation learning, perception, and generation.
Synchronized egocentric video, pose, and motion data for action understanding, visuomotor learning, and world-model research.
Explore sample catalogs · Request access
Need a closer look?