Data typesAny data type, one calibration gate.
The guide, the training module and the test are the same shape whatever the file is. Audio carries real numbers because it has shipped; the rest are open on the same bench.
Audio
DeliveredVerbatim transcription, word-level alignment, speaker attribution and entity spans, to one written convention per project.
- Verbatim transcription — fillers, repeats and false starts kept
- Word alignment: start and end on every word
- Speaker turns and diarization
- Entity and intent spans on the tokens that carry the transaction
3,000+ hours transcribed · 1,000+ hours aligned · 95%+ QA-sample accuracy
Image
OpenBoxes, polygons, masks and attributes, with the guide and the gold set that make two annotators land on the same answer.
- Bounding boxes and polygons
- Segmentation masks
- Attribute and classification labels
- Document and receipt field extraction
No volumes listed until the first order ships.
Video
OpenTemporal segmentation, tracking and captions — including the embodied tasks on the Physical Intelligence pages.
- Task and sub-task boundaries
- Object tracking across frames
- Natural-language captions per segment
- Success and failure labels per episode
Text
OpenPreference judgments, rubric scores and span labels by people who work in the sector, against a rubric a stranger can apply.
- Pairwise preference and rubric scoring
- Entity and intent spans
- Classification to a sector taxonomy
- Red-team and policy labels
No volumes listed until the first order ships.