Skip to main content

Catalogue

Explore our dataset collections

Purpose-built formats for AI training and evaluation across speech, screen, video, audio, image and multimodal data.

Speech and dialogue

AUDIO01

Natural English Dialogue Sets

Natural multi-speaker conversation with the turn structure, timing and speaker context needed for ASR, diarisation and conversational systems.

AUDIODIALOGUE
View dataset
AUDIO02

Cross-Language Dialogue Sets

Multilingual conversation organised under a consistent data structure while preserving the language-specific information each locale requires.

AUDIOMULTILINGUAL
View dataset
AUDIO03

Accent and Regional Speech Sets

Speech organised across defined accent and regional groups for adaptation, coverage analysis and comparative model evaluation.

AUDIOACCENT COVERAGE
View dataset
AUDIO04

Rare and Low-Resource Language Sets

Speech data for languages and varieties where existing machine-learning resources are limited, fragmented or difficult to standardise.

AUDIOLANGUAGE COVERAGE
View dataset
AUDIO05

Instruction and Command Voice Sessions

Spoken instructions and short follow-up exchanges that preserve intent, correction, confirmation and clarification behaviour.

AUDIOCOMMAND SPEECH
View dataset
AUDIO06

Bilingual Code-Switching Dialogue Sets

Bilingual conversation in which language changes can be represented within turns rather than reduced to a single session label.

AUDIOCODE-SWITCHING
View dataset
AUDIO07

Expressive and Styled Speech Sets

Speech captured under defined delivery styles for work on prosody, pacing, emphasis and style-conditioned modelling.

AUDIOPROSODY
View dataset
AUDIO08

Customer Service Dialogue Simulations

Scenario-based support conversations that preserve roles, changing intent, escalation and the outcome of the exchange.

AUDIOSUPPORT DIALOGUE
View dataset

Screen and software

SCREEN09

Software Task Screenflows

Desktop task recordings that connect interface state, user action and task outcome across a continuous workflow.

SCREENCOMPUTER USE
View dataset
SCREEN10

Browser Navigation and Web Task Flows

Browser-based task sequences for models that need to reason across navigation, page state and multi-step interaction.

SCREENWEB AGENTS
View dataset
SCREEN11

Error Recovery and Correction Screenflows

Screen workflows centred on failed actions, visible error states, corrections and the path back towards task completion.

SCREENERROR RECOVERY
View dataset
SCREEN12

Mobile App Task Screenflows

Mobile task sequences with device context and touch interaction where available, built for models that operate inside app interfaces.

SCREENMOBILE
View dataset
SCREEN13

Guided Software Walkthrough Sessions

Narrated software workflows that connect spoken explanation with screen state, visible text and user action.

SCREENNARRATED WORKFLOWS
View dataset

Video

VIDEO14

First-Person Activity Clips

Egocentric activity video for models that need to understand actions, objects and task progression from the actor's point of view.

VIDEOEGOCENTRIC
View dataset
VIDEO15

Hand and Object Interaction Clips

Close-range manipulation video focused on hand-object contact, fine-grained actions and changes in object state.

VIDEOMANIPULATION
View dataset
VIDEO16

Procedural Demonstration Video Sets

Multi-step demonstrations structured around procedure order, temporal boundaries, progress and completion.

VIDEOPROCEDURES
View dataset
VIDEO17

Multi-Environment Activity Clips

Activity video captured across defined settings or conditions for robustness, domain-shift and context-sensitive model work.

VIDEOENVIRONMENT DIVERSITY
View dataset

Audio and multimodal

AUDIO18

Acoustic Environment Sets

Ambient audio organised around acoustic scenes, recording conditions and sound events where the task requires them.

AUDIOACOUSTIC SCENES
View dataset
MULTIMODAL19

Multimodal Paired Capture Sets

Linked audio, video, screen, image or text records with explicit relationships between modalities.

MULTIMODALPAIRED CAPTURE
View dataset
CUSTOM20

Buyer-Scoped Custom Collection

A collection designed from your own objective, schema, inputs, targets and delivery requirements rather than a predefined format.

CUSTOMBUYER-SCOPED
View dataset

Image

IMAGE21

Product and Packaging Image Sets

Product imagery across views, packaging surfaces and visual conditions for recognition, retrieval and packaging understanding.

IMAGEPRODUCTS
View dataset
IMAGE22

Everyday Object Photo Sets

Object images across viewpoints and states, with persistent instance relationships where the task depends on them.

IMAGEOBJECTS
View dataset
IMAGE23

Real-World Text and Signage Photos

Text-bearing scenes with explicit links between image regions, transcription, language and script for OCR and scene-text models.

IMAGESCENE TEXT
View dataset

Looking for something different?

If the task does not fit an existing format, we can define a collection around your own data contract or model requirement.