Person filming a video with a smartphone
Video Data for AI

Video Data Collection
for AI Training

Get short videos filmed by real people on their own phones in the countries you choose. Describe the action, set the pay and approve only the clips that fit.

Results in hours

Video tasks usually get first clips within minutes of going live.

190+ countries

Film short actions from everyday people in the countries you choose.

Review every clip

Approve only the videos that match your brief. Reject what does not.

190+

Countries

13,487

Tasks completed

12,000+

Registered workers

~4 Hours

Avg time to results

How It Works

How Video Data Collection Works

Launch a video task yourself and start receiving clips from real contributors, without long negotiations.

Person at a laptop writing video task instructions
01 · Create

Describe what people should film

Explain the action, camera angle, length and setting you need. Add an example video so workers see exactly what a good clip looks like.

  • Actions, activities or places
  • Example video for reference
  • Clear length and angle rules
Laptop with a world map on screen for country targeting
02 · Target

Choose countries and clip count

Select the countries your contributors should come from, set how many clips each person sends and how many submissions you need in total.

  • Any number of countries in one task
  • Several clips per submission
  • You set the total volume
Person recording a video of a scene with a smartphone
03 · Film

Contributors film on their own phones

Your task becomes available to workers in the selected countries. They record the scenes you described in their homes, streets and workplaces.

  • Different phones and cameras
  • Real homes, streets and workplaces
  • New people in every submission
Person watching video clips on a monitor
04 · Review

Approve only the clips that fit

Watch every submission, approve the videos that meet your requirements and reject the rest. Keep your dataset clean from the start.

  • Watch every clip
  • Approve or reject each submission
  • Clean data for training
Video Data Types

Types of Video Data You Can Collect

Set up a task for the exact clips your model needs. Contributors follow your instructions and film on their own phones.

Actions and Gestures

Movements performed by many different people

Contributors record gestures, hand signs, poses or simple actions you describe. Every person moves a little differently, which is exactly what your model needs.

  • Your own list of actions
  • Many people, many styles
  • Repeatable short clips
Create a video task
Person making a hand gesture toward a phone camera
First-Person Video

Point-of-view clips for physical AI

Workers film everyday tasks from their own point of view, with the phone held or mounted at eye or chest level. Used to train robotics and world models.

  • Egocentric view of real tasks
  • Natural hand movements
  • Homes and workplaces worldwide
Create a video task
Hands cooking filmed from above in a first-person view
Everyday Activities

Real life, recorded as it happens

Collect clips of cooking, cleaning, shopping, exercising or working. Contributors film the activities you choose in their own environments.

  • Activities you define
  • Real homes and places
  • Local routines by country
Create a video task
Person cleaning or cooking at home for an activity clip
Hands and Objects

How people handle everyday things

Contributors film their hands opening, sorting, assembling or using objects. Useful for manipulation, product recognition and robotics models.

  • Close-up hand movements
  • Real objects, not props
  • Steps from start to finish
Create a video task
Hands assembling or sorting objects in close-up video
Talking to Camera

Short clips with face and voice

People record themselves speaking a phrase or answering a question on camera. One clip gives you video and audio for multimodal models.

  • Face and speech together
  • Scripted or free answers
  • Diverse ages and backgrounds
Create a video task
Person recording a selfie video speaking to the camera
Walkthroughs

Rooms, shops and streets on video

Contributors walk through apartments, stores, offices or streets while filming. Train models for navigation, mapping and scene understanding.

  • Indoor and outdoor routes
  • Local places by country
  • Steady, continuous shots
Create a video task
Person walking through a room or street filming with a phone
Task Example

Videos With Descriptions
in One Task

Ask contributors to add a short description of what happens in each clip. You get videos with context, ready for labeling or training.

Your setup

What you set up

Task settings

Choose Video collection, set how many clips each worker sends and write instructions with what to film, how long and from which angle.

  • Video collection task type
  • Number of clips per submission
  • Instructions with an example video

Your view

What you receive

Submission

Each submission arrives with the video clips and the worker's description, ready to review and approve.

  • Clips and text in one submission
  • Worker country on every submission
  • Approve or reject each set
Targeting

Target the Right Contributors

Choose where your clips come from and tell contributors exactly what to film. Every submission comes from a different person.

Countries and regions

Accept workers from all countries or pick specific ones by continent or individually. Capture local homes, streets and routines in each market.

Contributor requirements

Describe who should film in your instructions, such as age group, setting or phone type. Reject submissions that do not match.

Example videos

Upload a reference clip that shows the angle, length and pace you expect, so contributors get it right the first time.

One contributor, one submission

Each worker can complete your task only once. Set the number of submissions you need and get clips from that many different people.

Quality Control

Quality Control You Stay in Charge Of

You decide what counts as a good clip and review every submission before it joins your dataset.

Person writing notes at a laptop to define video requirements

Clear instructions

Describe the action, length, angle and what must not appear in the frame.

Video player on a screen showing a sample clip

Example clips

Show a sample video so contributors see the pace and framing you expect.

Person reviewing video clips on a laptop

Approve or reject

Watch each submission and accept only clips that meet your requirements.

Second person checking videos on a laptop at home

Second review

Launch a separate task where other workers check clips and descriptions for you.

Transparency

Collected With Consent

Every clip comes from a real person who chose to take your task and knew what it was for.

Contributors know the purpose

Your task description explains what people will film and that the videos will be used to train AI.

Voluntary participation

Nobody is assigned to your task. Each worker decides to take it and uploads their videos on their own.

Known source of every clip

Each submission comes from a specific worker in the country you selected, filmed specifically for your task.

Pricing

Estimate Your Video Data Collection Cost

You set the pay per task. Adjust the numbers to see what your project will cost, with the 15% platform fee included.

min

How long it takes a worker to complete one task.

$

Minimum pay is $0.10 per task. Need more clips later? Extend the same task instead of creating a new one.

FAQ

Frequently Asked Questions

Video data collection is gathering short clips from real people to train and test AI models that understand movement, actions and scenes. Good datasets include many people, places, devices and lighting conditions.

You set the pay for each task, starting from $0.10, plus a 15% platform fee. Longer or more complex clips usually need a higher pay rate. A campaign starts from $0.30, so you can test first.

You decide the length in your instructions. Short clips of a few seconds to a few minutes work best: they are easier to film, upload and review.

It is first-person (egocentric) video filmed from the contributor's own point of view while they do a task, such as cooking or sorting items. This type of data is used to train robotics and physical AI models.

Yes. Ask contributors in your instructions to describe what happens in each clip, and every submission will include both video and text.

You can accept workers from all countries or choose specific continents and countries when you create the task.

Describe the resolution, orientation, frame rate and conditions you need in your instructions. Add an example clip to show the result you expect.

State in your instructions that only the contributor may appear in the frame and that faces of others, documents or screens must stay out of view. Reject any submission that breaks these rules.

Yes. Extend your existing task instead of creating a new one. Workers who already took part still cannot submit again.

If you need a special format, a review flow of your own or help with your instructions, contact us through the form below and we will set it up with you.

Contact

Planning a Large Video Data Project?

Need high volume, first-person footage or help writing your task? Send a message and we will set it up with you.