Reka Responsible AI, Model Risk, Ethics & Governance Framework

How do we prepare data for a World Language Action model?

How do we prepare data for a World Language Action model?

How do we prepare data for a World Language Action model?

EPISODE 03

·

·

7:46

A key ingredient in training a new world model is the data. In this episode the team walks through how raw video becomes training data for a World Language Action model — what gets thrown away, what gets labelled, and where a human still has to sit in the loop.

It ends where the work actually is right now: scaling the pipeline from terabytes to petabytes without letting label quality quietly collapse along the way.

[TRANSCRIPT]

Test

EPISODE DETAILS

SERIES

Working Notes

RUNTIME

7:46

PUBLISHED

IN THIS ONE

Julian Lopez, Konrad Jamrozik

Data pipelineWLA models
Read the research