Alia 360° 16K panorama — single shot equirectangular
360°Alia exo stream · unwrapped equirectangular16K
90°180°270°360°

Real-world human data that robot foundation models and world models train on.

DreamVu is an applied research company. We built the camera, we publish the research, and we run the operation that delivers the data.

( 00 ) QualityBuilt to frontier-lab specification.Full numeric specification under NDA
G1
Calibration
Verified per rig, per session.
G2
Alignment
All camera streams confirmed in sync before annotation starts.
G3
Annotation review
A person reviews every batch.
G4
Rejection
Batches that fall below the threshold are captured again.
( 01 )Capture Library

Look at the data.

Every clip below was captured in a real working environment. Nothing is staged and nothing is synthetic.

Ego + Exo-A + Exo-B · synchronized
Stocking and replenishment
Ego + Exo-A + Exo-B · synchronized
Pick and pack
Ego + Exo-A + Exo-B · synchronized
Food prep
( 02 )The DreamVu Pipeline

We customize the pipeline for every data challenge.

Each stage uses the best method available — open source where open source is better, our own where it is not. We publish research in this field, so we know the difference, and we swap components when something better comes out. You tell us what your model needs to learn, and we build the pipeline for it.

Stage 01

Capture

Alia is our own camera and our first choice. If a program needs other cameras, we use those too.

Stage 02

Annotate

AI does the first pass. A person reviews every batch. 20,000 hours a month.

Stage 03

Deliver

The format your training stack uses. OpenUSD, LeRobot (RLDS), Open X-Embodiment.

Applied research — how we decide what each stage does
  • VLM training datasetsVision-language fine-tuning
  • VLA training datasetsManipulation and action policies
  • Simulation-ready USDAny USD-compatible simulator
  • High-resolution walkthrough videoWorld model and generative video training
  • Custom formatsOn specification

If you need something that is not on this list, tell us and we will scope it.

See how the pipeline works

( 03 )The Capture System

An exocentric view with real 3D geometry.

One shot from Alia is all of this. A single 360° stereo capture produces a 16K RGB panorama of the entire room and metric depth for every pixel — no scanning rig, no multi-camera stitching, no post-processed geometry. Below, both outputs from the same capture.

DRAG
Fig. 03.1 — 360° RGB · single shot · 16K equirect
DRAG
Fig. 03.2 — Metric depth · every pixel · native 8K equirect

The same instant, two ways: 16K color across the full sphere and metric depth for every pixel. Only possible with Alia.

Single-shot 360° stereo with depth
The whole sphere and its 3D geometry, from one camera position. The optical design was published at CVPR in 2016.
32+ patents
The optics are protected. No one else can capture this.
( 04 )Scale and Access

Capacity and access, already built.

20,000 hrs
per month — capacity for richly annotated data
Hundreds
of venues in our capture network
Retail to
industrial
the range of environments we capture in

Getting a capture team into a working pharmacy, an auto plant, or a hospital ward takes agreements, training, and compliance work. We have already done that work.

Examples of where we capture

Grocery
Pharmacy Retail
Hardware Stores
Fashion Retail
Drycleaning & Garment Services
QSR & Restaurant Kitchens
Precision Manufacturing
Auto Parts Manufacturing
Industrial Remanufacturing
Warehousing & Materials Handling
Healthcare (Clinical)
Senior & Home Care
( 05 )Research

We publish what we learn.

We test our capture and annotation methods on public models and publish the results. Four papers so far, alongside our granted patents and the earlier computer vision work our team brought to the company.

66.6%
error reduction
Cosmos-Reason2-2B · PRISM
2.19×
improvement
GR00T N1.6 · SABER
6 / 7
metrics won
Video world models · RetailSMV

Measured on public foundation and robot models. Published, with reproducible results.

( 06 )Frequently Asked

Common questions.

What does DreamVu do?
We capture real people doing real work, and deliver it as training data for robot foundation models and world models. We design the capture, run it, annotate it, and deliver it in the format you train on.
What can I buy?
A capture program. There is no catalog. You tell us what your model needs to learn and which environments it needs to learn from. We scope the capture, the annotation, and the delivery format around that.
How is your data different?
Our camera, Alia, records the whole scene as a single-shot 360° stereo panorama with depth, protected by more than 32 patents. We synchronize it with a head-mounted camera, so you get the worker’s point of view and the full 3D room around them. Conventional setups give you flat video from one fixed angle.
Do you sell cameras?
No. The camera is how we capture. The data is what we sell.
Who is this for?
Teams training robot foundation models, vision-language-action models, vision-language models, and world models. Anyone whose system has to understand or act in the physical world.

Tell us what your model needs to learn.

Capture programs, research collaboration, and dataset partnerships.

Talk to us