Text · dataset · 2026
Stera-10M
Listed in Hugging Face Datasets
Description
Stera-10M Visualizer: platform.fpvlabs.ai/dataset/stera-10m/viz Dataset
Summary
Read the rest (1 more)
Stera-10M is an open egocentric multimodal dataset for embodied AI, robotics, world models, and spatial intelligence, captured end-to-end on commodity iPhone Pro hardware through the open Stera platform. It contains 200 hours of synchronized first-person recordings across 500+ sessions from 20 contributors in 20+ unique environments, with 10 million RGB frames, LiDAR depth… See the full description on the dataset page: huggingface.co/datasets/fpvlabs/stera-10m.
Links
Get the data
- Hugging Face dataset page huggingface.co/datasets/fpvlabs/stera-10m ↗
Gated (auto) - request access on the dataset page
landing page · request access · from Hugging Face
Where it is published
- DOI doi.org/10.57967/hf/10149 ↗
DOI / persistent id · from Hugging Face
Documentation and papers
- arXiv:2605.05945 arxiv.org/abs/2605.05945 ↗
publication · from Hugging Face
Catalogue records · 1
- Hub API huggingface.co/api/datasets/fpvlabs/stera-10m ↗
metadata API · from Hugging Face
Topics
- Stated by source
- 3d · audio · depth estimation · image to text · robotics · text · video · video classification
- From keywords
- Audio · Computer Science & AI · Text · Video
Provenance · 1 source records, 19 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| Hugging Face Datasets | fpvlabs/stera-10m | 8 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| access_level | source · Hugging Face | connector:huggingface@1.0.0 | /gated |
| concepts[field].local:field:computer-science-ai | mapping · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:3d | source · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:audio | source · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:text | source · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].hf_modality:video | source · Hugging Face | connector:huggingface@1.0.0 | |
| concepts[modality].local:modality:audio | mapping · Hugging Face | vocabulary-mapper@1.0.0 | keywords['audio'] |
| concepts[modality].local:modality:text | mapping · Hugging Face | vocabulary-mapper@1.0.0 | keywords['text'] |
| concepts[modality].local:modality:video | mapping · Hugging Face | vocabulary-mapper@1.0.0 | keywords['video'] |
| concepts[task].hf_task:depth-estimation | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:image-to-text | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:robotics | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| concepts[task].hf_task:video-classification | source · Hugging Face | connector:huggingface@1.0.0 | /tags[task_categories:*] |
| created_date | source · Hugging Face | connector:huggingface@1.0.0 | |
| description | source · Hugging Face | connector:huggingface@1.0.0 | /description |
| license_text | source · Hugging Face | connector:huggingface@1.0.0 | |
| publication_date | source · Hugging Face | connector:huggingface@1.0.0 | |
| title | source · Hugging Face | connector:huggingface@1.0.0 | /id |
| updated_date | source · Hugging Face | connector:huggingface@1.0.0 |