#203 · Primary category: Computer Vision
Ego4d
Ego4d dataset repository. Download the dataset, visualize, extract features & example usage of the dataset
Project last updated:07/25/26
GitHub Stars
639
Forks
65
Contributors
36
License
MIT
Why we included this project
Researchers and engineers working on first-person perception will find this the reference implementation for the Ego4D and Ego-Exo4D datasets, the largest public collections of egocentric video. The repository is not a model you run directly; it is the official tooling for getting the data into your own pipeline. It ships command-line downloaders for both datasets, a visualizer for browsing clips and annotations, and feature-extraction utilities, so you can move from license approval to usable tensors without hand-writing download scripts. The bundled benchmark tasks, covering episodic memory, hand-object interaction, audio-visual diarization, social interaction, and forecasting, give you ready-made evaluation targets if you are building models for wearable or assistive AI. Teams planning to train on egocentric video should start here to understand the annotation formats and the exact data access workflow before committing to a larger experiment.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
opencv
Open Source Computer Vision Library
RuView
π RuView turns commodity WiFi signals into real-time spatial intelligence, vital sign monitoring, and presence detection — all without a single pixel of video.
PaddleOCR
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
MinerU
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
tesseract
Tesseract Open Source OCR Engine (main repository)