#300 · Primary category: Computer Vision
VisDrone-Dataset
The dataset for drone based detection and tracking is released, including both image/video, and annotations.
Project last updated:09/24/23
GitHub Stars
2.5K
Forks
242
Contributors
1
License
Other
Why we included this project
The VisDrone benchmark is the dataset most aerial-vision teams reach for when they need realistic drone footage rather than ground-level camera images. It packs 288 video clips and more than 10,000 static frames shot from UAVs in 14 Chinese cities, all manually annotated with over 2.6 million bounding boxes for pedestrians, cars, bicycles, and tricycles. Because each frame also carries attributes like occlusion, visibility, and object class, the same data supports detection, single- and multi-object tracking, and crowd counting, which makes it a fair common ground for comparing approaches. Teams building surveillance, traffic monitoring, or agricultural drone systems will find the per-task ground truth easy to load into an evaluation pipeline. The images span dense urban scenes and sparse country settings, so models trained here tend to hold up across conditions.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
opencv
Open Source Computer Vision Library
RuView
π RuView turns commodity WiFi signals into real-time spatial intelligence, vital sign monitoring, and presence detection — all without a single pixel of video.
PaddleOCR
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
MinerU
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
tesseract
Tesseract Open Source OCR Engine (main repository)