paper-with-me

Papers

An Open Source and Open Hardware Deep Learning-powered Visual Navigation Engine for Autonomous Nano-UAVs

2019-05-10 · Daniele Palossi, Francesco Conti, Luca Benini

Nano-size unmanned aerial vehicles (UAVs), with few centimeters of diameter and sub-10 Watts of total power budget, have so far been considered incapable of running sophisticated visual-based autonomous navigation software without external aid from base-stations, ad-hoc local positioning infrastructure, and powerful external computation servers. In this work, we present what is, to the best of our knowledge, the first 27g nano-UAV system able to run aboard an end-to-end, closed-loop visual pipeline for autonomous navigation based on a state-of-the-art deep-learning algorithm, built upon the open-source CrazyFlie 2.0 nano-quadrotor. Our visual navigation engine is enabled by the combination of an ultra-low power computing device (the GAP8 system-on-chip) with a novel methodology for the deployment of deep convolutional neural networks (CNNs). We enable onboard real-time execution of a state-of-the-art deep CNN at up to 18Hz. Field experiments demonstrate that the system's high responsiveness prevents collisions with unexpected dynamic obstacles up to a flight speed of 1.5m/s. In addition, we also demonstrate the capability of our visual navigation engine of fully autonomous indoor navigation on a 113m previously unseen path. To share our key findings with the embedded and robotics communities and foster further developments in autonomous nano-UAVs, we publicly release all our code, datasets, and trained networks.

📄 PDF Abstract BibTeX arXiv:1905.04166

Code (2)

pulp-platform/pulp-dronet 공식 구현 tf
joanfmendo/prop-dronet

Tasks

Autonomous NavigationVisual Navigation

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Docling Technical Report

2024-08-19 · Christoph Auer, Maksym Lysak, Ahmed Nassar, Michele Dolfi 외

This technical report introduces Docling, an easy to use, self-contained, MIT-licensed open-source package for PDF document conversion. It is powered by state-of-the-art specialized AI models for layout analysis (DocLayN…

ALOHA 2: An Enhanced Low-Cost Hardware for Bimanual Teleoperation

2024-02-07 · ALOHA 2 Team, Jorge Aldaco, Travis Armstrong, Robert Baruch 외

Diverse demonstration datasets have powered significant advances in robot learning, but the dexterity and scale of such data can be limited by the hardware cost, the hardware robustness, and the ease of teleoperation. We…

MuJoCo

OpenGlass: Ultra-Low-Power On-Device AI Eyewear with Event-based Vision

2026-06-05 · Pietro Bonazzi, Julian Moosmann, Ahmet Celik, Philipp Mayer 외 arxiv

Smart eyewear enables unobtrusive, context-aware interaction through multimodal sensors and on-device intelligence, but is severely limited by power, memory, and compute constraints in a compact form factor. Open-hardwar…

Hand Gesture RecognitionEvent-based vision

Docling: An Efficient Open-Source Toolkit for AI-driven Document Conversion

2025-01-27 · Nikolaos Livathinos, Christoph Auer, Maksym Lysak, Ahmed Nassar 외

We introduce Docling, an easy-to-use, self-contained, MIT-licensed, open-source toolkit for document conversion, that can parse several types of popular document formats into a unified, richly structured representation. …

NeoRacer: An Open, Standardized 1:12 Scale Autonomous Race Car for Benchmarking and Education

2026-07-29 · Koneshka Bandyopadhyay, Ansh Mehta, Bassel El Mabsout, Renato Mancuso arxiv

Many scientific fields rely on standard benchmarks and shared platforms to improve review and reproducibility, but autonomous systems research still lacks widely accepted open hardware. Where standardization has emerged,…