paper-with-me

홈 › Papers

Deeplite Neutrino: An End-to-End Framework for Constrained Deep Learning Model Optimization

2021-01-11 · Anush Sankaran, Olivier Mastropietro, Ehsan Saboori, Yasser Idris, Davis Sawyer, MohammadHossein AskariHemmat, Ghouthi Boukli Hacene

Designing deep learning-based solutions is becoming a race for training deeper models with a greater number of layers. While a large-size deeper model could provide competitive accuracy, it creates a lot of logistical challenges and unreasonable resource requirements during development and deployment. This has been one of the key reasons for deep learning models not being excessively used in various production environments, especially in edge devices. There is an immediate requirement for optimizing and compressing these deep learning models, to enable on-device intelligence. In this research, we introduce a black-box framework, Deeplite Neutrino for production-ready optimization of deep learning models. The framework provides an easy mechanism for the end-users to provide constraints such as a tolerable drop in accuracy or target size of the optimized models, to guide the whole optimization process. The framework is easy to include in an existing production pipeline and is available as a Python Package, supporting PyTorch and Tensorflow libraries. The optimization performance of the framework is shown across multiple benchmark datasets and popular deep learning models. Further, the framework is currently used in production and the results and testimonials from several clients are summarized.

📄 PDF Abstract BibTeX arXiv:2101.04073

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningModel Optimization

Similar Papers 제목 키워드 기반

Accelerating Deep Learning Model Inference on Arm CPUs with Ultra-Low Bit Quantization and Runtime

2022-07-18 · Saad Ashfaq, MohammadHossein AskariHemmat, Sudhakar Sah, Ehsan Saboori 외

Deep Learning has been one of the most disruptive technological advancements in recent times. The high performance of deep learning models comes at the expense of high computational, storage and power requirements. Sensi…

Quantization

DeepliteRT: Computer Vision at the Edge

2023-09-19 · Saad Ashfaq, Alexander Hoffman, Saptarshi Mitra, Sudhakar Sah 외

The proliferation of edge devices has unlocked unprecedented opportunities for deep learning model deployment in computer vision applications. However, these complex models require considerable power, memory and compute …

Quantization

Transfer Learning for Neutrino Scattering: Domain Adaptation with GANs

2025-08-18 · Jose L. Bonilla, Krzysztof M. Graczyk, Artur M. Ankowski, Rwik Dharmapal Banerjee 외 arxiv

Transfer learning (TL) is used to extrapolate the physics information encoded in a Generative Adversarial Network (GAN) trained on synthetic neutrino-carbon inclusive scattering data to related processes such as neutrino…

Transfer LearningDomain Adaptation

Application of Neural Networks for the Reconstruction of Supernova Neutrino Energy Spectra Following Fast Neutrino Flavor Conversions

2024-01-30 · Sajad Abbar, Meng-Ru Wu, Zewei Xiong

Neutrinos can undergo fast flavor conversions (FFCs) within extremely dense astrophysical environments such as core-collapse supernovae (CCSNe) and neutron star mergers (NSMs). In this study, we explore FFCs in a \emph{m…

Deep-Learning-Based Kinematic Reconstruction for DUNE

2020-12-11 · Junze Liu, Jordan Ott, Julian Collado, Benjamin Jargowsky 외

In the framework of three-active-neutrino mixing, the charge parity phase, the neutrino mass ordering, and the octant of $\theta_{23}$ remain unknown. The Deep Underground Neutrino Experiment (DUNE) is a next-generation …

Deep Learning