paper-with-me

홈 › Papers

TITAN: Bringing The Deep Image Prior to Implicit Representations

2022-11-01 · Lorenzo Luzi, Daniel LeJeune, Ali Siahkoohi, Sina AlEMohammad, Vishwanath Saragadam, Hossein Babaei, Naiming Liu, Zichao Wang, Richard G. Baraniuk

We study the interpolation capabilities of implicit neural representations (INRs) of images. In principle, INRs promise a number of advantages, such as continuous derivatives and arbitrary sampling, being freed from the restrictions of a raster grid. However, empirically, INRs have been observed to poorly interpolate between the pixels of the fit image; in other words, they do not inherently possess a suitable prior for natural images. In this paper, we propose to address and improve INRs' interpolation capabilities by explicitly integrating image prior information into the INR architecture via deep decoder, a specific implementation of the deep image prior (DIP). Our method, which we call TITAN, leverages a residual connection from the input which enables integrating the principles of the grid-based DIP into the grid-free INR. Through super-resolution and computed tomography experiments, we demonstrate that our method significantly improves upon classic INRs, thanks to the induced natural image bias. We also find that by constraining the weights to be sparse, image quality and sharpness are enhanced, increasing the Lipschitz constant.

📄 PDF Abstract BibTeX arXiv:2211.00219

Code (1)

dlej/titan-implicit-prior 공식 구현 pytorch

Tasks

DecoderSuper-Resolution

Methods 이 논문이 사용한 방법론

Residual Connection 설명 없음

Similar Papers 제목 키워드 기반

TitaNet: Neural Model for speaker representation with 1D Depth-wise separable convolutions and global context

2021-10-08 · Nithin Rao Koluguri, Taejin Park, Boris Ginsburg

In this paper, we propose TitaNet, a novel neural network architecture for extracting speaker representations. We employ 1D depth-wise separable convolutions with Squeeze-and-Excitation (SE) layers with global context fo…

speaker-diarizationSpeaker DiarizationSpeaker Verification

TITAN: Future Forecast using Action Priors

2020-03-31 · CVPR 2020 6 · Srikanth Malla, Behzad Dariush, Chiho Choi

We consider the problem of predicting the future trajectory of scene agents from egocentric views obtained from a moving platform. This problem is important in a variety of domains, particularly for autonomous systems ma…

A Time Series is Worth Five Experts: Heterogeneous Mixture of Experts for Traffic Flow Prediction

2024-09-26 · Guangyu Wang, Yujie Chen, Ming Gao, Zhiqiao Wu 외

Accurate traffic prediction faces significant challenges, necessitating a deep understanding of both temporal and spatial cues and their complex interactions across multiple variables. Recent advancements in traffic pred…

Mixture-of-ExpertsPredictionTime SeriesTraffic Prediction

Task-oriented Prompt Enhancement via Script Generation

2024-09-24 · Chung-Yu Wang, Alireza DaghighFarsoodeh, Hung Viet Pham

Large Language Models (LLMs) have demonstrated remarkable abilities across various tasks, leveraging advanced reasoning. Yet, they struggle with task-oriented prompts due to a lack of specific prior knowledge of the task…

Code GenerationScript GenerationZero-Shot Learning

Multimodal Whole Slide Foundation Model for Pathology

2024-11-29 · Tong Ding, Sophia J. Wagner, Andrew H. Song, Richard J. Chen 외

The field of computational pathology has been transformed with recent advances in foundation models that encode histopathology region-of-interests (ROIs) into versatile and transferable feature representations via self-s…

Cross-Modal RetrievalmodelPrognosisRetrieval+3