paper-with-me

Papers

Accurate Deep Representation Quantization with Gradient Snapping Layer for Similarity Search

2016-10-30 · Shicong Liu, Hongtao Lu

Recent advance of large scale similarity search involves using deeply learned representations to improve the search accuracy and use vector quantization methods to increase the search speed. However, how to learn deep representations that strongly preserve similarities between data pairs and can be accurately quantized via vector quantization remains a challenging task. Existing methods simply leverage quantization loss and similarity loss, which result in unexpectedly biased back-propagating gradients and affect the search performances. To this end, we propose a novel gradient snapping layer (GSL) to directly regularize the back-propagating gradient towards a neighboring codeword, the generated gradients are un-biased for reducing similarity loss and also propel the learned representations to be accurately quantized. Joint deep representation and vector quantization learning can be easily performed by alternatively optimize the quantization codebook and the deep neural network. The proposed framework is compatible with various existing vector quantization approaches. Experimental results demonstrate that the proposed framework is effective, flexible and outperforms the state-of-the-art large scale similarity search methods.

📄 PDF Abstract BibTeX arXiv:1610.09645

Code (0)

등록된 구현이 없습니다.

Tasks

Quantization

Similar Papers 제목 키워드 기반

Snapping Actuators with Asymmetric and Sequenced Motion

2026-02-20 · Xin Li, Ye Jin, Mohsen Jafarpour, Hugo de Souza Oliveira 외 arxiv

Snapping instabilities in soft structures offer a powerful pathway to achieve rapid and energy-efficient actuation. In this study, an eccentric dome-shaped snapping actuator is developed to generate controllable asymmetr…

Quartet II: Accurate LLM Pre-Training in NVFP4 by Improved Unbiased Gradient Estimation

2026-01-30 · Andrei Panferov, Erik Schultheis, Soroush Tabesh, Dan Alistarh arxiv

The NVFP4 lower-precision format, supported in hardware by NVIDIA Blackwell GPUs, promises to allow, for the first time, end-to-end fully-quantized pre-training of massive models such as LLMs. Yet, existing quantized tra…

Accurate Neural Training with 4-bit Matrix Multiplications at Standard Formats

2021-12-19 · Brian Chmiel, Ron Banner, Elad Hoffer, Hilla Ben Yaacov 외

Quantization of the weights and activations is one of the main methods to reduce the computational footprint of Deep Neural Networks (DNNs) training. Current methods enable 4-bit quantization of the forward phase. Howeve…

Quantization

Tendon-Driven Reciprocating and Non-Reciprocating Motion via Snapping Metabeams

2026-02-20 · Mohsen Jafarpour, Ayberk Yüksek, Shahab Eshghi, Stanislav Gorb 외 arxiv

Snapping beams enable rapid geometric transitions through nonlinear instability, offering an efficient means of generating motion in soft robotic systems. In this study, a tendon-driven mechanism consisting of spiral-bas…

Distribution-Aware Hadamard Quantization for Hardware-Efficient Implicit Neural Representations

2025-08-19 · Wenyong Zhou, Jiachen Ren, Taiqiang Wu, Yuxin Cheng 외 arxiv

Implicit Neural Representations (INRs) encode discrete signals using Multi-Layer Perceptrons (MLPs) with complex activation functions. While INRs achieve superior performance, they depend on full-precision number represe…

Image Reconstruction