paper-with-me

Papers Zero-shot Text Retrieval

“Zero-shot Text Retrieval” 태그가 달린 논문 7편 · 필터 해제

CLIP-PING: Boosting Lightweight Vision-Language Models with Proximus Intrinsic Neighbors Guidance

2024-12-05 · Chu Myaet Thwal, Ye Lin Tun, Minh N. H. Nguyen, Eui-Nam Huh 외

Beyond the success of Contrastive Language-Image Pre-training (CLIP), recent trends mark a shift toward exploring the applicability of lightweight vision-language models for resource-constrained scenarios. These models o…

Contrastive Learningcross-modal alignmentCross-Modal RetrievalLinear evaluation+6

LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment

2023-10-03 · Bin Zhu, Bin Lin, Munan Ning, Yang Yan 외

The video-language (VL) pretraining has achieved remarkable improvement in multiple downstream tasks. However, the current VL pretraining framework is hard to extend to multiple modalities (N modalities, N>=3) beyond vis…

Audio ClassificationContrastive LearningMultimodal Deep LearningScene Classification (unified classes)+11

Keras GPT Copilot: Integrating the Power of Large Language Models in Deep Learning Model Development

2023-05-15 · Zenodo GitHub 2023 5 · Fabi Prezja

Keras GPT Copilot is the first Python package designed to integrate an LLM copilot within the model development workflow, offering iterative feedback options for enhancing the performance of your Keras deep learning mode…

Data-to-Text GenerationText GenerationZero-shot Text Retrieval

AltCLIP: Altering the Language Encoder in CLIP for Extended Language Capabilities

2022-11-12 · Zhongzhi Chen, Guang Liu, Bo-Wen Zhang, Fulong Ye 외

In this work, we present a conceptually simple and effective method to train a strong bilingual/multilingual multimodal representation model. Starting from the pre-trained multimodal representation model CLIP released by…

Contrastive LearningCross-Modal RetrievalImage ClassificationImage Retrieval+9

Chinese CLIP: Contrastive Vision-Language Pretraining in Chinese

2022-11-02 · An Yang, Junshu Pan, Junyang Lin, Rui Men 외

The tremendous success of CLIP (Radford et al., 2021) has promoted the research and application of contrastive learning for vision-language pretraining. In this work, we construct a large-scale dataset of image-text pair…

Contrastive Learningimage-classificationImage ClassificationImage Retrieval+7

LaPraDoR: Unsupervised Pretrained Dense Retriever for Zero-Shot Text Retrieval

2022-03-11 · Findings (ACL) 2022 5 · Canwen Xu, Daya Guo, Nan Duan, Julian McAuley

In this paper, we propose LaPraDoR, a pretrained dual-tower dense retriever that does not require any supervised data for training. Specifically, we first present Iterative Contrastive Learning (ICoL) that iteratively tr…

Contrastive LearningRe-RankingRetrievalText Retrieval+1

FLAVA: A Foundational Language And Vision Alignment Model

2021-12-08 · CVPR 2022 1 · Amanpreet Singh, Ronghang Hu, Vedanuj Goswami, Guillaume Couairon 외

State-of-the-art vision and vision-and-language models rely on large-scale visio-linguistic pretraining for obtaining good performance on a variety of downstream tasks. Generally, such models are often either cross-modal…

Image RetrievalImage-to-Text RetrievalVisual ReasoningZero-shot Image Retrieval+2
1–7 / 7