paper-with-me

홈 › Papers

Comparing Lexical and Semantic Vector Search Methods When Classifying Medical Documents

2025-05-16 · Lee Harris, Philippe De Wilde, James Bentham

Classification is a common AI problem, and vector search is a typical solution. This transforms a given body of text into a numerical representation, known as an embedding, and modern improvements to vector search focus on optimising speed and predictive accuracy. This is often achieved through neural methods that aim to learn language semantics. However, our results suggest that these are not always the best solution. Our task was to classify rigidly-structured medical documents according to their content, and we found that using off-the-shelf semantic vector search produced slightly worse predictive accuracy than creating a bespoke lexical vector search model, and that it required significantly more time to execute. These findings suggest that traditional methods deserve to be contenders in the information retrieval toolkit, despite the prevalence and success of neural models.

📄 PDF Abstract BibTeX arXiv:2505.11582

Code (0)

등록된 구현이 없습니다.

Tasks

Information Retrieval

Methods 이 논문이 사용한 방법론

Focus 설명 없음
SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Collocation Classification with Unsupervised Relation Vectors

2019-07-01 · ACL 2019 7 · Luis Espinosa Anke, Steven Schockaert, Leo Wanner

Lexical relation classification is the task of predicting whether a certain relation holds between a given pair of words. In this paper, we explore to which extent the current distributional landscape based on word embed…

ClassificationGeneral ClassificationRelationRelation Classification+1

LessLex: Linking Multilingual Embeddings to SenSe Representations of LEXical Items

2020-06-01 · CL 2020 6 · Davide Colla, Enrico Mensa, Daniele P. Radicioni

We present LESSLEX, a novel multilingual lexical resource. Different from the vast majority of existing approaches, we ground our embeddings on a sense inventory made available from the BabelNet semantic network. In this…

text similarity

Can Large Language Models (LLMs) Describe Pictures Like Children? A Comparative Corpus Study

2025-08-19 · Hanna Woloszyn, Benjamin Gagl arxiv

The role of large language models (LLMs) in education is increasing, yet little attention has been paid to whether LLM-generated text resembles child language. This study evaluates how LLMs replicate child-like language …

Semantic Similarity

Comparative Probing of Lexical Semantics Theories for Cognitive Plausibility and Technological Usefulness

2020-11-16 · COLING 2020 8 · António Branco, João Rodrigues, Małgorzata Salawa, Ruben Branco 외

Lexical semantics theories differ in advocating that the meaning of words is represented as an inference graph, a feature mapping or a vector space, thus raising the question: is it the case that one of these approaches …

Evaluating Distributed Representations for Multi-Level Lexical Semantics: A Research Proposal

2024-06-02 · Zhu Liu

Modern neural networks (NNs), trained on extensive raw sentence data, construct distributed representations by compressing individual words into dense, continuous, high-dimensional vectors. These representations are expe…

Sentence