paper-with-me

Papers Keyword Extraction

“Keyword Extraction” 태그가 달린 논문 191편 · 필터 해제

Letting the Data Speak: Extracting Keywords from Crowdsourced Collections with AI

2026-07-10 · Miguel Arana-Catania, Catherine Conisbee, Matthew Kidd arxiv

Identifying and assigning keywords at scale is a technical, practical, and ethical challenge for crowdsourced collections. This article reports the findings of the "Extracting Keywords from Crowdsourced Collections" proj…

Keyword Extraction

Building a Multimodal Dataset of Academic Paper for Keyword Extraction

2026-06-30 · Jingyu Zhang, Xinyi Yan, Yi Xiang, Yingyi Zhang 외 arxiv

Up to this point, keyword extraction task typically relies solely on textual data. Neglecting visual details and audio features from image and audio modalities leads to deficiencies in information richness and overlooks …

Keyword Extraction

Weave of Formal Thought

2026-06-24 · Alexandre Bouayad arxiv

Large language models (LLMs) attain remarkable surface fluency on code, yet they neither formally guarantee the syntactic validity of their output nor leverage the hierarchical structure defining the target language. Whi…

Keyword Extraction

ReLeVAnT: Relevance Lexical Vectors for Accurate Legal Text Classification

2026-04-24 · Ishaan Gakhar, Harsh Nandwani arxiv

The classification of legal documents from an unstructured data corpus has several crucial applications in downstream tasks. Documents relevant to court filings are key in use cases such as drafting motions, memos, and o…

Binary ClassificationText ClassificationKeyword Extraction

Enhancing Unsupervised Keyword Extraction in Academic Papers through Integrating Highlights with Abstract

2026-04-21 · Yi Xiang, Chengzhi Zhang arxiv

Automatic keyword extraction from academic papers is a key area of interest in natural language processing and information retrieval. Although previous research has mainly focused on utilizing abstract and references for…

Information RetrievalKeyword Extraction

Bangla Key2Text: Text Generation from Keywords for a Low Resource Language

2026-04-21 · Tonmoy Talukder, G M Shahariar arxiv

This paper introduces \textit{Bangla Key2Text}, a large-scale dataset of $2.6$ million Bangla keyword--text pairs designed for keyword-driven text generation in a low-resource language. The dataset is constructed using a…

Keyword ExtractionText Generation

Enhancing MLLM Spatial Understanding via Active 3D Scene Exploration for Multi-Perspective Reasoning

2026-04-08 · Jiahua Chen, Qihong Tang, Weinong Wang, Qi Fan arxiv

Although Multimodal Large Language Models have achieved remarkable progress, they still struggle with complex 3D spatial reasoning due to the reliance on 2D visual priors. Existing approaches typically mitigate this limi…

Keyword ExtractionSpatial Reasoning3D Reconstruction

Analyzing Cancer Patients' Experiences with Embedding-based Topic Modeling and LLMs

2026-01-17 · Teodor-Călin Ionescu, Lifeng Han, Jan Heijdra Suasnabar, Anne Stiggelbout 외 arxiv

This study investigates the use of neural topic modeling and LLMs to uncover meaningful themes from patient storytelling data, to offer insights that could contribute to more patient-oriented healthcare practices. We ana…

Keyword Extraction

NepEMO: A Multi-Label Emotion and Sentiment Analysis on Nepali Reddit with Linguistic Insights and Temporal Trends

2025-12-28 · Sameer Sitoula, Tej Bahadur Shahi, Laxmi Prasad Bhatt, Anisha Pokhrel 외 arxiv

Social media (SM) platforms (e.g. Facebook, Twitter, and Reddit) are increasingly leveraged to share opinions and emotions, specifically during challenging events, such as natural disasters, pandemics, and political elec…

Sentiment AnalysisKeyword Extraction

OmniMER: Auxiliary-Enhanced LLM Adaptation for Indonesian Multimodal Emotion Recognition

2025-12-22 · Xueming Yan, Boyan Xu, Yaochu Jin, Lixian Xiao 외 arxiv

Indonesian, spoken by over 200 million people, remains underserved in multimodal emotion recognition research despite its dominant presence on Southeast Asian social media platforms. We introduce IndoMER, the first multi…

Multimodal Emotion RecognitionKeyword Extraction

KG-DF: A Black-box Defense Framework against Jailbreak Attacks Based on Knowledge Graphs

2025-11-09 · Shuyuan Liu, Jiawei Chen, Xiao Yang, Hang Su 외 arxiv

With the widespread application of large language models (LLMs) in various fields, the security challenges they face have become increasingly prominent, especially the issue of jailbreak. These attacks induce the model t…

Keyword ExtractionGeneral KnowledgeKnowledge GraphsSemantic Parsing

Rethinking Schema Linking: A Context-Aware Bidirectional Retrieval Approach for Text-to-SQL

2025-10-16 · Md Mahadi Hasan Nahid, Davood Rafiei, Weiwei Zhang, Yong Zhang arxiv

Schema linking -- the process of aligning natural language questions with database schema elements -- is a critical yet underexplored component of Text-to-SQL systems. While recent methods have focused primarily on impro…

Keyphrase ExtractionKeyword Extraction

PromptGuard at BLP-2025 Task 1: A Few-Shot Classification Framework Using Majority Voting and Keyword Similarity for Bengali Hate Speech Detection

2025-10-10 · Rakib Hossan, Shubhashis Roy Dipta arxiv

The BLP-2025 Task 1A requires Bengali hate speech classification into six categories. Traditional supervised approaches need extensive labeled datasets that are expensive for low-resource languages. We developed PromptGu…

Hate Speech DetectionKeyword Extraction

LLM as Attention-Informed NTM and Topic Modeling as long-input Generation: Interpretability and long-Context Capability

2025-10-03 · Xuan Xu, Zhongliang Yang, Haolun Li, Beilin Chu 외 arxiv

Topic modeling aims to produce interpretable topic representations and topic--document correspondences from corpora, but classical neural topic models (NTMs) remain constrained by limited representation assumptions and s…

Keyword ExtractionTopic Models

Adaptive Graph Convolution and Semantic-Guided Attention for Multimodal Risk Detection in Social Networks

2025-09-21 · Cuiqianhe Du, Chia-En Chiang, Tianyi Huang, Zikun Cui arxiv

This paper focuses on the detection of potentially dangerous tendencies of social media users in an innovative multimodal way. We integrate Natural Language Processing (NLP) and Graph Neural Networks (GNNs) together. Fir…

Keyword Extraction

DeKeyNLU: Enhancing Natural Language to SQL Generation through Task Decomposition and Keyword Extraction

2025-09-18 · Jian Chen, Zhenyan Chen, Xuming Hu, Peilin Zhou 외 arxiv

Natural Language to SQL (NL2SQL) provides a new model-centric paradigm that simplifies database access for non-technical users by converting natural language queries into SQL commands. Recent advancements, particularly t…

Natural Language QueriesKeyword Extraction

GOSU: Retrieval-Augmented Generation with Global-Level Optimized Semantic Unit-Centric Framework

2025-08-30 · Xuecheng Zou, Ke Liu, Bingbing Wang, Huafei Deng 외 arxiv

Building upon the standard graph-based Retrieval-Augmented Generation (RAG), the introduction of heterogeneous graphs and hypergraphs aims to enrich retrieval and generation by leveraging the relationships between multip…

Coreference ResolutionKeyword Extraction

Beyond the Black Box: Integrating Lexical and Semantic Methods in Quantitative Discourse Analysis with BERTopic

2025-08-26 · Thomas Compton arxiv

Quantitative Discourse Analysis has seen growing adoption with the rise of Large Language Models and computational tools. However, reliance on black box software such as MAXQDA and NVivo risks undermining methodological …

Dimensionality ReductionKeyword Extraction

A Large-Scale Benchmark for Evaluating Large Language Models on Medical Question Answering in Romanian

2025-08-22 · Ana-Cristina Rogoz, Radu Tudor Ionescu, Alexandra-Valentina Anghel, Ionut-Lucian Antone-Iordache 외 arxiv

We introduce MedQARo, the first large-scale medical QA benchmark in Romanian, alongside a comprehensive evaluation of state-of-the-art large language models (LLMs). We construct a high-quality and large-scale dataset com…

Keyword ExtractionQuestion Answering

Cost-Efficient Serving of LLM Agents via Test-Time Plan Caching

2025-06-17 · Qizheng Zhang, Michael Wornow, Kunle Olukotun

LLM-based agentic applications have shown increasingly remarkable capabilities in complex workflows but incur substantial costs due to extensive planning and reasoning requirements. Existing LLM caching techniques (like …

Keyword Extraction
1–20 / 191 다음 →