From keywords to semantics: Perceptions of large language models in data discovery
Current approaches to data discovery match keywords between metadata and queries. This matching requires researchers to know the exact wording that other researchers previously used, creating a challenging process that could lead to missing relevant data. Large Language Models (LLMs) could enhance data discovery by removing this requirement and allowing researchers to ask questions with natural language. However, we do not currently know if researchers would accept LLMs for data discovery. Using a human-centered artificial intelligence (HCAI) focus, we ran focus groups (N = 27) to understand researchers' perspectives towards LLMs for data discovery. Our conceptual model shows that the potential benefits are not enough for researchers to use LLMs instead of current technology. Barriers prevent researchers from fully accepting LLMs, but features around transparency could overcome them. Using our model will allow developers to incorporate features that result in an increased acceptance of LLMs for data discovery.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Toward Explainable Users: Using NLP to Enable AI to Understand Users' Perceptions of Cyber Attacks
To understand how end-users conceptualize consequences of cyber security attacks, we performed a card sorting study, a well-known technique in Cognitive Sciences, where participants were free to group the given consequen…
SentenceStereoMap: Quantifying the Awareness of Human-like Stereotypes in Large Language Models
Large Language Models (LLMs) have been observed to encode and perpetuate harmful associations present in the training data. We propose a theoretically grounded framework called StereoMap to gain insights into their perce…
Reconsidering Annotator Disagreement about Racist Language: Noise or Signal?
An abundance of methodological work aims to detect hateful and racist language in text. However, these tools are hampered by problems like low annotator agreement and remain largely disconnected from theoretical work on …
DescriptiveSemantic Arabic Information Retrieval Framework
The continuous increasing in the amount of the published and stored information requires a special Information Retrieval (IR) frameworks to search and get information accurately and speedily. Currently, keywords-based te…
Information RetrievalRetrievalFully Unsupervised Training of Few-shot Keyword Spotting
For training a few-shot keyword spotting (FS-KWS) model, a large labeled dataset containing massive target keywords has known to be essential to generalize to arbitrary target keywords with only a few enrollment samples.…
Keyword SpottingMetric LearningSpeech Synthesis