Conception: Multilingually-Enhanced, Human-Readable Concept Vector Representations
To date, the most successful word, word sense, and concept modelling techniques have used large corpora and knowledge resources to produce dense vector representations that capture semantic similarities in a relatively low-dimensional space. Most current approaches, however, suffer from a monolingual bias, with their strength depending on the amount of data available across languages. In this paper we address this issue and propose Conception, a novel technique for building language-independent vector representations of concepts which places multilinguality at its core while retaining explicit relationships between concepts. Our approach results in high-coverage representations that outperform the state of the art in multilingual and cross-lingual Semantic Word Similarity and Word Sense Disambiguation, proving particularly robust on low-resource languages. Conception {--} its software and the complete set of representations {--} is available at https://github.com/SapienzaNLP/conception.
Code (1)
Tasks
Word Sense DisambiguationWord SimilaritySimilar Papers 제목 키워드 기반
Finnish 5th and 6th graders' misconceptions about Artificial Intelligence
Research on children's initial conceptions of AI is in an emerging state, which, from a constructivist viewpoint, challenges the development of pedagogically sound AI-literacy curricula, methods, and materials. To contri…
MisconceptionsFix the Mind, Not the Move: Interpretable AI Assistance via Knowledge-Gap Localization
AI assistants in human-AI collaboration often correct suboptimal human actions through behavioral feedback (e.g., alerts or steering-wheel nudges in assistive driving). Such interventions can mitigate immediate errors, b…
Simulating Students or Sycophantic Problem Solving? On Misconception Faithfulness of LLM Simulators
Large language models (LLMs) can fluently generate student-like responses, making them attractive as simulated students for training and evaluating AI tutors and human educators. Yet such simulators are typically evaluat…
Reinforcement LearningProbing the Mind Behind the (Literal and Figurative) Lightbulb
After doing away with the evolutionary scaffold for BVSR, what remains is a notion of "blindness" that does not distinguish BVSR from other theories of creativity, and an assumption that creativity can be understood by t…
Data-Mining Textual Responses to Uncover Misconception Patterns
An important, yet largely unstudied, problem in student data analysis is to detect misconceptions from students' responses to open-response questions. Misconception detection enables instructors to deliver more targeted …
Misconceptions