paper-with-me

홈 › Papers

Semantic Communities and Boundary-Spanning Lyrics in K-pop: A Graph-Based Unsupervised Analysis

2026-02-13 · Oktay Karakuş arxiv

Large-scale lyric corpora present unique challenges for data-driven analysis, including the absence of reliable annotations, multilingual content, and high levels of stylistic repetition. Most existing approaches rely on supervised classification, genre labels, or coarse document-level representations, limiting their ability to uncover latent semantic structure. We present a graph-based framework for unsupervised discovery and evaluation of semantic communities in K-pop lyrics using line-level semantic representations. By constructing a similarity graph over lyric texts and applying community detection, we uncover stable micro-theme communities without genre, artist, or language supervision. We further identify boundary-spanning songs via graph-theoretic bridge metrics and analyse their structural properties. Across multiple robustness settings, boundary-spanning lyrics exhibit higher lexical diversity and lower repetition compared to core community members, challenging the assumption that hook intensity or repetition drives cross-theme connectivity. Our framework is language-agnostic and applicable to unlabeled cultural text corpora.

📄 PDF Abstract BibTeX arXiv:2602.12881

Code (0)

등록된 구현이 없습니다.

Tasks

Community Detection

Similar Papers 제목 키워드 기반

Bob's Confetti: Phonetic Memorization Attacks in Music and Video Generation

2025-07-23 · Jaechul Roh, Zachary Novack, Yuefeng Peng, Niloofar Mireshghallah 외 arxiv

Generative AI systems for music and video commonly use text-based filters to prevent regurgitation of copyrighted material. We expose a significant vulnerability in this approach by introducing Adversarial PhoneTic Promp…

Cross-Modal RetrievalSemantic SimilarityVideo Generation

Music- and Lyrics-driven Dance Synthesis

2023-09-30 · Wenjie Yin, Qingyuan Yao, Yi Yu, Hang Yin 외

Lyrics often convey information about the songs that are beyond the auditory dimension, enriching the semantic meaning of movements and musical themes. Such insights are important in the dance choreography domain. Howeve…

Triplet

Melody-Conditioned Lyrics Generation with SeqGANs

2020-10-28 · Yihao Chen, Alexander Lerch

Automatic lyrics generation has received attention from both music and AI communities for years. Early rule-based approaches have~---due to increases in computational power and evolution in data-driven models---~mostly b…

Lyrics: Boosting Fine-grained Language-Vision Alignment and Comprehension via Semantic-aware Visual Objects

2023-12-08 · Junyu Lu, Dixiang Zhang, Songxin Zhang, Zejian Xie 외

Large Vision Language Models (LVLMs) have demonstrated impressive zero-shot capabilities in various vision-language dialogue scenarios. However, the absence of fine-grained visual object detection hinders the model from …

Image Captioningobject-detectionObject DetectionReferring Expression Comprehension+3

LM2D: Lyrics- and Music-Driven Dance Synthesis

2024-03-14 · Wenjie Yin, Xuejiao Zhao, Yi Yu, Hang Yin 외

Dance typically involves professional choreography with complex movements that follow a musical rhythm and can also be influenced by lyrical content. The integration of lyrics in addition to the auditory dimension, enric…

Motion GenerationPose EstimationRhythm