paper-with-me

Text Clustering

3개 벤치마크 · 논문 134편 · 이 태스크의 논문 보기 →

Benchmarks

MTEB

결과 31개

20 Newsgroups

결과 2개

Most implemented

MTEB: Massive Text Embedding Benchmark

2022-10-13 · 구현 5개

Papers

CLUBench: A Clustering Benchmark

2026-05-28 · Feng Xiao, Dazhi Fu, Chris Ding, Jicong Fan arxiv

Clustering is a fundamental problem in data science with a long-standing research history, yielding numerous insightful algorithms. Despite this progress, a systematic and large-scale empirical evaluation that jointly co…

Text ClusteringDeep Clustering

TextClusterLab: An Integrated Framework for Reliable Text Clustering Studies

2026-05-17 · Daoming Wan, Yizheng Huang, Jimmy X. Huang arxiv

In recent years, text clustering has become a critical technique for applications including intent discovery, topic mining, and recommendation systems. However, evaluating text clustering algorithms remains challenging s…

Recommendation SystemsIntent DiscoveryText Clustering

AdaGraph: A Graph-Native Clustering Algorithm That Overcomes the Curse of Dimensionality and Enables Scientific Discovery

2026-05-05 · Ahmed Elmahdi arxiv

We present AdaGraph, a graph-native clustering algorithm born from the Structure-Centric Machine Learning (SC-ML) paradigm -- a new field of unsupervised learning that replaces geometry-centric (distance-based) computati…

Dimensionality ReductionText Clustering

Topeax -- An Improved Clustering Topic Model with Density Peak Detection and Lexical-Semantic Term Importance

2026-01-29 · Márton Kardos arxiv

Text clustering is today the most popular paradigm for topic modelling, both in academia and industry. Despite clustering topic models' apparent success, we identify a number of issues in Top2Vec and BERTopic, which rema…

Text ClusteringTopic Models

Optimized Algorithms for Text Clustering with LLM-Generated Constraints

2026-01-16 · Chaoqi Jia, Weihong Wu, Longkun Guo, Zhigang Lu 외 arxiv

Clustering is a fundamental tool that has garnered significant interest across a wide range of applications including text analysis. To improve clustering accuracy, many researchers have incorporated background knowledge…

Text Clustering

A Large-Language-Model Framework for Automated Humanitarian Situation Reporting

2025-12-22 · Ivan Decostanzi, Yelena Mejova, Kyriaki Kalimeri arxiv

Timely and accurate situational reports are essential for humanitarian decision-making, yet current workflows remain largely manual, resource intensive, and inconsistent. We present a fully automated framework that uses …

Question GenerationText Clustering

전체 134편 보기 →