Text Classification
68개 벤치마크 · 논문 3,862편 · 이 태스크의 논문 보기 →
Benchmarks
MTEB
AG News
DBpedia
R8
TREC-6
20NEWS
UK Key Stage Readability
IMDb
MR
Ohsumed
Yahoo! Answers
Climabench
NewsDiscourse
R52
Yelp-5
DODF Data
Lot-insts
MVICTOR (type)
SVICTOR (type)
Yelp-2
Amazon-2
HateXplain
RCV1
arXiv-10
Amazon-5
BLURB
Bala-Copa
IMDb Movie Reviews
Overruling
SST-2
Sogou News
Terms of Service
GLUE SST2
MuLD (Character Type)
Searchsnippets
TREC-50
This is not a Dataset
20 Newsgroups
BANKING77
FMC-MWO2KG
Facebook Media
GLUE COLA
GLUE MRPC
GLUE RTE
GLUE STSB
Hyperpartisan
NICE-2
NICE-45
Patents
SILICONE Benchmark
STOPS-2
STOPS-41
TRAC2-Benghali. Task 2.
TRAC2-English. Task2.
TREC-10
Twitter-US
WNUT-2020 Task 2
Most implemented
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Semi-supervised Sequence Learning
RoBERTa: A Robustly Optimized BERT Pretraining Approach
Universal Language Model Fine-tuning for Text Classification
Bag of Tricks for Efficient Text Classification
FastText.zip: Compressing text classification models
Papers
Making Clinical Language Models Auditable: Concept-Guided Fine-Tuning for Robust Prediction
Clinical language models can achieve strong in-hospital accuracy yet fail under deployment shifts because they exploit note-specific artifacts (e.g., templates, separators, boilerplate) that do not reflect patient state.…
Mortality PredictionText ClassificationHelaBERT: Enhancing Sinhala Language Understanding with Dual Pooling Classification Head
We present HelaBERT, a family of two BERT-based masked language models pre-trained from scratch on approximately 1 billion tokens of Sinhala text sourced from MADLAD-400, CulturaX, and a custom corpus comprising news art…
Text ClassificationSentiment AnalysisMIL-BERT: Classification of Arbitrarily Large Text with Performance and Explanatory Guarantees
Many text classification decisions are viable based on constituent excerpts alone. Taking inspiration from the field of multiple instance learning, we present an algorithm for training a neural network to classify text b…
Multiple Instance LearningText ClassificationTask-CoEvolve: Efficient Harness Optimization via Adaptive Validation Task Selection
We present a novel approach to efficient LLM harness optimization through adaptive validation task selection. Harness optimization iteratively rewrites the harness code based on validation performance, enabling substanti…
Text ClassificationFraQ: Efficient Coordinate-Space Recompression for Federated Low-Rank Adaptation
Federated fine-tuning with Low-Rank Adaptation (LoRA) enables efficient collaborative adaptation of Large Language Models (LLMs) without centralizing private data. However, LoRA's two-factor parameterization creates an a…
Text ClassificationRelative Parameter Importance in Task-Agnostic Replay-Free Continual Learning
Achieving continual learning (CL) with deep neural networks requires balancing stability and plasticity while enabling knowledge transfer. In this work, we focus on offline learning algorithms under the constraints: (I) …
Incremental LearningText ClassificationContinual LearningText Generation