paper-with-me

Papers

輔助建立混合性系統之自然語言處理系統深度評估平台 --- 以評估依存關係分析器為例 (The Platform providing NLP System Deep Comparative Evaluation and Auxiliary Information for Hybrid NLP System Building: Trial on Dependency Parser Evaluation) [In Chinese]

2018-10-01 · ROCLINGIJCLCLP 2018 10 · Yi-siang Wang
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Dependency Parsing

Similar Papers 제목 키워드 기반

LLM BiasScope: A Real-Time Bias Analysis Platform for Comparative LLM Evaluation

2026-03-12 · Himel Ghosh, Nick Elias Werner arxiv

As large language models (LLMs) are deployed widely, detecting and understanding bias in their outputs is critical. We present LLM BiasScope, a web application for side-by-side comparison of LLM outputs with real-time bi…

Bias Detection

SEAGLE: A Platform for Comparative Evaluation of Semantic Encoders for Information Retrieval

2019-11-01 · IJCNLP 2019 11 · Fabian David Schmidt, Markus Dietsche, Simone Paolo Ponzetto, Goran Glava{\v{s}}

We introduce Seagle, a platform for comparative evaluation of semantic text encoding models on information retrieval (IR) tasks. Seagle implements (1) word embedding aggregators, which represent texts as algebraic aggreg…

Information RetrievalRetrievalSentenceSentence Retrieval+1

A New Era: Intelligent Tutoring Systems Will Transform Online Learning for Millions

2022-03-03 · Francois St-Hilaire, Dung Do Vu, Antoine Frau, Nathan Burns 외

Despite artificial intelligence (AI) having transformed major aspects of our society, less than a fraction of its potential has been explored, let alone deployed, for education. AI-powered learning can provide millions o…

Active LearningMultiple-choice

Evaluating LLMs When They Do Not Know the Answer: Statistical Evaluation of Mathematical Reasoning via Comparative Signals

2026-02-03 · Zihan Dong, Zhixian Zhang, Yang Zhou, Can Jin 외 arxiv

Evaluating mathematical reasoning in LLMs is constrained by limited benchmark sizes and inherent model stochasticity, yielding high-variance accuracy estimates and unstable rankings across platforms. On difficult problem…

Mathematical Reasoning

LinguistAgent: A Reflective Multi-Model Platform for Automated Linguistic Annotation

2026-02-05 · Bingru Li arxiv

Data annotation remains a significant bottleneck in the Humanities and Social Sciences, particularly for complex semantic tasks such as metaphor identification. While Large Language Models (LLMs) show promise, a signific…

Prompt Engineering