輔助建立混合性系統之自然語言處理系統深度評估平台 --- 以評估依存關係分析器為例 (The Platform providing NLP System Deep Comparative Evaluation and Auxiliary Information for Hybrid NLP System Building: Trial on Dependency Parser Evaluation) [In Chinese]
Code (0)
등록된 구현이 없습니다.
Tasks
Dependency ParsingSimilar Papers 제목 키워드 기반
LLM BiasScope: A Real-Time Bias Analysis Platform for Comparative LLM Evaluation
As large language models (LLMs) are deployed widely, detecting and understanding bias in their outputs is critical. We present LLM BiasScope, a web application for side-by-side comparison of LLM outputs with real-time bi…
Bias DetectionSEAGLE: A Platform for Comparative Evaluation of Semantic Encoders for Information Retrieval
We introduce Seagle, a platform for comparative evaluation of semantic text encoding models on information retrieval (IR) tasks. Seagle implements (1) word embedding aggregators, which represent texts as algebraic aggreg…
Information RetrievalRetrievalSentenceSentence Retrieval+1A New Era: Intelligent Tutoring Systems Will Transform Online Learning for Millions
Despite artificial intelligence (AI) having transformed major aspects of our society, less than a fraction of its potential has been explored, let alone deployed, for education. AI-powered learning can provide millions o…
Active LearningMultiple-choiceEvaluating LLMs When They Do Not Know the Answer: Statistical Evaluation of Mathematical Reasoning via Comparative Signals
Evaluating mathematical reasoning in LLMs is constrained by limited benchmark sizes and inherent model stochasticity, yielding high-variance accuracy estimates and unstable rankings across platforms. On difficult problem…
Mathematical ReasoningLinguistAgent: A Reflective Multi-Model Platform for Automated Linguistic Annotation
Data annotation remains a significant bottleneck in the Humanities and Social Sciences, particularly for complex semantic tasks such as metaphor identification. While Large Language Models (LLMs) show promise, a signific…
Prompt Engineering