paper-with-me

Papers

Computer Assisted Annotation of Tension Development in TED Talks through Crowdsourcing

2019-11-01 · WS 2019 11 · Seungwon Yoon, Wonsuk Yang, Jong Park

We propose a method of machine-assisted annotation for the identification of tension development, annotating whether the tension is increasing, decreasing, or staying unchanged. We use a neural network based prediction model, whose predicted results are given to the annotators as initial values for the options that they are asked to choose. By presenting such initial values to the annotators, the annotation task becomes an evaluation task where the annotators inspect whether or not the predicted results are correct. To demonstrate the effectiveness of our method, we performed the annotation task in both in-house and crowdsourced environments. For the crowdsourced environment, we compared the annotation results with and without our method of machine-assisted annotation. We find that the results with our method showed a higher agreement to the gold standard than those without, though our method had little effect at reducing the time for annotation. Our codes for the experiment are made publicly available.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

TEDxTN: A Three-way Speech Translation Corpus for Code-Switched Tunisian Arabic - English

2025-11-13 · Fethi Bougares, Salima Mdhaffar, Haroun Elleuch, Yannick Estève arxiv

In this paper, we introduce TEDxTN, the first publicly available Tunisian Arabic to English speech translation dataset. This work is in line with the ongoing effort to mitigate the data scarcity obstacle for a number of …

Speech Recognition

CataractSAM-2: A Domain-Adapted Model for Anterior Segment Surgery Segmentation and Scalable Ground-Truth Annotation

2026-03-23 · Mohammad Eslami, Dhanvinkumar Ganeshkumar, Saber Kazeminasab, Michael G. Morley 외 arxiv

We present CataractSAM-2, a domain-adapted extension of Meta's Segment Anything Model 2, designed for real-time semantic segmentation of cataract ophthalmic surgery videos with high accuracy. Positioned at the intersecti…

Real-Time Semantic SegmentationZero-shot Generalization

Shallow Discourse Annotation for Chinese TED Talks

2020-03-09 · LREC 2020 5 · Wanqiu Long, Xinyi Cai, James E. M. Reid, Bonnie Webber 외

Text corpora annotated with language-related properties are an important resource for the development of Language Technology. The current work contributes a new resource for Chinese Language Technology and for Chinese-En…

Translation

The SI TEDx-UM speech database: a new Slovenian Spoken Language Resource

2016-05-01 · LREC 2016 5 · Andrej {\v{Z}}gank, Mirjam Sepesy Mau{\v{c}}ec, Darinka Verdonik

This paper presents a new Slovenian spoken language resource built from TEDx Talks. The speech database contains 242 talks in total duration of 54 hours. The annotation and transcription of acquired spoken material was g…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modelling+2

Can We Make Computers Laugh at Talks?

2016-12-01 · WS 2016 12 · Chong Min Lee, Su-Youn Yoon, Lei Chen

Considering the importance of public speech skills, a system which makes a prediction on where audiences laugh in a talk can be helpful to a person who prepares for a talk. We investigated a possibility that a state-of-t…