paper-with-me

TriBERT

홈페이지 · 논문 3편

TriBERT dataset consists of 12,049 training, 2,527 validation and 2,560 test Human-Machine collaborative texts. Each text contains both human-written and LLM-generated parts, which can appear in different orders (human → AI, AI → human). Therefore, each sample has between 1 and 3 boundaries, indicating the sentences where authorship changes. The texts were created using humanwritten essays with LLM-generated sections added using ChatGPT.

Texts English

벤치마크

Boundary Detection on TriBERT (in-domain) 결과 2개