paper-with-me

Papers

Analyzing Similarity in Mathematical Content To Enhance the Detection of Academic Plagiarism

2018-01-25 · Isele Maurice-Roman

Despite the effort put into the detection of academic plagiarism, it continues to be a ubiquitous problem spanning all disciplines. Various tools have been developed to assist human inspectors by automatically identifying suspicious documents. However, to our knowledge currently none of these tools use mathematical content for their analysis. This is problematic, because mathematical content potentially represents a significant amount of the scientific contribution in academic documents. Hence, ignoring mathematical content limits the detection of plagiarism considerably, especially in disciplines with frequent use of mathematics. This paper aims to help close this gap by providing an overview of existing approaches in mathematical information retrieval and an analysis of their applicability for different possible cases of mathematical plagiarism. I find that whereas syntax-based approaches perform particularly well in detecting undisguised plagiarism, structure-based and hybrid approaches promise to also detect forms of disguised mathematical plagiarism, such as plagiarism with renamed identifiers. However, more research in this area is needed to enable the detection of more complex mathematical plagiarism: the scope of current approaches is restricted to the formula-level, an extension to the section-level is needed. Additionally, the general detection of equivalence transformations is currently not feasible. Despite these remaining problems, I conclude that the presented approaches could already be used for a basic automated detection system targeting mathematical plagiarism and therefore enhance current plagiarism detection systems.

📄 PDF Abstract BibTeX arXiv:1801.08439

Code (0)

등록된 구현이 없습니다.

Tasks

Information RetrievalRetrieval

Similar Papers 제목 키워드 기반

Improving Academic Plagiarism Detection for STEM Documents by Analyzing Mathematical Content and Citations

2019-06-27 · Norman Meuschke, Vincent Stange, Moritz Schubotz, Michael Karmer 외

Identifying academic plagiarism is a pressing task for educational and research institutions, publishers, and funding agencies. Current plagiarism detection systems reliably find instances of copied and moderately reword…

Math

Analyzing Non-Textual Content Elements to Detect Academic Plagiarism

2021-06-10 · Norman Meuschke

Identifying academic plagiarism is a pressing problem, among others, for research institutions, publishers, and funding organizations. Detection approaches proposed so far analyze lexical, syntactical, and semantic text …

Mathtext similarity

Taxonomy of Mathematical Plagiarism

2024-01-30 · Ankit Satpute, Andre Greiner-Petter, Noah Gießing, Isabel Beckenbach 외

Plagiarism is a pressing concern, even more so with the availability of large language models. Existing plagiarism detection systems reliably find copied and moderately reworded text but fail for idea plagiarism, especia…

MathQuestion AnsweringRecommendation Systems

TEIMMA: The First Content Reuse Annotator for Text, Images, and Math

2023-05-22 · Ankit Satpute, André Greiner-Petter, Moritz Schubotz, Norman Meuschke 외

This demo paper presents the first tool to annotate the reuse of text, images, and mathematical formulae in a document pair -- TEIMMA. Annotating content reuse is particularly useful to develop plagiarism detection algor…

Math

SearchLLM: Detecting LLM Paraphrased Text by Measuring the Similarity with Regeneration of the Candidate Source via Search Engine

2026-01-23 · Hoang-Quoc Nguyen-Son, Minh-Son Dao, Koji Zettsu arxiv

With the advent of large language models (LLMs), it has become common practice for users to draft text and utilize LLMs to enhance its quality through paraphrasing. However, this process can sometimes result in the loss …