paper-with-me

Papers

ClaimCompare: A Data Pipeline for Evaluation of Novelty Destroying Patent Pairs

2024-07-16 · Arav Parikh, Shiri Dori-Hacohen

A fundamental step in the patent application process is the determination of whether there exist prior patents that are novelty destroying. This step is routinely performed by both applicants and examiners, in order to assess the novelty of proposed inventions among the millions of applications filed annually. However, conducting this search is time and labor-intensive, as searchers must navigate complex legal and technical jargon while covering a large amount of legal claims. Automated approaches using information retrieval and machine learning approaches to detect novelty destroying patents present a promising avenue to streamline this process, yet research focusing on this space remains limited. In this paper, we introduce a novel data pipeline, ClaimCompare, designed to generate labeled patent claim datasets suitable for training IR and ML models to address this challenge of novelty destruction assessment. To the best of our knowledge, ClaimCompare is the first pipeline that can generate multiple novelty destroying patent datasets. To illustrate the practical relevance of this pipeline, we utilize it to construct a sample dataset comprising of over 27K patents in the electrochemical domain: 1,045 base patents from USPTO, each associated with 25 related patents labeled according to their novelty destruction towards the base patent. Subsequently, we conduct preliminary experiments showcasing the efficacy of this dataset in fine-tuning transformer models to identify novelty destroying patents, demonstrating 29.2% and 32.7% absolute improvement in MRR and P@1, respectively.

📄 PDF Abstract BibTeX arXiv:2407.12193

Code (1)

RIET-lab/claim-compare 공식 구현

Tasks

Information RetrievalNavigate

Methods 이 논문이 사용한 방법론

BASE 설명 없음

Similar Papers 제목 키워드 기반

LLM generation novelty through the lens of semantic similarity

2025-10-31 · Philipp Davydov, Ameya Prabhu, Matthias Bethge, Elisa Nguyen 외 arxiv

Generation novelty is a key indicator of an LLM's ability to generalize, yet measuring it against full pretraining corpora is computationally challenging. Existing evaluations often rely on lexical overlap, failing to de…

Semantic SimilaritySemantic Retrieval

Counterfactual contrastive learning: robust representations via causal image synthesis

2024-03-14 · Melanie Roschewitz, Fabio De Sousa Ribeiro, Tian Xia, Galvin Khara 외

Contrastive pretraining is well-known to improve downstream task performance and model generalisation, especially in limited label settings. However, it is sensitive to the choice of augmentation pipeline. Positive pairs…

Contrastive LearningcounterfactualCounterfactual InferenceImage Generation

A Unified Evaluation Framework for Novelty Detection and Accommodation in NLP with an Instantiation in Authorship Attribution

2023-05-08 · Neeraj Varshney, Himanshu Gupta, Eric Robertson, Bing Liu 외

State-of-the-art natural language processing models have been shown to achieve remarkable performance in 'closed-world' settings where all the labels in the evaluation set are known at training time. However, in real-wor…

Authorship AttributionNovelty Detection

Evaluating Novelty in AI-Generated Research Plans Using Multi-Workflow LLM Pipelines

2025-12-24 · Devesh Saraogi, Rohit Singhee, Dhruv Kumar arxiv

The integration of Large Language Models (LLMs) into the scientific ecosystem raises fundamental questions about the creativity and originality of AI-generated research. Recent work has identified ``smart plagiarism'' as…

Creating and Destroying Party Brands

2014-06-01 · WS 2014 6 · Justin Grimmer