paper-with-me

홈 › Papers

K-repeating Substrings: a String-Algorithmic Approach to Privacy-Preserving Publishing of Textual Data

2014-12-01 · PACLIC 2014 12 · Yusuke Matsubara, Koiti Hasida
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Privacy Preserving

Similar Papers 제목 키워드 기반

Double-Ended Palindromic Trees: A Linear-Time Data Structure and Its Applications

2022-10-05 · Qisheng Wang, Ming Yang, Xinrui Zhu

The palindromic tree (a.k.a. eertree) is a linear-size data structure that provides access to all palindromic substrings of a string. In this paper, we propose a generalized version of eertree, called double-ended eertre…

Chinese Word Segmentation by Mining Maximized Substrings

2013-10-01 · IJCNLP 2013 10 · Mo Shen, Daisuke Kawahara, Sadao Kurohashi
Chinese Word Segmentation

An Operator for Entity Extraction in MapReduce

2015-12-15 · Ndapandula Nakashole

Dictionary-based entity extraction involves finding mentions of dictionary entities in text. Text mentions are often noisy, containing spurious or missing words. Efficient algorithms for detecting approximate entity ment…

Entity Extraction using GAN

Lexis: An Optimization Framework for Discovering the Hierarchical Structure of Sequential Data

2016-02-17 · Payam Siyari, Bistra Dilkina, Constantine Dovrolis

Data represented as strings abounds in biology, linguistics, document mining, web search and many other fields. Such data often have a hierarchical structure, either because they were artificially designed and composed i…

Text Compression

Diffusion Models Preferentially Memorize Prototypical Examples or: Why Does My Diffusion Model Love Slop?

2026-05-28 · Marta Aparicio Rodriguez, Anastasia Borovykh, Grigorios A. Pavliotis, Daniel J. Korchinski arxiv

Generative models have a persistent limitation: their tendency to memorize training data can create legal liabilities and erode creative diversity. Understanding which samples are memorized in whole or in part, and under…