paper-with-me

홈 › Papers

PaCo2: A Fully Automated tool for gathering Parallel Corpora from the Web

2012-05-01 · LREC 2012 5 · I{\~n}aki San Vicente, Iker Manterola

The importance of parallel corpora in the NLP field is fully acknowledged. This paper presents a tool that can build parallel corpora given just a seed word list and a pair of languages. Our approach is similar to others proposed in the literature, but introduces a new phase to the process. While most of the systems leave the task of finding websites containing parallel content up to the user, PaCo2 (Parallel Corpora Collector) takes care of that as well. The tool is language independent as far as possible, and adapting the system to work with new languages is fairly straightforward. Evaluation of the different modules has been carried out for Basque-Spanish, Spanish-English and Portuguese-English language pairs. Even though there is still room for improvement, results are positive. Results show that the corpora created have very good quality translations units, and the quality is maintained for the various language pairs. Details of the corpora created up until now are also provided.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

PaCoNet: Deep Data Extraction for Parallel Coordinates

2026-08-06 · Poonam Poonam, Hannah Kniesel, Pere-Pau Vázquez, Timo Ropinski arxiv

Extracting data from visualizations has long challenged computer vision, with current research focused on bar, line, and pie charts, among other low-dimensional visualizations. However, parallel coordinates as a widely u…

PaCo-RL: Advancing Reinforcement Learning for Consistent Image Generation with Pairwise Reward Modeling

2025-12-02 · Bowen Ping, Chengyou Jia, Minnan Luo, Changliang Xia 외 arxiv

Consistent image generation requires faithfully preserving identities, styles, and logical coherence across multiple images, which is essential for applications such as storytelling and character design. Supervised train…

Reinforcement LearningImage Generation

PaCoRe: Learning to Scale Test-Time Compute with Parallel Coordinated Reasoning

2026-01-09 · Jingcheng Hu, Yinmin Zhang, Shijie Shang, Xiaobo Yang 외 arxiv

We introduce Parallel Coordinated Reasoning (PaCoRe), a training-and-inference framework designed to overcome a central limitation of contemporary language models: their inability to scale test-time compute (TTC) far bey…

Reinforcement Learning

DeepACO: Neural-enhanced Ant Systems for Combinatorial Optimization

2023-09-25 · NeurIPS 2023 11 · Haoran Ye, Jiarui Wang, Zhiguang Cao, Helan Liang 외

Ant Colony Optimization (ACO) is a meta-heuristic algorithm that has been successfully applied to various Combinatorial Optimization Problems (COPs). Traditionally, customizing ACO for a specific problem requires the exp…

Combinatorial OptimizationDeep Reinforcement Learning

Object-level Scene Deocclusion

2024-06-11 · Zhengzhe Liu, Qing Liu, Chirui Chang, Jianming Zhang 외

Deoccluding the hidden portions of objects in a scene is a formidable task, particularly when addressing real-world scenes. In this paper, we present a new self-supervised PArallel visible-to-COmplete diffusion framework…

3D Scene ReconstructionObjectSelf-Supervised Learning