paper-with-me

홈 › Papers

Co-Driven Recognition of Semantic Consistency via the Fusion of Transformer and HowNet Sememes Knowledge

2023-02-21 · Fan Chen, Yan Huang, Xinfang Zhang, Kang Luo, Jinxuan Zhu, Ruixian He

Semantic consistency recognition aims to detect and judge whether the semantics of two text sentences are consistent with each other. However, the existing methods usually encounter the challenges of synonyms, polysemy and difficulty to understand long text. To solve the above problems, this paper proposes a co-driven semantic consistency recognition method based on the fusion of Transformer and HowNet sememes knowledge. Multi-level encoding of internal sentence structures via data-driven is carried out firstly by Transformer, sememes knowledge base HowNet is introduced for knowledge-driven to model the semantic knowledge association among sentence pairs. Then, interactive attention calculation is carried out utilizing soft-attention and fusion the knowledge with sememes matrix. Finally, bidirectional long short-term memory network (BiLSTM) is exploited to encode the conceptual semantic information and infer the semantic consistency. Experiments are conducted on two financial text matching datasets (BQ, AFQMC) and a cross-lingual adversarial dataset (PAWSX) for paraphrase identification. Compared with lightweight models including DSSM, MwAN, DRCN, and pre-training models such as ERNIE etc., the proposed model can not only improve the accuracy of semantic consistency recognition effectively (by 2.19%, 5.57% and 6.51% compared with the DSSM, MWAN and DRCN models on the BQ dataset), but also reduce the number of model parameters (to about 16M). In addition, driven by the HowNet sememes knowledge, the proposed method is promising to adapt to scenarios with long text.

📄 PDF Abstract BibTeX arXiv:2302.10570

Code (1)

platanus-hy/sememes_codriven_text_matching 공식 구현 pytorch

Tasks

Paraphrase IdentificationSentenceText Matching

Methods 이 논문이 사용한 방법론

Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
Label Smoothing Label Smoothing is a regularization technique that introduces noise for the labels. This accounts for the fact that datasets may have mistakes in them, so maximizing the…
Dense Connections Dense Connections, or Fully Connected Connections, are a type of layer in a deep neural network that use a linear operation where every input is connected to every output…
Absolute Position Encodings Absolute Position Encodings are a type of position embeddings for [Transformer-based models] where positional encodings are…
Position-Wise Feed-Forward Layer 설명 없음
Adam 설명 없음
Multi-Head Attention 설명 없음

Similar Papers 제목 키워드 기반

MIND: Multimodal Intent-Driven Network via Diffusion Transformers for Medical Image Fusion

2026-07-30 · Yunzhan Fu, Xiangyu Shen, Yifei Sun, Yuhan Chen 외 arxiv

Medical image fusion aims to integrate complementary information from diverse imaging modalities to support clinical diagnosis. Existing methods typically apply uniform fusion rules globally, lacking a deep understanding…

Brain Tumor Segmentation

Multi Self-supervised Pre-fine-tuned Transformer Fusion for Better Intelligent Transportation Detection

2023-10-17 · Juwu Zheng, Jiangtao Ren

Intelligent transportation system combines advanced information technology to provide intelligent services such as monitoring, detection, and early warning for modern transportation. Intelligent transportation detection …

object-detectionObject DetectionSelf-Supervised Learning

KGEdit: Ambiguity-Aware Knowledge Graphs for Training-Free Precise Video Generation and Editing

2026-05-28 · Mingshu Cai, Miao Zhang, Chenghe Yang, Yixuan Li 외 arxiv

In recent years, training-free video generation has progressed remarkably. However, when handling complex textual instructions, existing methods still suffer from semantic ambiguity, incorrect concept binding, and cross-…

Knowledge GraphsVideo Generation

CoSyncDiT: Cognitive Synchronous Diffusion Transformer for Movie Dubbing

2026-04-14 · Gaoxiang Cong, Liang Li, Jiaxin Ye, Zhedong Zhang 외 arxiv

Movie dubbing aims to synthesize speech that preserves the vocal identity of a reference audio while synchronizing with the lip movements in a target video. Existing methods fail to achieve precise lip-sync and lack natu…

FutrTrack: A Camera-LiDAR Fusion Transformer for 3D Multiple Object Tracking

2025-10-22 · Martha Teiko Teye, Ori Maoz, Matthias Rottmann arxiv

We propose FutrTrack, a modular camera-LiDAR multi-object tracking framework that builds on existing 3D detectors by introducing a transformer-based smoother and a fusion-driven tracker. Inspired by query-based tracking …

Multiple Object TrackingMulti-Object Tracking