paper-with-me

홈 › Papers

Are Graph Embeddings the Panacea? An Empirical Survey from the Data Fitness Perspective

2024-04-25 · Pacific-Asia Conference on Knowledge Discovery and Data Mining 2024 4 · Qiang Sun, Du Q. Huynh, Mark Reynolds, Wei Liu

Graph representation learning has emerged as a machine learning go-to technique, outperforming traditional tabular view of data across many domains. Current surveys on graph representation learning predominantly have an algorithmic focus with the primary goal of explaining foundational principles and comparing performances, yet the natural and practical question “Are graph embeddings the panacea?” has been so far neglected. In this paper, we propose to examine graph embedding algorithms from a data fitness perspective by offering a methodical analysis that aligns network characteristics of data with appropriate embedding algorithms. The overarching objective is to provide researchers and practitioners with comprehensive and methodical investigations, enabling them to confidently answer pivotal questions confronting node classification problems: 1) Is there a potential benefit of applying graph representation learning? 2) Is structural information alone sufficient? 3) Which embedding technique would best suit my dataset? Through 1400 experiments across 35 datasets, we have evaluated four network embedding algorithms – three popular GNN-based algorithms (GraphSage, GCN, GAE) and node2vec – over traditional classification methods, namely SVM, KNN, and Random Forest (RF). Our results indicate that the cohesiveness of the network, the representation of relation information, and the number of classes in a classification problem play significant roles in algorithm selection.

📄 PDF Abstract BibTeX

Code (1)

PascalSun/PAKDD-2024 pytorch

Tasks

Graph ClassificationGraph EmbeddingGraph LearningGraph Neural NetworkGraph Representation LearningNetwork EmbeddingNode ClassificationRepresentation Learning

Methods 이 논문이 사용한 방법론

GCN A Graph Convolutional Network, or GCN, is an approach for semi-supervised learning on graph-structured data. It is based on an efficient variant of [convolutional neural…
SVM A Support Vector Machine, or SVM, is a non-parametric supervised learning model. For non-linear classification and regression, they utilise the kernel trick to map inputs…
Focus 설명 없음
node2vec node2vec is a framework for learning graph embeddings for nodes in graphs. Node2vec maximizes a likelihood objective over mappings which preserve neighbourhood distances in…

Similar Papers 제목 키워드 기반

PANACEA: An Automated Misinformation Detection System on COVID-19

2023-02-28 · Runcong Zhao, Miguel Arana-Catania, Lixing Zhu, Elena Kochkina 외

In this demo, we introduce a web-based misinformation detection system PANACEA on COVID-19 related claims, which has two modules, fact-checking and rumour detection. Our fact-checking module, which is supported by novel …

Fact CheckingMisinformationNatural Language InferenceRumour Detection

Panacea+: Panoramic and Controllable Video Generation for Autonomous Driving

2024-08-14 · Yuqing Wen, Yucheng Zhao, Yingfei Liu, Binyuan Huang 외

The field of autonomous driving increasingly demands high-quality annotated video training data. In this paper, we propose Panacea+, a powerful and universally applicable framework for generating video data in driving sc…

3D Object Detection3D Object TrackingAutonomous DrivingLane Detection+6

The Panacea Threat Intelligence and Active Defense Platform

2020-04-20 · Adam Dalton, Ehsan Aghaei, Ehab Al-Shaer, Archna Bhatia 외

We describe Panacea, a system that supports natural language processing (NLP) components for active defenses against social engineering attacks. We deploy a pipeline of human language technology, including Ask and Framin…

AttributeDialogue Generationnamed-entity-recognitionNamed Entity Recognition+1

Panacea: A foundation model for clinical trial search, summarization, design, and recruitment

2024-06-25 · Jiacheng Lin, Hanwen Xu, Zifeng Wang, Sheng Wang 외

Clinical trials are fundamental in developing new drugs, medical devices, and treatments. However, they are often time-consuming and have low success rates. Although there have been initial attempts to create large langu…

Clinical Knowledge

Panacea: Pareto Alignment via Preference Adaptation for LLMs

2024-02-03 · Yifan Zhong, Chengdong Ma, Xiaoyuan Zhang, Ziran Yang 외

Current methods for large language model alignment typically use scalar human preference labels. However, this convention tends to oversimplify the multi-dimensional and heterogeneous nature of human preferences, leading…

Language ModellingLarge Language Model