paper-with-me

홈 › Papers

Discriminating sample groups with multi-way data

2016-06-26 · Tianmeng Lyu, Eric F. Lock, Lynn E. Eberly

High-dimensional linear classifiers, such as the support vector machine (SVM) and distance weighted discrimination (DWD), are commonly used in biomedical research to distinguish groups of subjects based on a large number of features. However, their use is limited to applications where a single vector of features is measured for each subject. In practice data are often multi-way, or measured over multiple dimensions. For example, metabolite abundance may be measured over multiple regions or tissues, or gene expression may be measured over multiple time points, for the same subjects. We propose a framework for linear classification of high-dimensional multi-way data, in which coefficients can be factorized into weights that are specific to each dimension. More generally, the coefficients for each measurement in a multi-way dataset are assumed to have low-rank structure. This framework extends existing classification techniques, and we have implemented multi-way versions of SVM and DWD. We describe informative simulation results, and apply multi-way DWD to data for two very different clinical research studies. The first study uses metabolite magnetic resonance spectroscopy data over multiple brain regions to compare patients with and without spinocerebellar ataxia, the second uses publicly available gene expression time-course data to compare treatment responses for patients with multiple sclerosis. Our method improves performance and simplifies interpretation over naive applications of full rank linear classification to multi-way data. An R package is available at https://github.com/lockEF/MultiwayClassification .

📄 PDF Abstract BibTeX arXiv:1606.08046

Code (1)

lockEF/MultiwayClassification 공식 구현

Tasks

General Classification

Methods 이 논문이 사용한 방법론

SVM A Support Vector Machine, or SVM, is a non-parametric supervised learning model. For non-linear classification and regression, they utilise the kernel trick to map inputs…

Similar Papers 제목 키워드 기반

Discriminating between Similar Languages using Weighted Subword Features

2017-04-01 · WS 2017 4 · Adrien Barbaresi

The present contribution revolves around a contrastive subword n-gram model which has been tested in the Discriminating between Similar Languages shared task. I present and discuss the method used in this 14-way language…

Language IdentificationText Categorization

Discriminating between Similar Languages with Word-level Convolutional Neural Networks

2017-04-01 · WS 2017 4 · Marcelo Criscuolo, S Alu{\'\i}sio, ra Maria

Discriminating between Similar Languages (DSL) is a challenging task addressed at the VarDial Workshop series. We report on our participation in the DSL shared task with a two-stage system. In the first stage, character …

Language IdentificationQuestion AnsweringText Classification

Checking Functional Modularity in DNN By Biclustering Task-specific Hidden Neurons

2019-09-11 · NeurIPS Workshop Neuro_AI 2019 12 · Jialin Lu, Martin Ester

While real brain networks exhibit functional modularity, we investigate whether functional mod- ularity also exists in Deep Neural Networks (DNN) trained through back-propagation. Under the hypothesis that DNN are also o…

Discriminating between Similar Languages Using a Combination of Typed and Untyped Character N-grams and Words

2017-04-01 · WS 2017 4 · Helena Gomez, Ilia Markov, Jorge Baptista, Grigori Sidorov 외

This paper presents the cic{\_}ualg{'}s system that took part in the Discriminating between Similar Languages (DSL) shared task, held at the VarDial 2017 Workshop. This year{'}s task aims at identifying 14 languages acro…

General ClassificationInformation RetrievalMachine Translation

An Unsupervised Morphological Criterion for Discriminating Similar Languages

2016-12-01 · WS 2016 12 · Adrien Barbaresi

In this study conducted on the occasion of the Discriminating between Similar Languages shared task, I introduce an additional decision factor focusing on the token and subtoken level. The motivation behind this submissi…

Language IdentificationText Categorization