paper-with-me

Papers

Multi-modal Collaborative Optimization and Expansion Network for Event-assisted Single-eye Expression Recognition

2025-05-17 · Runduo Han, Xiuping Liu, Shangxuan Yi, Yi Zhang, Hongchen Tan

In this paper, we proposed a Multi-modal Collaborative Optimization and Expansion Network (MCO-E Net), to use event modalities to resist challenges such as low light, high exposure, and high dynamic range in single-eye expression recognition tasks. The MCO-E Net introduces two innovative designs: Multi-modal Collaborative Optimization Mamba (MCO-Mamba) and Heterogeneous Collaborative and Expansion Mixture-of-Experts (HCE-MoE). MCO-Mamba, building upon Mamba, leverages dual-modal information to jointly optimize the model, facilitating collaborative interaction and fusion of modal semantics. This approach encourages the model to balance the learning of both modalities and harness their respective strengths. HCE-MoE, on the other hand, employs a dynamic routing mechanism to distribute structurally varied experts (deep, attention, and focal), fostering collaborative learning of complementary semantics. This heterogeneous architecture systematically integrates diverse feature extraction paradigms to comprehensively capture expression semantics. Extensive experiments demonstrate that our proposed network achieves competitive performance in the task of single-eye expression recognition, especially under poor lighting conditions.

📄 PDF Abstract BibTeX arXiv:2505.12007

Code (1)

hrdhrd/MCO-E-Net 공식 구현

Tasks

Deep AttentionMambaMixture-of-Experts

Methods 이 논문이 사용한 방법론

Mamba Foundation models, now powering most of the exciting applications in deep learning, are almost universally based on the Transformer architecture and its core attention module.…

Similar Papers 제목 키워드 기반

Multi-Agent Amodal Completion: Direct Synthesis with Fine-Grained Semantic Guidance

2025-09-22 · Hongxing Fan, Lipeng Wang, Haohua Chen, Zehuan Huang 외 arxiv

Amodal completion, generating invisible parts of occluded objects, is vital for applications like image editing and AR. Prior methods face challenges with data needs, generalization, or error accumulation in progressive …

Image Editing

ISEP: Implicit Support Expansion for Offline Reinforcement Learning via Stochastic Policy Optimization

2026-05-18 · Yifei Chen, Shaoqin Zhu, Xiaoqiang Ji arxiv

Offline reinforcement learning methods typically enforce strict constraints to ensure safety; yet this rigidity often prevents the discovery of optimal behaviors outside the immediate support of the behavior policy. To a…

Reinforcement Learning

CoRE: Concept-Reasoning Expansion for Continual Brain Lesion Segmentation

2026-04-28 · Qianqian Chen, Anglin Liu, Jingyang Zhang, Yudong Zhang arxiv

Accurate brain lesion segmentation in MRI is vital for effective clinical diagnosis and treatment planning. Due to high annotation costs and strict data privacy regulations, universal models require employing Continual L…

Lesion SegmentationContinual Learning

Asynchronous Collaborative Graph Representation for Frames and Events

2025-01-01 · CVPR 2025 1 · Dianze Li, Jianing Li, Xu Liu, Xiaopeng Fan 외

Integrating frames and events has become a widely accepted solution for various tasks in challenging scenarios. However, most multimodal methods directly convert events into image-like formats synchronized with frame…

Depth EstimationDomain Adaptationobject-detectionObject Detection

Focus Through Motion: RGB-Event Collaborative Token Sparsification for Efficient Object Detection

2025-09-04 · Nan Yang, Yang Wang, Zhanwen Liu, Yuchao Dai 외 arxiv

Existing RGB-Event detection methods process the low-information regions of both modalities (background in images and non-event regions in event data) uniformly during feature extraction and fusion, resulting in high com…

Object Detection