paper-with-me

Papers Attribute

“Attribute” 태그가 달린 논문 5,387편 · 필터 해제

MGFFD-VLM: Multi-Granularity Prompt Learning for Face Forgery Detection with VLM

2025-07-16 · Tao Chen, Jingyi Zhang, Decheng Liu, Chunlei Peng

Recent studies have utilized visual large language models (VLMs) to answer not only "Is this face a forgery?" but also "Why is the face a forgery?" These studies introduced forgery-related attributes, such as forgery loc…

AttributeFace SwappingPrompt LearningVisual Question Answering (VQA)

Non-Adaptive Adversarial Face Generation

2025-07-16 · Sunpill Kim, Seunghun Paik, Chanwoo Hwang, Minsu Kim 외

Adversarial attacks on face recognition systems (FRSs) pose serious security and privacy threats, especially when these systems are used for identity verification. In this paper, we propose a novel method for generating …

AttributeFace GenerationFace Recognition

Attributes Shape the Embedding Space of Face Recognition Models

2025-07-15 · Pierrick Leroy, Antonio Mastropietro, Marco Nurisso, Francesco Vaccarino

Face Recognition (FR) tasks have made significant progress with the advent of Deep Neural Networks, particularly through margin-based triplet losses that embed facial images into high-dimensional feature spaces. During t…

AttributeFace RecognitionTriplet

COLIBRI Fuzzy Model: Color Linguistic-Based Representation and Interpretation

2025-07-15 · Pakizar Shamoi, Nuray Toganas, Muragul Muratbekova, Elnara Kadyrgali 외

Colors are omnipresent in today's world and play a vital role in how humans perceive and interact with their surroundings. However, it is challenging for computers to imitate human color perception. This paper introduces…

AttributeMarketing

Ref-Long: Benchmarking the Long-context Referencing Capability of Long-context Language Models

2025-07-13 · Junjie Wu, Gefei Gu, Yanan Zheng, Dit-yan Yeung 외

Long-context language models (LCLMs) have exhibited impressive capabilities in long-context understanding tasks. Among these, long-context referencing -- a crucial task that requires LCLMs to attribute items of interest …

AttributeBenchmarkingLong-Context Understanding

Model Parallelism With Subnetwork Data Parallelism

2025-07-11 · Vaibhav Singh, Zafir Khalid, Edouard Oyallon, Eugene Belilovsky

Distributed pre-training of large models at scale often imposes heavy memory demands on individual nodes and incurs significant intra-node communication costs. We propose a novel alternative approach that reduces the mem…

AttributeFederated Learningmodel

Bradley-Terry and Multi-Objective Reward Modeling Are Complementary

2025-07-10 · Zhiwei Zhang, Hui Liu, Xiaomin Li, Zhenwei Dai 외

Reward models trained on human preference data have demonstrated strong effectiveness in aligning Large Language Models (LLMs) with human intent under the framework of Reinforcement Learning from Human Feedback (RLHF). H…

Attributeregression

Evaluating Attribute Confusion in Fashion Text-to-Image Generation

2025-07-09 · Ziyue Liu, Federico Girella, Yiming Wang, Davide Talon

Despite the rapid advances in Text-to-Image (T2I) generation models, their evaluation remains challenging in domains like fashion, involving complex compositional generation. Recent automated T2I evaluation methods lever…

Attributecross-modal alignmentImage GenerationQuestion Answering+5

LIRA: Inferring Segmentation in Large Multi-modal Models with Local Interleaved Region Assistance

2025-07-08 · Zhang Li, Biao Yang, Qiang Liu, Zhiyin Ma 외

While large multi-modal models (LMMs) demonstrate promising capabilities in segmentation and comprehension, they still struggle with two limitations: inaccurate segmentation and hallucinated comprehension. These challeng…

AttributeSegmentationSemantic Segmentation

Emergent Semantics Beyond Token Embeddings: Transformer LMs with Frozen Visual Unicode Representations

2025-07-07 · A. Bochkov

Understanding the locus of semantic representation in large language models (LLMs) is crucial for interpretability and architectural innovation. The dominant paradigm posits that trainable input embeddings serve as found…

AttributeMMLU

An analysis of vision-language models for fabric retrieval

2025-07-07 · Francesco Giuliari, Asif Khan Pattan, Mohamed Lamine Mekhalfi, Fabio Poiesi

Effective cross-modal retrieval is essential for applications like information retrieval and recommendation systems, particularly in specialized domains such as manufacturing, where product information often consists of …

AttributeCross-Modal RetrievalImage RetrievalInformation Retrieval+3

MambaFusion: Height-Fidelity Dense Global Fusion for Multi-modal 3D Object Detection

2025-07-06 · Hanshi Wang, Jin Gao, Weiming Hu, Zhipeng Zhang

We present the first work demonstrating that a pure Mamba block can achieve efficient Dense Global Fusion, meanwhile guaranteeing top performance for camera-LiDAR multi-modal 3D object detection. Our motivation stems fro…

3D Object DetectionAttributeLong-range modelingMamba+3

Helping CLIP See Both the Forest and the Trees: A Decomposition and Description Approach

2025-07-04 · Leyan Xue, Zongbo Han, Guangyu Wang, QinGhua Hu 외

Vision-Language Models (VLMs) like CLIP achieve cross-modal semantic alignment through contrastive learning, exhibiting robust zero-shot generalization. Traditional prompt engineering, however, predominantly relies on co…

AttributeContrastive LearningPrompt EngineeringTest-time Adaptation+1

IndianBailJudgments-1200: A Multi-Attribute Dataset for Legal NLP on Indian Bail Orders

2025-07-03 · Sneha Deshmukh, Prathmesh Kamble

Legal NLP remains underdeveloped in regions like India due to the scarcity of structured datasets. We introduce IndianBailJudgments-1200, a new benchmark dataset comprising 1200 Indian court judgments on bail decisions, …

AttributeFairnessJurisprudenceLegal Reasoning

The Trilemma of Truth in Large Language Models

2025-06-30 · Germans Savcisens, Tina Eliassi-Rad

We often attribute human characteristics to large language models (LLMs) and claim that they "know" certain things. LLMs have an internal probabilistic knowledge that represents information retained during training. How …

AttributeConformal PredictionKnowledge DistillationMultiple Instance Learning+1

FOCUS: Fine-grained Optimization with Semantic Guided Understanding for Pedestrian Attributes Recognition

2025-06-28 · Hongyan An, Kuan Zhu, Xin He, Haiyun Guo 외

Pedestrian attribute recognition (PAR) is a fundamental perception task in intelligent transportation and security. To tackle this fine-grained task, most existing methods focus on extracting regional features to enrich …

AttributeContrastive LearningPedestrian Attribute Recognition

SAC: A Framework for Measuring and Inducing Personality Traits in LLMs with Dynamic Intensity Control

2025-06-26 · Adithya Chittem, Aishna Shrivastava, Sai Tarun Pendela, Jagat Sesh Challa 외

Large language models (LLMs) have gained significant traction across a wide range of fields in recent years. There is also a growing expectation for them to display human-like personalities during interactions. To meet t…

Attribute

Text2Cypher Across Languages: Evaluating Foundational Models Beyond English

2025-06-26 · Makbule Gulcin Ozsoy, William Tai

Recent advances in large language models have enabled natural language interfaces that translate user questions into database queries, such as Text2SQL, Text2SPARQL, and Text2Cypher. While these interfaces enhance databa…

AttributeText2Sparql

Style-Aligned Image Composition for Robust Detection of Abnormal Cells in Cytopathology

2025-06-26 · Qiuyi Qi, Xin Li, Ming Kong, Zikang Xu 외

Challenges such as the lack of high-quality annotations, long-tailed data distributions, and inconsistent staining styles pose significant obstacles to training neural networks to detect abnormal cells in cytopathology r…

AttributeCell Detection

IPFormer-VideoLLM: Enhancing Multi-modal Video Understanding for Multi-shot Scenes

2025-06-26 · Yujia Liang, Jile Jiao, Zhicheng Wang, Xuetao Feng 외

Video Large Language Models (VideoLLMs) have demonstrated remarkable understanding capabilities, but are found struggling to tackle multi-shot scenarios,e.g., video clips with varying camera angles or scene changes. This…

AttributeQuestion AnsweringVideo Understanding
1–20 / 5,387 다음 →