paper-with-me

Papers Language Modeling

“Language Modeling” 태그가 달린 논문 14,182편 · 필터 해제

Visual-Language Model Knowledge Distillation Method for Image Quality Assessment

2025-07-21 · Yongkang Hou, Jiarun Song

Image Quality Assessment (IQA) is a core task in computer vision. Multimodal methods based on vision-language models, such as CLIP, have demonstrated exceptional generalization capabilities in IQA tasks. To address the i…

Image Quality AssessmentKnowledge DistillationLanguage ModelingLanguage Modelling

Making Language Model a Hierarchical Classifier and Generator

2025-07-17 · Yihong Wang, Zhonglin Jiang, Ningyuan Xi, Yue Zhao 외

Decoder-only language models, such as GPT and LLaMA, generally decode on the last layer. Motivated by human's hierarchical thinking capability, we propose that a hierarchical decoder architecture could be built with diff…

DecoderLanguage ModelingLanguage Modellingmodel+3

VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning

2025-07-17 · Senqiao Yang, Junyi Li, Xin Lai, Bei Yu 외

Recent advancements in vision-language models (VLMs) have improved performance by increasing the number of visual tokens, which are often significantly longer than text tokens. However, we observe that most real-world sc…

Language ModelingLanguage ModellingOptical Character Recognition (OCR)reinforcement-learning+2

The Generative Energy Arena (GEA): Incorporating Energy Awareness in Large Language Model (LLM) Human Evaluations

2025-07-17 · Carlos Arriaga, Gonzalo Martínez, Eneko Sendin, Javier Conde 외

The evaluation of large language models is a complex task, in which several approaches have been proposed. The most common is the use of automated benchmarks in which LLMs have to answer multiple-choice questions of diff…

Language ModelingLanguage ModellingLarge Language ModelMultiple-choice

Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities

2025-07-17 · Hao Sun, Mihaela van der Schaar

In the era of Large Language Models (LLMs), alignment has emerged as a fundamental yet challenging problem in the pursuit of more reliable, controllable, and capable machine intelligence. The recent success of reasoning …

Language ModelingLanguage ModellingLarge Language ModelReinforcement Learning (RL)

Assay2Mol: large language model-based drug design using BioAssay context

2025-07-16 · Yifan Deng, Spencer S. Ericksen, Anthony Gitter

Scientific databases aggregate vast amounts of quantitative data alongside descriptive text. In biochemistry, molecule screening assays evaluate the functional responses of candidate molecules against disease targets. Un…

DescriptiveDrug DesignDrug DiscoveryIn-Context Learning+3

Describe Anything Model for Visual Question Answering on Text-rich Images

2025-07-16 · Yen-Linh Vu, Dinh-Thang Duong, Truong-Binh Duong, Anh-Khoi Nguyen 외

Recent progress has been made in region-aware vision-language modeling, particularly with the emergence of the Describe Anything Model (DAM). DAM is capable of generating detailed descriptions of any specific image areas…

DescriptiveLanguage ModelingLanguage ModellingQuestion Answering+2

InstructFLIP: Exploring Unified Vision-Language Model for Face Anti-spoofing

2025-07-16 · Kun-Hsiang Lin, Yu-Wen Tseng, Kang-Yang Huang, Jhih-Ciang Wu 외

Face anti-spoofing (FAS) aims to construct a robust system that can withstand diverse attacks. While recent efforts have concentrated mainly on cross-domain generalization, two significant challenges persist: limited sem…

Domain GeneralizationFace Anti-SpoofingLanguage ModelingLanguage Modelling

Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility

2025-07-16 · Michael A. Lepori, Jennifer Hu, Ishita Dasgupta, Roma Patel 외

Language models (LMs) are used for a diverse range of tasks, from question answering to writing fantastical stories. In order to reliably accomplish these tasks, LMs must be able to discern the modal category of a senten…

Language ModelingLanguage ModellingQuestion Answering

KptLLM++: Towards Generic Keypoint Comprehension with Large Language Model

2025-07-15 · Jie Yang, Wang Zeng, Sheng Jin, Lumin Xu 외

The emergence of Multimodal Large Language Models (MLLMs) has revolutionized image understanding by bridging textual and visual modalities. However, these models often struggle with capturing fine-grained semantic inform…

Keypoint DetectionLanguage ModelingLanguage ModellingLarge Language Model+1

Tactical Decision for Multi-UGV Confrontation with a Vision-Language Model-Based Commander

2025-07-15 · Li Wang, Qizhen Wu, Lei Chen

In multiple unmanned ground vehicle confrontations, autonomously evolving multi-agent tactical decisions from situational awareness remain a significant challenge. Traditional handcraft rule-based methods become vulnerab…

Language ModelingLanguage ModellingLarge Language Modelreinforcement-learning+2

Mixture of Experts in Large Language Models

2025-07-15 · Danyang Zhang, Junhao Song, Ziqian Bi, Yingfang Yuan 외

This paper presents a comprehensive review of the Mixture-of-Experts (MoE) architecture in large language models, highlighting its ability to significantly enhance model performance while maintaining minimal computationa…

DiversityLanguage ModelingLanguage ModellingLarge Language Model+2

LRCTI: A Large Language Model-Based Framework for Multi-Step Evidence Retrieval and Reasoning in Cyber Threat Intelligence Credibility Verification

2025-07-15 · Fengxiao Tang, Huan Li, Ming Zhao, Zongzong Wu 외

Verifying the credibility of Cyber Threat Intelligence (CTI) is essential for reliable cybersecurity defense. However, traditional approaches typically treat this task as a static classification problem, relying on handc…

Language ModelingLanguage ModellingLarge Language ModelNatural Language Inference+1

LiLM-RDB-SFC: Lightweight Language Model with Relational Database-Guided DRL for Optimized SFC Provisioning

2025-07-15 · Parisa Fard Moshiri, Xinyu Zhu, Poonam Lohan, Burak Kantarci 외

Effective management of Service Function Chains (SFCs) and optimal Virtual Network Function (VNF) placement are critical challenges in modern Software-Defined Networking (SDN) and Network Function Virtualization (NFV) en…

Deep Reinforcement LearningLanguage ModelingLanguage ModellingLarge Language Model

KisMATH: Do LLMs Have Knowledge of Implicit Structures in Mathematical Reasoning?

2025-07-15 · Soumadeep Saha, Akshay Chaturvedi, Saptarshi Saha, Utpal Garain 외

Chain-of-thought traces have been shown to improve performance of large language models in a plethora of reasoning tasks, yet there is no consensus on the mechanism through which this performance boost is achieved. To sh…

GSM8KLanguage ModelingLanguage ModellingMathematical Reasoning

MLAR: Multi-layer Large Language Model-based Robotic Process Automation Applicant Tracking

2025-07-14 · Mohamed T. Younes, Omar Walid, Mai Hassan, Ali Hamdi

This paper introduces an innovative Applicant Tracking System (ATS) enhanced by a novel Robotic process automation (RPA) framework or as further referred to as MLAR. Traditional recruitment processes often encounter bott…

BenchmarkingLanguage ModelingLanguage ModellingLarge Language Model

Iceberg: Enhancing HLS Modeling with Synthetic Data

2025-07-14 · Zijian Ding, Tung Nguyen, Weikai Li, Aditya Grover 외

Deep learning-based prediction models for High-Level Synthesis (HLS) of hardware designs often struggle to generalize. In this paper, we study how to close the generalizability gap of these models through pretraining on …

Data AugmentationHigh-Level SynthesisLanguage ModelingLanguage Modelling+2

Kodezi Chronos: A Debugging-First Language Model for Repository-Scale, Memory-Driven Code Understanding

2025-07-14 · Ishraq Khan, Assad Chowdary, Sharoz Haseeb, Urvish Patel

Large Language Models (LLMs) have advanced code generation and software automation, but are fundamentally constrained by limited inference-time context and lack of explicit code structure reasoning. We introduce Kodezi C…

Code GenerationLanguage ModelingLanguage ModellingRetrieval

ByDeWay: Boost Your multimodal LLM with DEpth prompting in a Training-Free Way

2025-07-11 · Rajarshi Roy, Devleena Das, Ankesh Banerjee, Arjya Bhattacharjee 외

We introduce ByDeWay, a training-free framework designed to enhance the performance of Multimodal Large Language Models (MLLMs). ByDeWay uses a novel prompting strategy called Layered-Depth-Based Prompting (LDP), which i…

Depth EstimationHallucinationLanguage ModelingLanguage Modelling+2

Lizard: An Efficient Linearization Framework for Large Language Models

2025-07-11 · Chien Van Nguyen, Ruiyi Zhang, Hanieh Deilamsalehy, Puneet Mathur 외

We propose Lizard, a linearization framework that transforms pretrained Transformer-based Large Language Models (LLMs) into flexible, subquadratic architectures for infinite-context generation. Transformer-based LLMs fac…

Language ModelingLanguage ModellingMMLU
1–20 / 14,182 다음 →