paper-with-me

홈 › Papers

How Small Can You Go? LoRA Fine-Tuning 270M-8B Models for Merchant Information Extraction in Financial Transactions

2026-06-06 · Donghao Huang, Tomas Drietomsky, Benjamin Barrett, Zhaoxia Wang arxiv

Merchant information extraction turns noisy financial transaction descriptors into structured fields at production scale. Our deployed LoRA-fine-tuned LLaMA~3.1-8B reaches 96.95\% F1, but its memory and throughput motivate smaller replacements. We evaluate 23 retained fine-tuning runs plus a separately trained production reference, spanning Gemma~3 (270M--4B), Qwen~3.5 (0.8B--4B), Aya~3.35B, and LLaMA~3.1-8B across LoRA ranks, prompts, training templates, and serving environments. A rank-8 LLaMA fine-tune reaches 96.75\% F1, only 0.20 points below the rank-32 production reference. Qwen~3.5~4B with JSON-Only prompting reaches 96.60\% F1 and strict record-level exact match of 91.67\%, with a $3.8\times$ lower inverse-throughput time estimate than the rank-8 8B model. Qwen~3.5~0.8B reaches 94.75\% F1, and Qwen Think and Nothink templates differ by less than 0.004 F1. Across 14 Databricks endpoints, mean F1 change from local evaluation is $-0.0081$; Aya is the only family with a 2.7--5.1 point decline. These results show that compact fine-tuned models can preserve most extraction accuracy, but model selection must account for prompt choice, throughput, and serving-stack behavior.

📄 PDF Abstract BibTeX arXiv:2606.08051

Code (0)

등록된 구현이 없습니다.

Tasks

Information Extraction

Similar Papers 제목 키워드 기반

Leveraging Large Language Models for Active Merchant Non-player Characters

2024-12-15 · Byungjun Kim, Minju Kim, Dayeon Seo, Bugeun Kim

We highlight two significant issues leading to the passivity of current merchant non-player characters (NPCs): pricing and communication. While immersive interactions have been a focus, negotiations between merchant NPCs…

Knowledge Distillation

Instruction-Based Fine-tuning of Open-Source LLMs for Predicting Customer Purchase Behaviors

2025-01-28 · Halil Ibrahim Ergul, Selim Balcisoy, Burcin Bozkaya

In this study, the performance of various predictive models, including probabilistic baseline, CNN, LSTM, and finetuned LLMs, in forecasting merchant categories from financial transaction data have been evaluated. Utiliz…

Marketing

Merchant Category Identification Using Credit Card Transactions

2020-11-05 · Chin-Chia Michael Yeh, Zhongfang Zhuang, Yan Zheng, Liang Wang 외

Digital payment volume has proliferated in recent years with the rapid growth of small businesses and online shops. When processing these digital transactions, recognizing each merchant's real identity (i.e., business ty…

Time SeriesTime Series AnalysisTime Series Classification

SeqLLM: Augmenting LLMs with Behavioral-Sequence Modeling for High-Stakes Decisions at WeChat Pay

2026-08-04 · Guilin Li, Jiaxing Zhang, Matthias Hwai Yong Tan, Bo Wang 외 arxiv

Merchant risk control at large payment platforms screens tens of millions of merchants daily, where false positives harm legitimate merchants and false negatives leave harmful activity undetected. The hardest cases requi…

MERIT: A Merchant Incentive Ranking Model for Hotel Search & Ranking

2025-06-10 · Shigang Quan, Hailong Tan, Shui Liu, Zhenzhe Zheng 외

Online Travel Platforms (OTPs) have been working on improving their hotel Search & Ranking (S&R) systems that facilitate efficient matching between consumers and hotels. Existing OTPs focus almost exclusively on improvin…

Relation