paper-with-me

홈 › Papers

Domain-independent generation and classification of behavior traces

2020-11-03 · Daniel Borrajo, Manuela Veloso

Financial institutions mostly deal with people. Therefore, characterizing different kinds of human behavior can greatly help institutions for improving their relation with customers and with regulatory offices. In many of such interactions, humans have some internal goals, and execute some actions within the financial system that lead them to achieve their goals. In this paper, we tackle these tasks as a behavior-traces classification task. An observer agent tries to learn characterizing other agents by observing their behavior when taking actions in a given environment. The other agents can be of several types and the goal of the observer is to identify the type of the other agent given a trace of observations. We present CABBOT, a learning technique that allows the agent to perform on-line classification of the type of planning agent whose behavior is observing. In this work, the observer agent has partial and noisy observability of the environment (state and actions of the other agents). In order to evaluate the performance of the learning technique, we have generated a domain-independent goal-based simulator of agents. We present experiments in several (both financial and non-financial) domains with promising results.

📄 PDF Abstract BibTeX arXiv:2011.02918

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationGeneral Classification

Similar Papers 제목 키워드 기반

Learning Correct Behavior from Examples: Validating Sequential Execution in Autonomous Agents

2026-05-04 · Reshabh K Sharma, Gaurav Mittal, Yu Hu arxiv

As autonomous agents become increasingly sophisticated, validating their sequential behavior presents a significant challenge. Traditional testing approaches require manual specification, exact sequence matching, or thou…

Code Generation

ConPress: Learning Efficient Reasoning from Multi-Question Contextual Pressure

2026-02-01 · Jie Deng, Shining Liang, Jun Li, Hongzhi Li 외 arxiv

Large reasoning models (LRMs) typically solve reasoning-intensive tasks by generating long chain-of-thought (CoT) traces, leading to substantial inference overhead. We identify a reproducible inference-time phenomenon, t…

Reinforcement Learning

Modeling Student Learning with 3.8 Million Program Traces

2025-10-06 · Alexis Ross, Megha Srivastava, Jeremiah Blanchard, Jacob Andreas arxiv

As programmers write code, they often edit and retry multiple times, creating rich "interaction traces" that reveal how they approach coding tasks and provide clues about their level of skill development. For novice prog…

Code Generation

Self-Supervised Noise2Noise-Enhanced Denoising for Continuous-Scan Air-Plasma THz Spectroscopy

2026-08-17 · Adam Umra, Oways Alsoloh, Oliver Nagy, Aydin Sezgin 외 arxiv

Terahertz time-domain spectroscopy (THz-TDS) based on air-plasma generation and balanced air-biased coherent detection offers gap-free broadband coverage, but individual continuous-scan traces are strongly affected by pu…

Self-Supervised Learning

RedAct: Redacting Agent Capability Traces for Procedural Skill Protection

2026-06-09 · Shuwen Xu, Zhitao He, Yi R. Fung arxiv

Users rely on execution traces to observe agent behavior, diagnose failures, and ensure accountability. These traces contain rich procedural detail, including tool invocations, intermediate decisions, and error-recovery …