paper-with-me

홈 › Papers

Vibe-FDTR: An agent-oriented framework for reproducible frequency-domain thermoreflectance data analysis

2026-07-30 · Fuwei Yang, Weiheng Li, Bai Song arxiv

Frequency-domain thermoreflectance (FDTR) is a laser pump-probe technique widely used to measure thermal properties at the micro- and nanoscale; however, it relies on a complex data analysis procedure that demands substantial domain expertise and is susceptible to subtle human errors. Here, we present Vibe-FDTR, an agent-oriented framework that enables large language model (LLM) agents to perform reliable and reproducible FDTR analyses directly from natural language requests. This framework couples a configuration-driven FDTR code package, which enforces physical and parametric consistency, with procedural agent skills that translate user intentions into organized and verifiable analysis steps. We evaluate Vibe-FDTR using a controlled benchmark with two levels: synthetic single-step tasks and real-data multi-step tasks based on measurements of gold-coated graphite samples. Across the two levels, agents using Vibe-FDTR achieve success rates of 100% and 98.9%, respectively. In sharp contrast, ablating skills (Code-agent) reduces performance to 91.4% and 36.7%, which drops further to 38.6% and 0% when the domain package is also omitted (Agent-only). Beyond success rate, Vibe-FDTR also reduces computational cost by 87.7% relative to the Code-agent variant and cuts execution time by more than 60%. Finally, an optional expert mode supports experimental planning via autonomous sensitivity and uncertainty evaluations, and formulates physically grounded recommendations for underspecified tasks. These results demonstrate that encapsulating domain code and expert knowledge into agent skills offers a promising route toward low-barrier, autonomous, and trustworthy thermal metrology.

📄 PDF Abstract BibTeX arXiv:2607.28200

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

VibeSearchBench: Benchmarking Long-horizon Proactive Search in the Wild

2026-05-27 · Xiaohongshu Inc arxiv

LLM-based agents score well on search benchmarks, yet real users consistently find results unsatisfying, revealing a persistent evaluation-experience gap. We attribute this gap to existing benchmarks' reliance on over-sp…

Multi-object Tracking: Decoupling Features to Solve the Contradictory Dilemma of Feature Requirements

2023-02-27 · IEEE Transactions on Circuits and Systems for Video Technology 2023 2 · Yan Jin; Fang Gao; Jun Yu; Jiabao Wang; Feng Shuang

Multi-object tracking achieves the acquisition of target location information and identity information through two subtasks, detection and re-identification (ReID). The existing commonly used one-shot framework has speed…

Multi-Object TrackingObject Tracking

VibeWorlding: Can Multimodal Agents Construct 3D Open Worlds End-to-End?

2026-08-15 · Yansong Ning, Jingwen Ye, Zhongkai Wu, Yang Sun 외 hf

Constructing an interactive 3D open world from a user query is important. However, existing methods are primarily evaluated on idealized, simple queries, making it difficult to systematically analyze and compare how mult…

VibeTensor: System Software for Deep Learning, Fully Generated by AI Agents

2026-01-21 · Bing Xu, Terry Chen, Fengzhe Zhou, Tianqi Chen 외 arxiv

VIBETENSOR is an open-source research system software stack for deep learning, generated by LLM-powered coding agents under high-level human guidance. In this paper, "fully generated" refers to code provenance: implement…

SecureVibeBench: Benchmarking Secure Vibe Coding of AI Agents via Reconstructing Vulnerability-Introducing Scenarios

2025-09-26 · Junkai Chen, Huihui Huang, Yunbo Lyu, Junwen An 외 arxiv

Large language model-powered code agents are rapidly transforming software engineering, yet the security risks of their generated code have become a critical concern. Existing benchmarks have provided valuable insights, …