paper-with-me

홈 › Papers

Baikal: Structured Search for Deep Research over Data Lakes

2026-07-30 · Dhruv Agarwal, Rishitha Guttapalle Mohan, Aarti Kumari, Ashi Sinha, Athulya Anil, Kavitha Srinivas, Horst Samulowitz, Andrew McCallum arxiv

Deep research over data lakes requires an LLM agent to investigate evidence across thousands of heterogeneous tables and passages to synthesize a report. Existing methods perform iterative retrieval and generation, letting accumulated context determine what to investigate next, which can overexploit locally promising evidence and fail to cover distinct semantic regions under a fixed budget. To address this, we cast deep research over data lakes as a budgeted search problem and present Baikal - a framework that clusters heterogeneous evidence into semantic regions, then searches over them adaptively to balance exploration and exploitation. Within each selected region, Baikal generates and investigates region-grounded subquestions, using finding quality as rewards to update region-level value estimates and guide search under policies ranging from random and LLM-guided selection to Bayesian $ε$-greedy and UCB. We evaluate Baikal on 15 queries each over HybridQA and TAT-QA data lakes containing 10,993 and 2,757 tables, respectively, together with 227K Wikipedia passages and 13K financial report passages. We assess research quality with a new rubric covering groundedness, relevance, diversity, and utility, and use GPT-5-mini to score Baikal and strong baselines, including DeepSearcher and an OpenCode research agent with retrieval and clustering variants. Across both data lakes, Baikal performs strongly under several region-selection policies; its best configuration improves report scores over the strongest baselines by 28% on HybridQA and 36% on TAT-QA. Our analyses attribute these gains to organizing and exploring semantic evidence regions, which improves groundedness and diversity and yields more useful findings under the same subquestion budget. These results demonstrate the value of structured semantic exploration for systematic research and discovery over heterogeneous data lakes.

📄 PDF Abstract BibTeX arXiv:2607.27726

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Rejecting noise in Baikal-GVD data with neural networks

2022-10-10 · I. Kharuk, G. Rubtsov, G. Safronov

Baikal-GVD is a large ($\sim$1 km$^3$) underwater neutrino telescope installed in the fresh waters of Lake Baikal. The deep lake water environment is pervaded by background light, which is detectable by Baikal-GVD's phot…

From raw data to neutrino candidates: a neural-network pipeline for Baikal-GVD

2026-05-11 · A. Matseiko, G. Plotnikov, I. Kharuk arxiv

We present a neural-network-based data processing pipeline for Baikal-GVD, designed to improve event reconstruction quality and accelerate neutrino candidates selection. The pipeline comprises three stages: fast suppress…

Domain Adaptation

Harnessing the power of Topological Data Analysis to detect change points in time series

2019-10-28 · Umar Islambekov, Monisha Yuvaraj, Yulia R. Gel

We introduce a novel geometry-oriented methodology, based on the emerging tools of topological data analysis, into the change point detection framework. The key rationale is that change points are likely to be associated…

Change Point DetectionTime SeriesTime Series AnalysisTopological Data Analysis

DataSTORM: Deep Research on Large-Scale Databases using Exploratory Data Analysis and Data Storytelling

2026-04-07 · Shicheng Liu, Yucheng Jiang, Sajid Farook, Camila Nicollier Sanchez 외 arxiv

Deep research with Large Language Model (LLM) agents is emerging as a powerful paradigm for multi-step information discovery, synthesis, and analysis. However, existing approaches primarily focus on unstructured web data…

Personality Structured Interview for Large Language Model Simulation in Personality Research

2025-02-17 · Pengda Wang, Huiqi Zou, Hanjie Chen, Tianjun Sun 외

Although psychometrics researchers have recently explored the use of large language models (LLMs) as proxies for human participants, LLMs often fail to generate heterogeneous data with human-like diversity, which diminis…

Language ModelingLanguage ModellingLarge Language Model