paper-with-me

홈 › Papers

Orion-RAG: Path-Aligned Hybrid Retrieval for Graphless Data

2026-01-08 · Zhen Chen, Weihao Xie, Peilin Chen, Shiqi Wang, Jianping Wang arxiv

Retrieval-Augmented Generation (RAG) has proven effective for knowledge synthesis, yet it encounters significant challenges in practical scenarios where data is inherently discrete and fragmented. In most environments, information is distributed across isolated files like reports and logs that lack explicit links. Standard search engines process files independently, ignoring the connections between them. Furthermore, manually building Knowledge Graphs is impractical for such vast data. To bridge this gap, we present Orion-RAG. Our core insight is simple yet effective: we do not need heavy algorithms to organize this data. Instead, we use a low-complexity strategy to extract lightweight paths that naturally link related concepts. We demonstrate that this streamlined approach suffices to transform fragmented documents into semi-structured data, enabling the system to link information across different files effectively. Extensive experiments demonstrate that Orion-RAG consistently outperforms mainstream frameworks across diverse domains, supporting real-time updates and explicit Human-in-the-Loop verification with high cost-efficiency. Experiments on FinanceBench demonstrate superior precision with a 25.2% relative improvement over strong baselines.

📄 PDF Abstract BibTeX arXiv:2601.04764

Code (0)

등록된 구현이 없습니다.

Tasks

Knowledge Graphs

Similar Papers 제목 키워드 기반

Federated Graph Learning with Graphless Clients

2024-11-13 · Xingbo Fu, Song Wang, Yushun Dong, Binchi Zhang 외

Federated Graph Learning (FGL) is tasked with training machine learning models, such as Graph Neural Networks (GNNs), for multiple clients, each with its own graph data. Existing methods usually assume that each client h…

Graph LearningKnowledge DistillationTransfer Learning

ORION: Teaching Language Models to Reason Efficiently in the Language of Thought

2025-11-28 · Kumar Tanmay, Kriti Aggarwal, Paul Pu Liang, Subhabrata Mukherjee arxiv

Large Reasoning Models (LRMs) achieve strong performance in mathematics, code generation, and task planning, but their reliance on long chains of verbose "thinking" tokens leads to high latency, redundancy, and incoheren…

Reinforcement LearningCode Generation

ORION: Option-Regularized Deep Reinforcement Learning for Cooperative Multi-Agent Online Navigation

2026-01-03 · Shizhe Zhang, Jingsong Liang, Zhitao Zhou, Shuhan Ye 외 arxiv

Existing methods for multi-agent navigation typically assume fully known environments, offering limited support for partially known scenarios with outdated or imperfect prior maps, such as warehouses or factory floors. T…

Reinforcement Learning

Orion-14B: Open-source Multilingual Large Language Models

2024-01-20 · Du Chen, Yi Huang, Xiaopu Li, Yongqiang Li 외

In this study, we introduce Orion-14B, a collection of multilingual large language models with 14 billion parameters. We utilize a data scheduling approach to train a foundational model on a diverse corpus of 2.5 trillio…

Scheduling

Orion: A Unified Visual Agent for Multimodal Perception, Advanced Visual Reasoning and Execution

2025-11-18 · N Dinesh Reddy, Dylan Snyder, Lona Kiragu, Mirajul Mohin 외 arxiv

We introduce Orion, a visual agent that integrates vision-based reasoning with tool-augmented execution to achieve powerful, precise, multi-step visual intelligence across images, video, and documents. Unlike traditional…

Panoptic SegmentationObject DetectionVisual Reasoning