paper-with-me

홈 › Papers

EnterpriseLab: A Full-Stack Platform for developing and deploying agents in Enterprises

2026-03-23 · Ankush Agarwal, Harsh Vishwakarma, Suraj Nagaje, Chaitanya Devaguptapu arxiv

Deploying AI agents in enterprise environments requires balancing capability with data sovereignty and cost constraints. While small language models offer privacy-preserving alternatives to frontier models, their specialization is hindered by fragmented development pipelines that separate tool integration, data generation, and training. We introduce EnterpriseLab, a full-stack platform that unifies these stages into a closed-loop framework. EnterpriseLab provides (1) a modular environment exposing enterprise applications via Model Context Protocol, enabling seamless integration of proprietary and open-source tools; (2) automated trajectory synthesis that programmatically generates training data from environment schemas; and (3) integrated training pipelines with continuous evaluation. We validate the platform through EnterpriseArena, an instantiation with 15 applications and 140+ tools across IT, HR, sales, and engineering domains. Our results demonstrate that 8B-parameter models trained within EnterpriseLab match GPT-4o's performance on complex enterprise workflows while reducing inference costs by 8-10x, and remain robust across diverse enterprise benchmarks, including EnterpriseBench (+10%) and CRMArena (+10%). EnterpriseLab provides enterprises a practical path to deploying capable, privacy-preserving agents without compromising operational capability.

📄 PDF Abstract BibTeX arXiv:2603.21630

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MIND-Stack: Modular, Interpretable, End-to-End Differentiability for Autonomous Navigation

2025-05-27 · Felix Jahncke, Johannes Betz

Developing robust, efficient navigation algorithms is challenging. Rule-based methods offer interpretability and modularity but struggle with learning from large datasets, while end-to-end neural networks excel in learni…

Autonomous NavigationState Estimation

Edge Impulse: An MLOps Platform for Tiny Machine Learning

2022-11-02 · Shawn Hymel, Colby Banbury, Daniel Situnayake, Alex Elium 외

Edge Impulse is a cloud-based machine learning operations (MLOps) platform for developing embedded and edge ML (TinyML) systems that can be deployed to a wide range of hardware targets. Current TinyML workflows are plagu…

The Case for a Wholistic Serverless Programming Paradigm and Full Stack Automation for AI and Beyond -- The Philosophy of Jaseci and Jac

2022-06-16 · Jason Mars

In this work, the case is made for a wholistic top-down re-envisioning of the system stack from the programming language level down through the system architecture to bridge this complexity gap. The key goal of our desig…

Philosophy

Deploying Deep Reinforcement Learning Systems: A Taxonomy of Challenges

2023-08-23 · Ahmed Haj Yahmed, Altaf Allah Abbassi, Amin Nikanjam, Heng Li 외

Deep reinforcement learning (DRL), leveraging Deep Learning (DL) in reinforcement learning, has shown significant potential in achieving human-level autonomy in a wide range of domains, including robotics, computer visio…

Deep Reinforcement Learningreinforcement-learningReinforcement Learning

Industrial cuVSLAM Benchmark & Integration

2026-03-17 · Charbel Abi Hana, Kameel Amareen, Mohamad Mostafa, Dmitry Slepichev 외 arxiv

This work presents a comprehensive benchmark evaluation of visual odometry (VO) and visual SLAM (VSLAM) systems for mobile robot navigation in real-world logistical environments. We compare multiple visual odometry appro…

Robot NavigationVisual Odometry