paper-with-me

Papers

Robust Dual View Deep Agent

2018-04-13 · Ibrahim M. Sobh, Nevin M. Darwish

Motivated by recent advance of machine learning using Deep Reinforcement Learning this paper proposes a modified architecture that produces more robust agents and speeds up the training process. Our architecture is based on Asynchronous Advantage Actor-Critic (A3C) algorithm where the total input dimensionality is halved by dividing the input into two independent streams. We use ViZDoom, 3D world software that is based on the classical first person shooter video game, Doom, as a test case. The experiments show that in comparison to single input agents, the proposed architecture succeeds to have the same playing performance and shows more robust behavior, achieving significant reduction in the number of training parameters of almost 30%.

📄 PDF Abstract BibTeX arXiv:1804.05120

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

DualView: Preventing Indirect Prompt Injection in Personal AI Agents

2026-07-04 · Juhee Kim, Woohyuk Choi, Taehyun Kang, Youngmin Kim 외 arxiv

Personal AI agents that run on the user's local machine, such as OpenClaw, automate daily tasks including web search, email, and file management. Their access to computer resources, including the network, file system, an…

From Human-Centric to Agentic Code Review: The Impact of Different Generations of Generative AI Technology on Review Quality

2026-07-14 · Suzhen Zhong, Shayan Noei, Bram Adams, Ying Zou hf

Code review helps maintain software quality before code integration, but it also imposes a substantial workload on human reviewers. As generative artificial intelligence becomes part of software development, code review …

Streaming Multi-Agent Autoregressive Diffusion Model with World State Registers

2026-07-23 · Sicheng Mo, Yuheng Li, Ziyang Leng, Krishna Kumar Singh 외 hf

Multi-agent interactive world models should not only generate consistent observations, but also maintain world states that persist across agents and evolve across views. Existing autoregressive video diffusion pipelines …

Video Generation

MimiTalk: Revolutionizing Qualitative Research with Dual-Agent AI

2025-09-27 · Fengming Liu, Shubin Yu arxiv

We present MimiTalk, a dual-agent constitutional AI framework designed for scalable and ethical conversational data collection in social science research. The framework integrates a supervisor model for strategic oversig…

Question Generation

Duality and Stability in Complex Multiagent State-Dependent Network Dynamics

2020-07-19

Despite significant progress on stability analysis of conventional multiagent networked systems with weakly coupled state-network dynamics, most of the existing results have shortcomings in addressing multiagent systems …