paper-with-me

Papers

Guided Reasoning: A Non-Technical Introduction

2024-08-29 · Gregor Betz

We introduce the concept and a default implementation of Guided Reasoning. A multi-agent system is a Guided Reasoning system iff one agent (the guide) primarily interacts with other agents in order to improve reasoning quality. We describe Logikon's default implementation of Guided Reasoning in non-technical terms. This is a living document we'll gradually enrich with more detailed information and examples. Code: https://github.com/logikon-ai/logikon

📄 PDF Abstract BibTeX arXiv:2408.16331

Code (1)

logikon-ai/logikon 공식 구현

Similar Papers 제목 키워드 기반

A tutorial introduction to the minimum description length principle

2004-06-04 · Peter Grunwald

This tutorial provides an overview of and introduction to Rissanen's Minimum Description Length (MDL) Principle. The first chapter provides a conceptual, entirely non-technical introduction to the subject. It serves as a…

Introduction and Ranking Results of the ICSI 2014 Competition on Single Objective Optimization

2015-01-09 · Ying Tan, Junzhi Li, Zhongyang Zheng

This technical report includes the introduction and ranking results of the ICSI 2014 Competition on Single Objective Optimization.

OpenRFT: Adapting Reasoning Foundation Model for Domain-specific Tasks with Reinforcement Fine-Tuning

2024-12-22 · Yuxiang Zhang, YuQi Yang, Jiangming Shu, Yuhang Wang 외

OpenAI's recent introduction of Reinforcement Fine-Tuning (RFT) showcases the potential of reasoning foundation model and offers a new paradigm for fine-tuning beyond simple pattern imitation. This technical report prese…

GANterpretations

2020-11-06 · Pablo Samuel Castro

Since the introduction of Generative Adversarial Networks (GANs) [Goodfellow et al., 2014] there has been a regular stream of both technical advances (e.g., Arjovsky et al. [2017]) and creative uses of these generative m…

Score-based Diffusion Models via Stochastic Differential Equations -- a Technical Tutorial

2024-02-12 · Wenpin Tang, Hanyang Zhao

This is an expository article on the score-based diffusion models, with a particular focus on the formulation via stochastic differential equations (SDE). After a gentle introduction, we discuss the two pillars in the di…

reinforcement-learning