paper-with-me

Value prediction

1개 벤치마크 · 논문 115편 · 이 태스크의 논문 보기 →

Benchmarks

Py150

결과 2개

Most implemented

Value Prediction Network

2017-07-11 · 구현 2개

Papers

Pseudorandom Streams within Diffusion Models Act as Learnable Inputs That Affect Generation Quality

2026-08-03 · Shengzhi Deng, Chenqi Ye, Yanze Guo arxiv

Digital learning systems consume concrete pseudorandom values rather than abstract random variables. These values enter the realized loss and its gradient during training. If a pseudorandom stream contains structure that…

Value prediction

Auto-Fill: Learning to Predict Missing Values Accurately with Specialist Language Models

2026-07-22 · Yurong Liu, Yeye He, Haoyu Dong, Junjie Xing 외 arxiv

Predicting missing cell values in tabular data is a fundamental problem in data cleaning. While state-of-the-art reasoning models show great promise in predicting missing values in tables, by reasoning holistically acros…

Value prediction

NASDAQ: Normalized Observation Space Dynamics-Augmented Q-Learning

2026-06-19 · Xinwei Liu, Junyuan Liang, Zicong Hong, Jianting Zhang 외 arxiv

Augmenting model-free reinforcement learning (RL) with representations learned through observation dynamics prediction (observation-predictive RL) can improve sample efficiency and performance, with minor modifications a…

Reinforcement LearningValue prediction

How Should World Models Be Evaluated for Embodied Decision-Making? A Decision-Making-Centric Position

2026-06-13 · Yang Yu, Shiyuan Zhang, Yifei Sheng, Haoxiang Ren 외 arxiv

World models have become a central abstraction in modern AI. The term now refers to several different objects: action-conditioned environment models, latent imagination models, future-video predictors, interactive neural…

Instruction FollowingValue prediction

FlowBank: Query-Adaptive Agentic Workflows Optimization through Precompute-and-Reuse

2026-06-09 · Lingzhi Yuan, Chenghao Deng, Fangxu Yu, Souradip Chakraborty 외 arxiv

Large Language Model (LLM)-based multi-agent systems are increasingly powerful, but current agentic workflow optimization paradigms make an unsatisfying trade-off. Task-level methods spend substantial offline compute yet…

Value prediction

Q-Delta: Beyond Key-Value Associative State Evolution

2026-06-07 · Sumin Park, Seojin Kim, Noseong Park arxiv

Linear attention reformulates sequence modeling as recurrent state evolution, enabling efficient linear-time inference. Under the key-value associative paradigm, existing approaches restrict the role of the query to the …

Value prediction

전체 115편 보기 →