paper-with-me

Papers

Subjective Learning for Open-Ended Data

2021-08-27 · Tianren Zhang, Yizhou Jiang, Xin Su, Shangqi Guo, Feng Chen

Conventional supervised learning typically assumes that the learning task can be solved by learning a single function since the data is sampled from a fixed distribution. However, this assumption is invalid in open-ended environments where no task-level data partitioning is available. In this paper, we present a novel supervised learning framework of learning from open-ended data, which is modeled as data implicitly sampled from multiple domains with the data in each domain obeying a domain-specific target function. Since different domains may possess distinct target functions, open-ended data inherently requires multiple functions to capture all its input-output relations, rendering training a single global model problematic. To address this issue, we devise an Open-ended Supervised Learning (OSL) framework, of which the key component is a subjective function that allocates the data among multiple candidate models to resolve the "conflict" between the data from different domains, exhibiting a natural hierarchy. We theoretically analyze the learnability and the generalization error of OSL, and empirically validate its efficacy in both open-ended regression and classification tasks.

📄 PDF Abstract BibTeX arXiv:2108.12113

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Speech Quality and Testing Framework

2020-01-23 · Chandan K. A. Reddy, Ebrahim Beyrami, Harishchandra Dubey, Vishak Gopal 외

The INTERSPEECH 2020 Deep Noise Suppression Challenge is intended to promote collaborative research in real-time single-channel Speech Enhancement aimed to maximize the subjective (perceptual) quality of the enhanced spe…

Speech Enhancement

Extending RLVR to Open-Ended Tasks via Verifiable Multiple-Choice Reformulation

2025-11-04 · Mengyu Zhang, Siyu Ding, Weichong Yin, Yu Sun 외 arxiv

Reinforcement Learning with Verifiable Rewards(RLVR) has demonstrated great potential in enhancing the reasoning capabilities of large language models (LLMs). However, its success has thus far been largely confined to th…

Reinforcement Learning

RoleRMBench & RoleRM: Towards Reward Modeling for Profile-Based Role Play in Dialogue Systems

2025-12-11 · Hang Ding, Qiming Feng, Dongqi Liu, Qi Zhao 외 arxiv

Reward modeling has become a cornerstone of aligning large language models (LLMs) with human preferences. Yet, when extended to subjective and open-ended domains such as role play, existing reward models exhibit severe d…

The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Testing Framework, and Challenge Results

2020-05-16 · Chandan K. A. Reddy, Vishak Gopal, Ross Cutler, Ebrahim Beyrami 외

The INTERSPEECH 2020 Deep Noise Suppression (DNS) Challenge is intended to promote collaborative research in real-time single-channel Speech Enhancement aimed to maximize the subjective (perceptual) quality of the enhanc…

Speech Enhancement

Self-Referential Induction Increases Response Instability Relative to Unresolvable and Verifiable Questions in Large Language Models

2026-08-13 · Paras Balani, Subhrakanta Panda arxiv

Self-referential prompting has been shown to reliably induce large language models to produce first-person reports resembling subjective experience, but no prior work measures how consistent these reports are across repe…