paper-with-me

Papers

Developer Perspectives on Licensing and Copyright Issues Arising from Generative AI for Software Development

2024-11-16 · Trevor Stalnaker, Nathan Wintersgill, Oscar Chaparro, Laura A. Heymann, Massimiliano Di Penta, Daniel M German, Denys Poshyvanyk

Despite the utility that Generative AI (GenAI) tools provide for tasks such as writing code, the use of these tools raises important legal questions and potential risks, particularly those associated with copyright law. As lawmakers and regulators engage with those questions, the views of users can provide relevant perspectives. In this paper, we provide: (1) a survey of 574 developers on the licensing and copyright aspects of GenAI for coding, as well as follow-up interviews; (2) a snapshot of developers' views at a time when GenAI and perceptions of it are rapidly evolving; and (3) an analysis of developers' views, yielding insights and recommendations that can inform future regulatory decisions in this evolving field. Our results show the benefits developers derive from GenAI, how they view the use of AI-generated code as similar to using other existing code, the varied opinions they have on who should own or be compensated for such code, that they are concerned about data leakage via GenAI, and much more, providing organizations and policymakers with valuable insights into how the technology is being used and what concerns stakeholders would like to see addressed.

📄 PDF Abstract BibTeX arXiv:2411.10877

Code (0)

등록된 구현이 없습니다.

Tasks

MisconceptionsSurvey

Methods 이 논문이 사용한 방법론

AWARE We propose to theoretically and empirically examine the effect of incorporating weighting schemes into walk-aggregating GNNs. To this end, we propose a simple, interpretable, and…

Similar Papers 제목 키워드 기반

Content ARCs: Decentralized Content Rights in the Age of Generative AI

2025-03-14 · Kar Balan, Andrew Gilbert, John Collomosse

The rise of Generative AI (GenAI) has sparked significant debate over balancing the interests of creative rightsholders and AI developers. As GenAI models are trained on vast datasets that often include copyrighted mater…

CodeGenLink: A Tool to Find the Likely Origin and License of Automatically Generated Code

2025-10-01 · Daniele Bifolco, Guido Annicchiarico, Pierluigi Barbiero, Massimiliano Di Penta 외 arxiv

Large Language Models (LLMs) are widely used in software development tasks nowadays. Unlike reusing code taken from the Web, for LLMs' generated code, developers are concerned about its lack of trustworthiness and possib…

The KL3M Data Project: Copyright-Clean Training Resources for Large Language Models

2025-04-10 · Michael J Bommarito II, Jillian Bommarito, Daniel Martin Katz

Practically all large language models have been pre-trained on data that is subject to global uncertainty related to copyright infringement and breach of contract. This creates potential risk for users and developers due…

An investigation of licensing of datasets for machine learning based on the GQM model

2023-03-24 · Junyu Chen, Norihiro Yoshida, Hiroaki Takada

Dataset licensing is currently an issue in the development of machine learning systems. And in the development of machine learning systems, the most widely used are publicly available datasets. However, since the images …

PDMX: A Large-Scale Public Domain MusicXML Dataset for Symbolic Music Processing

2024-09-17 · Phillip Long, Zachary Novack, Taylor Berg-Kirkpatrick, Julian McAuley

The recent explosion of generative AI-Music systems has raised numerous concerns over data copyright, licensing music from musicians, and the conflict between open-source AI and large prestige companies. Such issues high…

Music GenerationTAG