paper-with-me

홈 › Papers

Brief analysis of DeepSeek R1 and it's implications for Generative AI

2025-02-04 · Sarah Mercer, Samuel Spillard, Daniel P. Martin

In late January 2025, DeepSeek released their new reasoning model (DeepSeek R1); which was developed at a fraction of the cost yet remains competitive with OpenAI's models, despite the US's GPU export ban. This report discusses the model, and what its release means for the field of Generative AI more widely. We briefly discuss other models released from China in recent weeks, their similarities; innovative use of Mixture of Experts (MoE), Reinforcement Learning (RL) and clever engineering appear to be key factors in the capabilities of these models. This think piece has been written to a tight time-scale, providing broad coverage of the topic, and serves as introductory material for those looking to understand the model's technical advancements, as well as it's place in the ecosystem. Several further areas of research are identified.

📄 PDF Abstract BibTeX arXiv:2502.02523

Code (0)

등록된 구현이 없습니다.

Tasks

GPUMixture-of-ExpertsReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

From ChatGPT to DeepSeek AI: A Comprehensive Analysis of Evolution, Deviation, and Future Implications in AI-Language Models

2025-04-04 · Simrandeep Singh, Shreya Bansal, Abdulmotaleb El Saddik, Mukesh Saini

The rapid advancement of artificial intelligence (AI) has reshaped the field of natural language processing (NLP), with models like OpenAI ChatGPT and DeepSeek AI. Although ChatGPT established a strong foundation for con…

Multiple-choice

Constitutional Precedent of Amicus Briefs

2016-06-15 · Allen Huang, Lars Roemheld

We investigate shared language between U.S. Supreme Court majority opinions and interest groups' corresponding amicus briefs. Specifically, we evaluate whether language that originated in an amicus brief acquired legal p…

When AI output tips to bad but nobody notices: Legal implications of AI's mistakes

2026-03-25 · Dylan J. Restrepo, Nicholas J. Restrepo, Frank Y. Huo, Neil F. Johnson arxiv

The adoption of generative AI across commercial and legal professions offers dramatic efficiency gains -- yet for law in particular, it introduces a perilous failure mode in which the AI fabricates fictitious case law, s…

Legal Reasoning

DeepSeek on a Trip: Inducing Targeted Visual Hallucinations via Representation Vulnerabilities

2025-02-11 · Chashi Mahiul Islam, Samuel Jacob Chacko, Preston Horne, Xiuwen Liu

Multimodal Large Language Models (MLLMs) represent the cutting edge of AI technology, with DeepSeek models emerging as a leading open-source alternative offering competitive performance to closed-source systems. While th…

HallucinationSSIM

Artificial Intelligence and Legal Analysis: Implications for Legal Education and the Profession

2025-02-04 · Lee Peoples

This article reports the results of a study examining the ability of legal and non-legal Large Language Models to perform legal analysis using the Issue-Rule-Application-Conclusion framework. LLMs were tested on legal re…

Legal Reasoning