paper-with-me

홈 › Papers

DimmWitted: A Study of Main-Memory Statistical Analytics

2014-03-28 · Ce Zhang, Christopher Ré

We perform the first study of the tradeoff space of access methods and replication to support statistical analytics using first-order methods executed in the main memory of a Non-Uniform Memory Access (NUMA) machine. Statistical analytics systems differ from conventional SQL-analytics in the amount and types of memory incoherence they can tolerate. Our goal is to understand tradeoffs in accessing the data in row- or column-order and at what granularity one should share the model and data for a statistical task. We study this new tradeoff space, and discover there are tradeoffs between hardware and statistical efficiency. We argue that our tradeoff study may provide valuable information for designers of analytics engines: for each system we consider, our prototype engine can run at least one popular task at least 100x faster. We conduct our study across five architectures using popular models including SVMs, logistic regression, Gibbs sampling, and neural networks.

📄 PDF Abstract BibTeX arXiv:1403.7550

Code (1)

HazyResearch/CaffeConTroll

Similar Papers 제목 키워드 기반

A Scalable Framework for Multilevel Streaming Data Analytics using Deep Learning

2019-07-15 · Shihao Ge, Haruna Isah, Farhana Zulkernine, Shahzad Khan

The rapid growth of data in velocity, volume, value, variety, and veracity has enabled exciting new opportunities and presented big challenges for businesses of all types. Recently, there has been considerable interest i…

Sentiment Analysis

MemLens: A Value-Aware Memory Management System with Interactive Analytics for LLM-based Agents

2026-07-28 · Shuyue Wei, Chang Liu, Zimu Zhou, Yongxin Tong 외 arxiv

Recently, memory management has become a key infrastructure for LLM-based agents, as it directly affects long-horizon reasoning, personalized responses, and knowledge reuse. However, existing LLM memory systems typically…

A Systematic Literature Review about Idea Mining: The Use of Machine-driven Analytics to Generate Ideas

2022-01-30 · Workneh Y. Ayele, Gustaf Juell-Skielse

Idea generation is the core activity of innovation. Digital data sources, which are sources of innovation, such as patents, publications, social media, websites, etc., are increasingly growing at unprecedented volume. Ma…

ArticlesInformation RetrievalMorphological AnalysisRetrieval+1

Darwin: A DRAM-based Multi-level Processing-in-Memory Architecture for Data Analytics

2023-05-23 · Donghyuk Kim, Jae-Young Kim, Wontak Han, Jongsoon Won 외

Processing-in-memory (PIM) architecture is an inherent match for data analytics application, but we observe major challenges to address when accelerating it using PIM. In this paper, we propose Darwin, a practical LRDIMM…

CPU

Jointly Optimizing Preprocessing and Inference for DNN-based Visual Analytics

2020-07-25 · Daniel Kang, Ankit Mathur, Teja Veeramacheneni, Peter Bailis 외

While deep neural networks (DNNs) are an increasingly popular way to query large corpora of data, their significant runtime remains an active area of research. As a result, researchers have proposed systems and optimizat…

CPUGPU