paper-with-me

홈 › Papers

Google Dataset Search by the Numbers

2020-06-12 · Omar Benjelloun, Shi-Yu Chen, Natasha Noy

Scientists, governments, and companies increasingly publish datasets on the Web. Google's Dataset Search extracts dataset metadata -- expressed using schema.org and similar vocabularies -- from Web pages in order to make datasets discoverable. Since we started the work on Dataset Search in 2016, the number of datasets described in schema.org has grown from about 500K to almost 30M. Thus, this corpus has become a valuable snapshot of data on the Web. To the best of our knowledge, this corpus is the largest and most diverse of its kind. We analyze this corpus and discuss where the datasets originate from, what topics they cover, which form they take, and what people searching for datasets are interested in. Based on this analysis, we identify gaps and possible future work to help make data more discoverable.

📄 PDF Abstract BibTeX arXiv:2006.06894

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Handwritten and Machine printed OCR for Geez Numbers Using Artificial Neural Network

2019-11-15 · Eyob Gebretinsae Beyene

Researches have been done on Ethiopic scripts. However studies excluded the Geez numbers from the studies because of different reasons. This paper presents offline handwritten and machine printed Geez number recognition …

Image RetrievalOptical Character Recognition (OCR)

The Causality Inference of Public Interest in Restaurants and Bars on COVID-19 Daily Cases in the US: A Google Trends Analysis

2020-07-27 · Milad Asgari Mehrabadi, Nikil Dutt, Amir M. Rahmani

The COVID-19 coronavirus pandemic has affected virtually every region of the globe. At the time of conducting this study, the number of daily cases in the United States is more than any other country, and the trend is in…

The Google Similarity Distance

2004-12-21 · Rudi Cilibrasi, Paul M. B. Vitanyi

Words and phrases acquire meaning from the way they are used in society, from their relative semantics to other words and phrases. For computers the equivalent of `society' is `database,' and the equivalent of `use' is `…

Binary ClassificationClusteringTranslation

A Framework for Fast Scalable BNN Inference using Googlenet and Transfer Learning

2021-01-04 · Karthik E

Efficient and accurate object detection in video and image analysis is one of the major beneficiaries of the advancement in computer vision systems with the help of deep learning. With the aid of deep learning, more powe…

GPUimage-classificationImage ClassificationObject+4

Deep Convolutional Neural Network Inference with Floating-point Weights and Fixed-point Activations

2017-03-08 · Liangzhen Lai, Naveen Suda, Vikas Chandra

Deep convolutional neural network (CNN) inference requires significant amount of memory and computation, which limits its deployment on embedded devices. To alleviate these problems to some extent, prior research utilize…