Sketching Datasets for Large-Scale Learning (long version)

Gribonval, Rémi, Chatalic, Antoine, Keriven, Nicolas, Schellekens, Vincent, Jacques, Laurent, Schniter, Philip

Aug-4-2020–arXiv.org Machine Learning

This article considers "sketched learning," or "compressive learning," an approach to large-scale machine learning where datasets are massively compressed before learning (e.g., clustering, classification, or regression) is performed. In particular, a "sketch" is first constructed by computing carefully chosen nonlinear random features (e.g., random Fourier features) and averaging them over the whole dataset. Parameters are then learned from the sketch, without access to the original dataset. This article surveys the current state-of-the-art in sketched learning, including the main concepts and algorithms, their connections with established signal-processing methods, existing theoretical guarantees---on both information preservation and privacy preservation, and important open problems.

data mining, latexit sha1, machine learning, (19 more...)

arXiv.org Machine Learning

Aug-4-2020

arXiv.org PDF

Add feedback

Country:
- North America > United States
  - Ohio > Franklin County
    - Columbus (0.04)
  - New York > New York County
    - New York City (0.04)
- Europe
  - Belgium (0.04)
  - United Kingdom > England
    - Oxfordshire > Oxford (0.04)
    - Cambridgeshire > Cambridge (0.04)
  - France > Auvergne-Rhône-Alpes
    - Isère > Grenoble (0.04)
    - Lyon > Lyon (0.04)

Genre:
- Overview (1.00)
- Research Report (0.83)

Industry:
- Information Technology > Security & Privacy (1.00)

Technology:
- Information Technology
  - Security & Privacy (1.00)
  - Data Science > Data Mining (1.00)
  - Artificial Intelligence
    - Representation & Reasoning (1.00)
    - Machine Learning
      - Performance Analysis > Accuracy (0.93)
      - Statistical Learning > Clustering (0.67)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found