Randomized Dimensionality Reduction for Euclidean Maximization and Diversity Measures

Gao, Jie, Jayaram, Rajesh, Kolbe, Benedikt, Sapir, Shay, Schwiegelshohn, Chris, Silwal, Sandeep, Waingarten, Erik

Jun-3-2025–arXiv.org Artificial Intelligence

Randomized dimensionality reduction is a widely-used algorithmic technique for speeding up large-scale Euclidean optimization problems. In this paper, we study dimension reduction for a variety of maximization problems, including max-matching, max-spanning tree, max TSP, as well as various measures for dataset diversity. For these problems, we show that the effect of dimension reduction is intimately tied to the \emph{doubling dimension} $λ_X$ of the underlying dataset $X$ -- a quantity measuring intrinsic dimensionality of point sets. Specifically, we prove that a target dimension of $O(λ_X)$ suffices to approximately preserve the value of any near-optimal solution,which we also show is necessary for some of these problems. This is in contrast to classical dimension reduction results, whose dependence increases with the dataset size $|X|$. We also provide empirical results validating the quality of solutions found in the projected space, as well as speedups due to dimensionality reduction.

artificial intelligence, dimension, machine learning, (17 more...)

arXiv.org Artificial Intelligence

Jun-3-2025

arXiv.org PDF

Add feedback

Country:
- Asia
  - Afghanistan > Parwan Province
    - Charikar (0.04)
  - Middle East > Israel (0.04)
- Europe
  - Denmark (0.04)
  - Germany > North Rhine-Westphalia
    - Cologne Region > Bonn (0.04)
  - Italy (0.04)
  - Slovenia > Drava
    - Municipality of Benedikt > Benedikt (0.04)
  - United Kingdom > England
    - Cambridgeshire > Cambridge (0.04)
- North America
  - Canada > British Columbia (0.04)
  - United States
    - California > Alameda County
      - Berkeley (0.04)
    - Florida > Miami-Dade County
      - Miami Beach (0.04)
    - Massachusetts > Middlesex County
      - Cambridge (0.04)
    - Nevada (0.04)
    - New Jersey > Middlesex County
      - Piscataway (0.04)
    - Oregon > Multnomah County
      - Portland (0.04)
    - Pennsylvania > Philadelphia County
      - Philadelphia (0.04)
    - Wisconsin > Dane County
      - Madison (0.04)

Genre:
- Research Report (0.82)

Technology:
- Information Technology
  - Artificial Intelligence > Machine Learning
    - Statistical Learning > Dimensionality Reduction (0.82)
  - Data Science (1.00)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found