• Home
  • Search
  • Dual Query: Practical Private Query Release for High Dimensional Data
  • Cite Icon46
  • https://doi.org/10.29012/jpc.v7i2.650Copy DOI Icon

Dual Query: Practical Private Query Release for High Dimensional Data

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

We present a practical, differentially private algorithm for answering a large number of queries on high dimensional datasets. Like all algorithms for this task, ours necessarily has worst-case complexity exponential in the dimension of the data. However, our algorithm packages the computationally hard step into a concisely defined integer program, which can be solved non-privately using standard solvers. We prove accuracy and privacy theorems for our algorithm, and then demonstrate experimentally that our algorithm performs well in practice. For example, our algorithm can efficiently and accurately answer millions of queries on the Netflix dataset, which has over 17,000 attributes; this is an improvement on the state of the art by multiple orders of magnitude.

Similar Papers
  • Research Article
  • Citations75

DBFS: An effective Density Based Feature Selection scheme for small sample size and high dimensional imbalanced data sets

  • Aug 17, 2012
  • Data & Knowledge Engineering
  • Mina Alibeigi +2
  • Research Article
  • Citations52

Distributed feature selection (DFS) strategy for microarray gene expression data to improve the classification performance

  • Apr 27, 2018
  • Clinical Epidemiology and Global Health
  • Sai Prasad Potharaju +1
  • Book Chapter
  • Citations22

A Binary Firefly Algorithm Based Feature Selection Method on High Dimensional Intrusion Detection Data

  • Jan 01, 2022
  • Yakub Kayode Saheed
  • PDF
  • Research Article
  • Citations12

Enhanced Firefly-K-Means Clustering with Adaptive Mutation and Central Limit Theorem for Automatic Clustering of High-Dimensional Datasets

  • Nov 30, 2022
  • Applied Sciences
  • Abiodun M Ikotun +1
  • Research Article
  • Citations19

Tracking recurrence of correlation structure in neuronal recordings

  • Oct 13, 2016
  • Journal of Neuroscience Methods
  • Samuel A Neymotin +4
  • Research Article
  • Citations9

High-dimensional sparse vine copula regression with application to genomic prediction.

  • Jan 29, 2024
  • Biometrics
  • Özge Sahin +1
  • PDF
  • Research Article
  • Citations6

A Novel Density-based Technique for Outlier Detection of High Dimensional Data Utilizing Full Feature Space

  • Mar 25, 2021
  • Information Technology and Control
  • Mujeeb Ur Rehman +1
  • Research Article
  • Citations28

Occam's razor in dimension reduction: Using reduced row Echelon form for finding linear independent features in high dimensional microarray datasets

  • Apr 22, 2017
  • Engineering Applications of Artificial Intelligence
  • Mohammad Kazem Ebrahimpour +3
  • Research Article
  • Citations6

An effective heuristic for developing hybrid feature selection in high dimensional and low sample size datasets.

  • Dec 26, 2024
  • BMC bioinformatics
  • Hyunseok Shin +1
  • Research Article
  • Citations4

Improved subspace-based and angle-based outlier detections for fuzzy datasets with a real case study

  • Apr 28, 2022
  • Journal of Intelligent & Fuzzy Systems
  • Alireza Fakharzadeh Jahromi +3
  • Preprint Article

Bayesian unidimensional scaling

  • Aug 09, 2017
  • F1000Research
  • Lan Huong Nguyen +1
  • Research Article
  • Citations5

A LoOP based outlier detection method for high dimensional fuzzy data set

  • Jan 13, 2017
  • Journal of Intelligent & Fuzzy Systems
  • Alireza Fakharzadeh Jahromi +1
  • Conference Article
  • Citations10

SSDP+: A Diverse and More Informative Subgroup Discovery Approach for High Dimensional Data

  • Jul 01, 2018
  • Tarcsio Lucas +2
  • Research Article

A new variant of radial visualization for supervised visualization of high dimensional data

  • Nov 15, 2019
  • Transport and Communications Science Journal
  • Long Tran Van +1
  • PDF
  • Research Article
  • Citations673

Evaluation of variable selection methods for random forests and omics data sets.

  • Oct 16, 2017
  • Briefings in bioinformatics
  • Frauke Degenhardt +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.