• Cite Icon52
  • https://doi.org/10.14778/3407790.3407854Copy DOI Icon

SAQE

Show More
  • Abstract
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

A private data federation enables clients to query the union of data from multiple data providers without revealing any extra private information to the client or any other data providers. Unfortunately, this strong end-to-end privacy guarantee requires cryptographic protocols that incur a significant performance overhead as high as 1,000 x compared to executing the same query in the clear. As a result, private data federations are impractical for common database workloads. This gap reveals the following key challenge in a private data federation: offering significantly fast and accurate query answers without compromising strong end-to-end privacy. To address this challenge, we propose SAQE, the Secure Approximate Query Evaluator, a private data federation system that scales to very large datasets by combining three techniques --- differential privacy, secure computation, and approximate query processing --- in a novel and principled way. First, SAQE adds novel secure sampling algorithms into the federation's query processing pipeline to speed up query workloads and to minimize the noise the system must inject into the query results to protect the privacy of the data. Second, we introduce a query planner that jointly optimizes the noise introduced by differential privacy with the sampling rates and resulting error bounds owing to approximate query processing. Our research shows that these three techniques are synergistic: sampling within certain accuracy bounds improves both query privacy and performance, meaning that SAQE executes over less data than existing techniques without sacrificing efficiency, privacy, or accuracy. Using our optimizer, we leverage this counter-intuitive result to identify an inflection point that maximizes all three criteria prior query evaluation. Experimentally, we show that this result enables SAQE to trade-off among these three criteria to scale its query processing to very large datasets with accuracy bounds dependent only on sample size, and not the raw data size.

Similar Papers
  • PDF
  • Research Article
  • Citations141

Approximate Query Processing: What is New and Where to Go?

  • Sep 14, 2018
  • Data Science and Engineering
  • Kaiyu Li +1
  • Conference Article
  • Citations2

An Agile Sample Maintenance Approach for Agile Analytics

  • Apr 01, 2020
  • Hanbing Zhang +5
  • Research Article

SPRINT: Scalable Secure & Differentially Private Inference for Transformers

  • Jan 01, 2026
  • Proceedings on Privacy Enhancing Technologies
  • Francesco Capano +2
  • Research Article
  • Citations19

DPLQ: Location‐based service privacy protection scheme based on differential privacy

  • Aug 20, 2021
  • IET Information Security
  • Qingyun Zhang +3
  • Book Chapter
  • Citations1

Learning-Based Optimization for Online Approximate Query Processing

  • Jan 01, 2022
  • Wenyuan Bi +5
  • Conference Article
  • Citations20

An Adaptive Differential Privacy Algorithm for Range Queries over Healthcare Data

  • Aug 01, 2017
  • Asma Alnemari +2
  • Research Article
  • Citations7

Query Evaluation under Differential Privacy

  • Oct 30, 2023
  • ACM SIGMOD Record
  • Wei Dong +1
  • Conference Article
  • Citations93

Composing Differential Privacy and Secure Computation

  • Oct 30, 2017
  • Xi He +3
  • Conference Article
  • Citations4

Securely Sampling Discrete Gaussian Noise for Multi-Party Differential Privacy

  • Nov 15, 2023
  • Chengkun Wei +4
  • Research Article
  • Citations2

A Novel Differentially Private Online Learning Algorithm for Group Lasso in Big Data

  • Jan 01, 2024
  • IET Information Security
  • Jinxia Li +1
  • PDF
  • Research Article

Differentially private range counting: where asymptotically better fails, integer covering prevails

  • Jun 19, 2025
  • Annals of Operations Research
  • Hafiz Asif +2
  • Research Article
  • Citations36

Investigating Statistical Privacy Frameworks from the Perspective of Hypothesis Testing

  • Jul 01, 2019
  • Proceedings on Privacy Enhancing Technologies
  • Changchang Liu +4
  • Research Article
  • Citations3

MISS: finding optimal sample sizes for approximate analytics

  • Oct 21, 2021
  • Distributed and Parallel Databases
  • Xuebin Su +1
  • PDF
  • Research Article
  • Citations1

Hybrid GNN–LSTM defense with differential privacy and secure multi-party computation for edge-optimized neuromorphic autonomous systems

  • Dec 16, 2025
  • Scientific Reports
  • Siwar Rekik +1
  • Conference Article
  • Citations8

COMPASS

  • Jul 09, 2018
  • Haoyuan Xing +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.