Generative Models for Effective ML on Private, Decentralized Datasets - Citegraph

Paper Info

Title
Generative Models for Effective ML on Private, Decentralized Datasets

Abstract
To improve real-world applications of machine learning, experienced modelers develop intuition about their datasets, their models, and how the two interact. Manual inspection of raw data—of representative samples, of outliers, of misclassifications—is an essential tool in a) identifying and fixing problems in the data, b) generating new modeling hypotheses, and c) assigning or refining human-provided labels. However, manual data inspection is risky for privacy-sensitive datasets, such as those representing the behavior of real-world individuals. Furthermore, manual data inspection is impossible in the increasingly important setting of federated learning, where raw examples are stored at the edge and the modeler may only access aggregated outputs such as metrics or model parameters. This paper demonstrates that generative models—trained using federated methods and with formal differential privacy guarantees—can be used effectively to debug data issues even when the data cannot be directly inspected. We explore these methods in applications to text with differentially private federated RNNs and to images using a novel algorithm for differentially private federated GANs.

Year	Venue	Keywords
2020	ICLR	generative models, federated learning, decentralized learning, differential privacy, privacy, security, GAN
DocType	Citations	PageRank
Conference	0	0.34
References	Authors
33	8

Authors (8 rows)

Cited by (0 rows)

References (33 rows)

Name	Order	Citations	PageRank
Sean Augenstein	1	11	1.57
H. Brendan McMahan	2	1630	105.15
Daniel Ramage	3	2109	93.77
Swaroop Ramaswamy	4	0	2.70
Peter Kairouz	5	189	22.27
Mingqing Chen	6	35	5.51
Rajiv Mathews	7	12	5.30
Blaise Agüera y Arcas	8	0	0.68

1