Hadamard Response: Estimating Distributions Privately, Efficiently, and with Little Communication

22ND INTERNATIONAL CONFERENCE ON ARTIFICIAL INTELLIGENCE AND STATISTICS, VOL 89(2018)

引用 121|浏览39
暂无评分
摘要
We study the problem of estimating k-ary distributions under ε-local differential privacy. n samples are distributed across users who send privatized versions of their sample to a central server. All previously known sample optimal algorithms require linear (in k) communication from each user in the high privacy regime (ε=O(1)), and run in time that grows as n· k, which can be prohibitive for large domain size k. We propose Hadamard Response (HR, a local privatization scheme that requires no shared randomness and is symmetric with respect to the users. Our scheme has order optimal sample complexity for all ε, a communication of at most log k+2 bits per user, and nearly linear running time of Õ(n + k). Our encoding and decoding are based on Hadamard matrices, and are simple to implement. The statistical performance relies on the coding theoretic aspects of Hadamard matrices, ie, the large Hamming distance between the rows. An efficient implementation of the algorithm using the Fast Walsh-Hadamard transform gives the computational gains. We compare our approach with Randomized Response (RR), RAPPOR, and subset-selection mechanisms (SS), both theoretically, and experimentally. For k=10000, our algorithm runs about 100x faster than SS, and RAPPOR.
更多
查看译文
关键词
estimating distributions,little communication
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络
Chat Paper
正在生成论文摘要