Search arXivSearch

arXiv · 2606.22802

Private Information Retrieval from Joint Systematic MDS-Coded with Non-Colluding Servers: Bounds and Constructions

Abstract

Consider a distributed storage system consisting of $N$ non-colluding servers that collectively store a database of $M$ files encoded using an $[N,K]$ maximum distance separable(MDS) code. A user wishes to retrieve one file privately by accessing the servers without revealing the identity of the requested file. A scheme designed for this purpose is called a joint MDS-coded private information retrieval(PIR) scheme, which was first introduced by Sun and Tian in 2019 to break the capacity $\frac{1-K/N}{1-(K/N)^M}$ of the separate MDS-coded PIR schemes established by Banawan and Ulukus. However, the capacity of joint MDS-coded PIR remains largely unexplored. In this paper, we study the capacity of joint MDS-coded PIR with systematic MDS array storage codes under prescribed storage patterns. Specifically, we first derive upper bounds on the capacity of joint MDS-coded PIR for $K=Mt$ and $K=Mt+1$, respectively. We then construct three joint MDS-coded PIR schemes for the cases $N\le K+t, K=Mt$, $N>K+t, K=Mt$ and $N\le K+t, K=Mt+1$. The proposed schemes require small file sizes and achieve higher retrieval rates: the first and third schemes exceed the capacity of separate MDS-coded PIR schemes, while the second scheme does so when the storage rate $\frac{K}{N}>r_M$ for some $0<r_M<\frac{M}{M+1}$. In particular, for $K=Mt$ and $N\leq K+t$, the proposed scheme achieves the derived upper bound, thereby establishing that the optimal joint MDS-coded PIR capacity under the considered storage pattern is $1-(1-\frac{1}{M})\frac{K}{N}$. Compared with capacity-achieving separate MDS-coded PIR schemes at the same storage-code rate, the proposed schemes may achieve a substantial relative retrieval-rate improvement: the maximum improvement can exceed $15\%$ when $M\geq 4$, exceed $20\%$ when $M\geq 9$, and asymptotically approach $1-2/e\approx 26.42\%$ as M increases.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Jingke Xu, Lirong Shi, Peng Lan, Weijun Fang. 2026-06-22. Private Information Retrieval from Joint Systematic MDS-Coded with Non-Colluding Servers: Bounds and Constructions. https://arxiv.org/abs/2606.22802

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Radiance-Field Guided Pretraining: Scaling Localization Models with Unlabeled Wireless Signals

Radio frequency (RF)-based indoor localization offers significant promise for applications such as indoor navigation, augmented reality, and pervasive computing. While deep learning has greatly enhanced localization accuracy and robustness, existing localization models still face major challenges in cross-scene generalization due to their reliance on scene-specific labeled data. To address this, we introduce Radiance-Field Reinforced Pretraining (RFRP). This novel self-supervised pretraining framework couples a large localization model (LM) with a neural radio-frequency radiance field (RF-NeRF) in an asymmetrical autoencoder architecture. In this design, the LM encodes received RF spectra into latent, position-relevant representations, while the RF-NeRF decodes them to reconstruct the original spectra. This alignment between input and output enables effective representation learning using large-scale, unlabeled RF data, which can be collected continuously with minimal effort. To this end, we collected RF samples at 7,327,321 positions across 100 diverse scenes using four common wireless technologies--RFID, BLE, WiFi, and IIoT. Data from 75 scenes were used for training, and the remaining 25 for evaluation. Experimental results show that the RFRP-pretrained LM reduces localization error by over 40% compared to non-pretrained models and by 21% compared to those pretrained using supervised learning.

cs.IT

Uniform Recovery of Structured Signals from Nonlinear Observations: Improved Error Rates

Consider the recovery of structured signals from nonlinear observations. Under Gaussian matrix and a large class of unknown nonlinear link functions, Plan and Vershynin (2016) showed that generalized Lasso achieves accurate nonuniform recovery of a fixed signal. More recently, Genzel and Stollenwerk (2023) showed that generalized Lasso is indeed capable of accurately recovering all structured signals. However, in some canonical settings with discontinuous link functions, their uniform recovery error rate is essentially slower than the nonuniform one. Specifically, in the recovery of $n$-dimensional $k$-sparse vectors from $m$ measurements, generalized Lasso with a perfectly tuned $\ell_1$ constraint achieves nonuniform error rate $ O(\sqrt{k\log(en/k)/m})$, while the uniform error rate of Genzel and Stollenwerk is no faster than $O((k\log(en/k)/m)^{1/4})$. In this paper, we narrow this gap by establishing improved uniform recovery guarantees under piecewise Lipschitz link functions with well-separated jump discontinuities. We analyze a projected gradient descent (PGD) algorithm whose projection can be onto a convex set or a cone, and our results for the PGD with a convex set are also valid for the generalized Lasso. In sparse recovery, the improved uniform error rates match the nonuniform rate $O(\sqrt{k\log(en/k)/m})$ up to logarithmic factors. Under the sign link function, we further show that iterative hard thresholding (a specific instance of the PGD) achieves uniform recovery error rate $O(\sqrt{k\log(en/k)/m})$, matching the nonuniform rate up to a universal constant. Technically, the uniform guarantees for the PGD are obtained by showing that the gradient maps satisfy the restricted approximate invertibility condition uniformly over all signals. We demonstrate that this is a general approach to uniform recovery under nonlinear observations.

cs.IT

Enhanced Feedback Mechanisms for Resource-Efficient Incremental Redundancy

Incremental redundancy (IR) can reduce error rates by spreading coded bits across multiple transmission attempts. However, conventional stop-and-wait operation with coarse feedback often over-provisions retransmissions, triggers unnecessary decoding attempts, and increases end-to-end latency. This paper develops enhanced feedback and scheduling mechanisms that predict the additional redundancy needed for successful decoding and allocate only the required resources. We study two complementary strategies. First, using channel statistics, we learn a one- or two-shot mapping from channel quality to the minimum redundancy budget. As a byproduct, we derive an achievable reliability lower bound on the error probability of hybrid automatic repeat request (HARQ) systems. Numerical results with polar-coded IR-HARQ scheme show that the bound can be closely approached by appropriately selecting the second-transmission redundancy over a wide SNR range with savings up to 60\% in retransmission size. Second, we propose a realization-aware early-feedback mechanism that uses first-transmission reliability information to make per-codeword decisions before decoding: whether the codeword is already decodable, if not, how many additional redundancy versions are needed, or whether decoding is unlikely and rate adaptation is preferable. Link-level simulations with 5G NR LDPC codes show that both predictors achieve high accuracy (about 96\% in our study), increasing the probability of successful decoding within at most two transmission occasions.

cs.IT