Lossless Compression with Probabilistic Circuits

From MaRDI portal
Publication:6383715

arXiv2111.11632MaRDI QIDQ6383715

Author name not available (Why is that?)

Publication date: 22 November 2021

Abstract: Despite extensive progress on image generation, common deep generative model architectures are not easily applied to lossless compression. For example, VAEs suffer from a compression cost overhead due to their latent variables. This overhead can only be partially eliminated with elaborate schemes such as bits-back coding, often resulting in poor single-sample compression rates. To overcome such problems, we establish a new class of tractable lossless compression models that permit efficient encoding and decoding: Probabilistic Circuits (PCs). These are a class of neural networks involving |p| computational units that support efficient marginalization over arbitrary subsets of the D feature dimensions, enabling efficient arithmetic coding. We derive efficient encoding and decoding schemes that both have time complexity mathcalO(log(D)cdot|p|), where a naive scheme would have linear costs in D and |p|, making the approach highly scalable. Empirically, our PC-based (de)compression algorithm runs 5-40 times faster than neural compression algorithms that achieve similar bitrates. By scaling up the traditional PC structure learning pipeline, we achieve state-of-the-art results on image datasets such as MNIST. Furthermore, PCs can be naturally integrated with existing neural compression algorithms to improve the performance of these base models on natural image datasets. Our results highlight the potential impact that non-standard learning architectures may have on neural data compression.




Has companion code repository: https://github.com/juice-jl/pressedjuice.jl








This page was built for publication: Lossless Compression with Probabilistic Circuits

Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q6383715)