Search papers, labs, and topics across Lattice.
This paper introduces Brevis, a novel approach to lossless tensor compression that leverages program synthesis to create a domain-specific language (DSL) tailored for tensor structures. By synthesizing a self-contained DSL program that reconstructs tensors bit-exactly, Brevis achieves a significant reduction in storage requirements, compressing 2.13 TB of checkpoint data to 1.41 TB, which translates to a 33.93% reduction. The method outperforms both general-purpose and tensor-specific compressors, achieving high speeds of 3.60 GB/s for compression and 6.61 GB/s for decompression while maintaining data integrity.
Brevis compresses tensor data by synthesizing a DSL program, achieving over 30% smaller archives than leading general-purpose compressors while ensuring bit-exact reconstruction.
Model checkpoints are growing in both number and size, which makes archival, transfer, and deployment increasingly costly. General-purpose compressors can reduce storage requirements but ignore tensor structure, whereas existing tensor-specific compressors rely on fixed and format-specific pipelines. We present Brevis, which formulates lossless tensor compression as program synthesis. We design a typed domain-specific language (DSL) that captures recurring tensor structures, such as repeated regions and floating-point fields, through a set of reversible operators. Given a tensor, Brevis synthesizes a self-contained DSL program that reconstructs it bit-exactly. A checkpoint-specific production prior, learned from a small representative sample of tensors, guides a bounded A* search to synthesize compact programs, which can later be executed directly for bit-exact decompression. On 10 public checkpoints spanning language, audio, and image generation models, Brevis reduces 2.13 TB of checkpoint data to 1.41 TB, a 33.93% storage reduction. It produces archives up to 30.87% smaller than those of four general-purpose compressors, including zstd and gzip, and smaller archives than the tensor-specific compressors ZipNN and DFloat11. Under a practical concurrency configuration, Brevis achieves 3.60 GB/s compression and 6.61 GB/s decompression while preserving every source byte.