I demonstrate how to compress a neural network using pruning in tensorflow.
-
Updated
Aug 2, 2017 - Python
I demonstrate how to compress a neural network using pruning in tensorflow.
The Unified Latent-State Memory Fabric (UL-SMF) is a hardware-software co-designed memory compression fabric that solves the memory bottleneck in long-context Transformer inference. By combining FSQ with dynamic 16-dimensional latent mapping, UL-SMF compresses Key-Value (KV) cache tensors by up to 384x while maintaining >94% semantic retention.
Ravdec is a Python module implementing a lossless data compression algorithm designed by Ravin Kumar on September 19, 2016. This algorithm is designed exclusively for textual data, including alphabets, numbers, and symbols.
A payload compression toolkit that makes it easy to create ideal data structures for LLMs; from training data to chain payloads.
Information theory and methods of data compression - PUT Poznań 2018
my projects for the course of "data and image compression" at UPC (Universitat Politecnica de Catalunya)
To associate your repository with the compression-methods topic, visit your repo's landing page and select "manage topics."