Nvidia researchers developed dynamic memory sparsification (DMS), a technique that compresses the KV cache in large language ...
New garbage collector promises a 10% to 40% reduction in garbage collection overhead in real-world programs that rely heavily on garbage collection.