1 comment

[ 2.9 ms ] story [ 13.5 ms ] thread
Some impressive results

> 1.58-bit FLUX achieves a 7.7× reduction in model storage and more than a 5.1× reduction in inference memory usage