TurboQuant is not a step change, it's more of a smaller incremental improvement ...

		zozbot234 30 days ago \| parent \| context \| favorite \| on: The threat is comfortable drift toward not underst... TurboQuant is not a step change, it's more of a smaller incremental improvement to KV quantization, and possibly (unsure) to quantization more generally. I'm actually more positive about SSD weights offload, which opens up very large local models for slow inference (good enough for slow chat) to virtually any hardware or amount of RAM.