Sinkhorn: Make LLMs even smaller through quantisation while maintaining accuracy(github.com)4 points by ilitirit 317 days ago | 1 comment