1 result found Sort:
An innovative library for efficient LLM inference via low-bit quantization
Created
2023-11-20
344 commits to main branch, last one 4 days ago