10 results found Sort:
- Filter by Primary Language:
- Python (4)
- C (3)
- C++ (2)
- +
SOTA low-bit LLM quantization (INT8/FP8/INT4/FP4/NF4) & sparsity; leading model compression techniques on TensorFlow, PyTorch, and ONNX Runtime
Created
2020-07-21
3,477 commits to master branch, last one 6 days ago
Must read research papers and links to tools and datasets that are related to using machine learning for compilers and systems optimisation
Created
2020-06-17
234 commits to master branch, last one 28 days ago
bpftune uses BPF to auto-tune Linux systems
Created
2023-05-09
528 commits to main branch, last one 19 days ago
Kernel Tuner
Created
2016-03-28
2,041 commits to master branch, last one 15 days ago
Machine Learning Framework for Operating Systems - Brings ML to Linux kernel
Created
2021-11-10
28 commits to main branch, last one 2 years ago
Stretching GPU performance for GEMMs and tensor contractions.
Created
2015-11-05
5,436 commits to develop branch, last one 5 hours ago
CLTune: An automatic OpenCL & CUDA kernel tuner
Created
2015-01-11
307 commits to master branch, last one about a year ago
Phoebe
Created
2021-01-05
216 commits to main branch, last one 3 years ago
Benchmark scripts for TVM
Created
2020-11-19
4 commits to main branch, last one 2 years ago
ebpf profiler for jvm
Created
2020-02-24
156 commits to master branch, last one 3 years ago