11 results found Sort:

257
2.2k
apache-2.0
33
SOTA low-bit LLM quantization (INT8/FP8/INT4/FP4/NF4) & sparsity; leading model compression techniques on TensorFlow, PyTorch, and ONNX Runtime
Created 2020-07-21
3,619 commits to master branch, last one a day ago
Must read research papers and links to tools and datasets that are related to using machine learning for compilers and systems optimisation
Created 2020-06-17
242 commits to master branch, last one about a month ago
69
1.3k
other
24
bpftune uses BPF to auto-tune Linux systems
Created 2023-05-09
545 commits to main branch, last one 17 hours ago
50
287
apache-2.0
10
Kernel Tuner
Created 2016-03-28
2,090 commits to master branch, last one about a month ago
26
235
apache-2.0
18
Machine Learning Framework for Operating Systems - Brings ML to Linux kernel
Created 2021-11-10
28 commits to main branch, last one 2 years ago
150
223
mit
56
Stretching GPU performance for GEMMs and tensor contractions.
Created 2015-11-05
5,537 commits to develop branch, last one a day ago
36
170
other
18
CLTune: An automatic OpenCL & CUDA kernel tuner
Created 2015-01-11
307 commits to master branch, last one about a year ago
10
123
apache-2.0
9
Alchemy Cat —— 🔥Config System for SOTA
Created 2019-12-07
388 commits to master branch, last one 3 months ago
15
88
bsd-3-clause
15
Phoebe
Created 2021-01-05
216 commits to main branch, last one 3 years ago
29
73
unknown
9
Benchmark scripts for TVM
Created 2020-11-19
4 commits to main branch, last one 2 years ago
3
67
apache-2.0
3
ebpf profiler for jvm
Created 2020-02-24
156 commits to master branch, last one 3 years ago