replicate/quantization-eetq
094
1---2library_name: kernels3license: apache-2.04---5 6> [!CAUTION]7> Starting from September 13, 2026, we will be removing the "model" type repositories of kernels (e.g., kernels-community/flash-attn3). Make sure you're using a latest version of kernels. If you face any disruption, please report them here: https://github.com/huggingface/kernels/issues/new.8 9This is the repository card of kernels-community/quantization-eetq that has been pushed on the Hub. It was built to be used with the [`kernels` library](https://github.com/huggingface/kernels). This card was automatically generated.10 11## How to use12 13```python14# make sure `kernels` is installed: `pip install -U kernels`15from kernels import get_kernel16 17kernel_module = get_kernel("kernels-community/quantization-eetq")18w8_a16_gemm = kernel_module.w8_a16_gemm19 20w8_a16_gemm(...)21```22 23## Available functions24- `w8_a16_gemm`25- `w8_a16_gemm_`26- `preprocess_weights`27- `quant_weights`28 29## Benchmarks30 31No benchmark available yet.32 