### H100 Inference Speedups with Quantization - [TensorRTLLM Speed up inference with SOTA quantization techniques in TRT-LLM](https://github.com/NVIDIA/TensorRT-LLM/blob/main/docs/source/blogs/quantization-in-TRT-LLM.md)
H100 Inference Speedups with Quantization