Posts

Showing posts with the label dynamic quantization

Benchmarking Dynamic Quantization for Larger Language Models