Posts

Showing posts with the label quantization

Benchmarking Dynamic Quantization for Larger Language Models