Quantized Transformers in Practice: Benchmarking Full- and Low-Precision LLMs across Two Processors | Synapse