Yipin Guo, Yilin Lang, Qinyuan Ren: GPTQT: Quantize Large Language Models Twice to Push the Efficiency. CIS-RAM 2024: 368-373