PT$^2$-LLM: Post-Training Ternarization for Large Language Models
arXiv:2510.03267v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown impressive capabilities across diverse tasks, but their large memory and compute demands hinder deployment. Ternarization has gained attention as a promising compression technique, delivering substantial size reduction and high…
