Hermes: Accelerating Long-Latency Load Requests via Perceptron-Based Off-Chip Load Prediction
arXiv:2209.00188v4 Announce Type: replace-cross Abstract: Long-latency load requests continue to limit the performance of high-performance processors. To increase the latency tolerance of a processor, architects have primarily relied on two key techniques: sophisticated data prefetchers and large on-chip caches. In…
