Enhancing AI Interpretability and Safety through Localised Architectures
arXiv:2606.07998v2 Announce Type: replace Abstract: Recent advances in generative AI, especially powerful Large Language Models (LLMs) and Large Reasoning Models (LRMs), raise concerns over the interpretability, safety and sustainability of these large and opaque AI models. The power of such…
