Archives AI News

Emergency Preemption Without Online Exploration: A Decision Transformer Approach

arXiv:2603.22315v1 Announce Type: new Abstract: Emergency vehicle (EV) response time is a critical determinant of survival outcomes, yet deployed signal preemption strategies remain reactive and uncontrollable. We propose a return-conditioned framework for emergency corridor optimization based on the Decision Transformer…

Universal Approximation Theorem for Input-Connected Multilayer Perceptrons

arXiv:2601.14026v2 Announce Type: replace Abstract: We present the Input-Connected Multilayer Perceptron (IC-MLP), a feedforward neural network architecture in which each hidden neuron receives, in addition to the outputs of the preceding layer, a direct affine connection from the raw input.…

FIPO: Eliciting Deep Reasoning with Future-KL Influenced Policy Optimization

arXiv:2603.19835v2 Announce Type: replace Abstract: We present Future-KL Influenced Policy Optimization (FIPO), a reinforcement learning algorithm designed to overcome reasoning bottlenecks in large language models. While GRPO style training scales effectively, it typically relies on outcome-based rewards (ORM) that distribute…

Equivariance via Minimal Frame Averaging for More Symmetries and Efficiency

arXiv:2406.07598v5 Announce Type: replace Abstract: We consider achieving equivariance in machine learning systems via frame averaging. Current frame averaging methods involve a costly sum over large frames or rely on sampling-based approaches that only yield approximate equivariance. Here, we propose…