#MachineLearning state-space models are incredible but custom CUDA kernels lock its performance to NVIDIA hardware.
My latest #JAX port maps the SSD algorithm to XLA passes, achieving true O(1) on-device caching across CPU, GPU & #TPU.
Pre-print 👉 huggingface.co/papers/2603....
#AI #SSM