Fast CUDA SoftDTW for PyTorch Launched
A new GPU-accelerated, memory-efficient SoftDTW implementation for PyTorch addresses speed and memory limits in time series alignment. It offers 67x speedup, 98% less GPU memory, and supports long sequences beyond 1024. Includes barycenters for DTW-space averaging with full autograd support.





