Deep Learning–Based Video Stabilization
Tarun Gangadhar Vadaparthi
Demo
Abstract
We estimate dense optical flow with RAFT and collapse it to mean (dx, dy) per frame pair. A BiLSTM (with Transformer/GRU baselines) is trained to predict smoothed motion sequences supervised by a local moving-average target, reducing jitter and improving visual stability.
Method Overview
Results