Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

SED:Lightweight Saliency prediction for Event-based data via Distillation

About

Event-based saliency prediction has gained attention recently, as combining event cameras with saliency estimation can act as an upstream stage that naturally improves the efficiency of downstream eventbased perception at the edge. However, current approaches are either neuromorphic, underperforming on event-based saliency benchmarks, or too heavy for resource-constrained edge applications due to their reliance on transformers or 3D convolutions. Drawing inspiration from efficient convolutional modules, SED and aiming to exploit the temporal information in event data, we propose a lightweight network, trained through knowledge distillation, built on a Depthwise Spatio-Temporal Block (DSTconv) -- a factorization of the 3D depthwise separable convolution. Relative to its teacher, our model reduces the model size from 180 MB to 0.32 MB (562x) and the parameter count from 45M to 81k (554x), while matching or outperforming it on the N-DHF1K and N-UCF Sports datasets. Moreover, it generalizes strongly beyond its training distribution, transferring from synthetic to real event data where a model trained from scratch fails.

Romaric Mazna, Jean Martinet, Michele Magno• 2026

Related benchmarks

TaskDatasetResultRank
Saliency PredictionN-DHF1K
AUC-J0.9047
25
Saliency PredictionN-UCF Sports
AUC-J0.9187
25
Showing 2 of 2 rows

Other info

Follow for update