Visual Place Recognition Using Rate-Encoded Spiking Neural Networks with Discrete STDP Learning

📅 2026-07-15
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
This work addresses the challenge that existing spike-timing-dependent plasticity (STDP)-based spiking neural networks (SNNs) struggle to achieve high recall under 100% precision in visual place recognition. The authors propose a tensor-native, discretized STDP-SNN pipeline incorporating closed-form deterministic tensor neuron assignment, a post-query state reset mechanism, and a velocity-compensated sliding-window frame aggregation strategy. By integrating rate coding, unsupervised STDP learning, and an efficient inference architecture, the method achieves 100.00% recall at 100% precision (R@100P) on the Nordland dataset under constant-velocity conditions, with only 0.20 milliseconds of latency, substantially improving recall performance under stringent accuracy requirements.
📝 Abstract
Spiking Neural Networks (SNNs) trained through unsupervised Spike-Timing-Dependent Plasticity (STDP) have been explored as solutions to visual loop closure problems, driven by the prospect of efficient on-device inference on neuromorphic devices. State-of-the-art STDP-based models deliver high classification accuracy but fail to reach the high Recall at 100% Precision (R@100P) needed for reliable autonomous navigation. We present a discrete, tensor-native implementation of the STDP-based SNN-VPR pipeline using PyTorch with snnTorch and evaluate it on a 100-place Nordland dataset using 15 independently-trained networks. The contribution of three decisions in the implementation is investigated. First, we show how to perform neuron assignment with a closed-form, deterministic tensor pipeline and show that it provides significantly higher R@100P than a standard argmax procedure. However, some of this gain comes from implementation differences compared to prior continuous-time models, which we measure independently. Second, ablation in isolation shows that state reset after each query helps improve R@100P regardless of the way neurons are assigned. Third, velocity-compensated sliding window aggregation over k consecutive frames reaches R@100P = 100.00% at k = 5 for constant-velocity traversal and an additional 0.20 ms latency. Taken together, these findings show the impact of inference stage design decisions in STDP-based SNN-VPR on recall precision, although the separate contribution of each mechanism and implementation differences is only partially disentangled and needs further examination.
Problem

Research questions and friction points this paper is trying to address.

Visual Place Recognition
Spiking Neural Networks
STDP
Recall at 100% Precision
Autonomous Navigation
Innovation

Methods, ideas, or system contributions that make the work stand out.

Spiking Neural Networks
STDP learning
Visual Place Recognition
tensor-native implementation
sliding window aggregation
💼 Related Jobs
No related jobs found.
A
Altzi Tsanko
Department of Production and Management Engineering, Democritus University of Thrace, Xanthi 67100, Greece
O
Oikonomou Katerina Maria
Department of Production and Management Engineering, Democritus University of Thrace, Xanthi 67100, Greece
A
Antonios Gasteratos
Department of Production and Management Engineering, Democritus University of Thrace, Xanthi 67100, Greece