arXiv Science⌕ Search

arXiv · 2012.00307

Edge Deep Learning for Neural Implants

Abstract

Implanted devices providing real-time neural activity classification and control are increasingly used to treat neurological disorders, such as epilepsy and Parkinson's disease. Classification performance is critical to identifying brain states appropriate for the therapeutic action. However, advanced algorithms that have shown promise in offline studies, in particular deep learning (DL) methods, have not been deployed on resource-restrained neural implants. Here, we designed and optimized three embedded DL models of commonly adopted architectures and evaluated their inference performance in a case study of seizure detection. A deep neural network (DNN), a convolutional neural network (CNN), and a long short-term memory (LSTM) network were designed to classify ictal, preictal, and interictal phases from the CHB-MIT scalp EEG database. After iterative model compression and quantization, the algorithms were deployed on a general-purpose, off-the-shelf microcontroller. Inference sensitivity, false positive rate, execution time, memory size, and power consumption were quantified. For seizure event detection, the sensitivity and FPR (h-1) for the DNN, CNN, and LSTM models were 87.36%/0.169, 96.70%/0.102, and 97.61%/0.071, respectively. Predicting seizures for early warnings was also feasible. The implemented compression and quantization achieved a significant saving of power and memory with an accuracy degradation of less than 0.5%. Edge DL models achieved performance comparable to many prior implementations that had no time or computational resource limitations. Generic microcontrollers can provide the required memory and computational resources, while model designs can be migrated to ASICs for further optimization. The results suggest that edge DL inference is a feasible option for future neural implants to improve classification performance and therapeutic outcomes.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Xilin Liu, Andrew G. Richardson. 2021-06-01. Edge Deep Learning for Neural Implants. https://doi.org/10.1088/1741-2552%2Fabf473

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Scalable Long-Term Beamforming for Massive Multi-User MIMO

Fully digital massive multiple-input multiple-output (MIMO) systems with large numbers (1000+) of antennas offer capacity gains from spatial multiplexing and beamforming, but receivers that scale to these array dimensions face challenges in both channel estimation overhead and digital computation. Long-term beamforming addresses both by projecting the data onto a low-dimensional subspace that can be tracked at a slow time scale from the long-term channel parameters. In this setting, we show how to compute, in closed form, the projection matrix that maximizes a capacity upper bound, using a matrix inverse square root; the same projection is shown to maximize the mean post-projection signal-to-interference-plus-noise ratio (SINR) exactly. Computationally efficient methods are then presented for the matrix computation, realizable with matrix-matrix multiplies and hence amenable to systolic array implementations in hardware. Bounds on the SINR degradation are derived, and ray tracing simulations in a realistic rural uplink setting show a small loss relative to instantaneous minimum mean-square error (MMSE) beamforming when the covariance is accurately estimated. The efficient Gram-domain form of the instantaneous MMSE receiver applies the maximum-ratio combining reduction before an inverse whose dimension is the total number of streams. Against this baseline, the method estimates and refreshes beamforming coefficients three orders of magnitude less often and decouples the real-time path across users. With the conjugate-gradient solve and a rank-one projection, its total arithmetic cost is 2% higher at the ten-user operating point and lower above a stream-dimension crossover that we characterize.

eess.SP↗

Curv-Tail: Lightweight Long-Tailed Encrypted Traffic Classification with Discrete Packet-Length Encoding and Lorentz Prototypes

Long-tailed encrypted traffic classification requires accurate recognition of infrequent classes under limited computational budgets. We propose Curv-Tail, a lightweight, end-to-end packet--byte framework trained without a separate pretraining stage. Mixed-resolution tokenization preserves exact packet-length identities within a bounded range and coarsens larger values to limit the vocabulary. An auxiliary objective predicts observed length tokens from contextual packet features before pooling, encouraging length-token retention beyond flow-level supervision. Compact temporal encoders process packet sequences and directional byte patches, and Lorentz prototypes with a shared learnable curvature magnitude classify their fused representation. On NUDT-Mobile and DataCon-Website under natural class frequencies, Curv-Tail achieves three-seed mean Tail-F1 scores of 85.70% and 46.30%, exceeding the strongest evaluated baselines by 2.91 and 1.89 percentage points, respectively. In 300-class profiling on an RTX 4090 with FP32 and batch size 256, Curv-Tail uses 98.48% fewer parameters and achieves 10.2 times the batch inference throughput of MM4Flow.

eess.SP↗

Airborne Particle Communication Through Time-varying Diffusion-Advection Channels

Particle-based communication using diffusion and advection has emerged as an alternative signaling paradigm recently. While most existing studies assume constant flow conditions, real macro-scale environments such as atmospheric winds exhibit time-varying behavior. In this work, airborne particle communication under time-varying advection is modeled as a linear time-varying (LTV) channel, and a closed-form, time-dependent channel impulse response is derived using the method of moving frames. Based on this formulation, the channel is characterized through its power delay profile, leading to the definition of channel dispersion time as a physically meaningful measure of channel memory and a guideline for symbol duration selection. System-level simulations under directed, time-varying wind conditions show that waveform design is critical for performance, enabling multi-symbol modulation using a single particle type when dispersion is sufficiently controlled. To quantify waveform distortion and guide the design of orthogonal signaling waveforms, the Orthogonality Loss Ratio (OLR) is introduced as a structural metric. The results demonstrate that time-varying diffusion-advection channels can be systematically modeled and engineered using communication-theoretic tools, providing a realistic foundation for particle-based communication in complex flow environments.

eess.SP↗