Skip to main navigation Skip to search Skip to main content

The Data-Schema Bottleneck: Benchmarking 16 Deep Learning Architectures for Real-Time Starlink LEO Telemetry Management

  • Tanmoy Debnath
  • , Sourabhi Debnath
  • , Miroslaw Narbutt
  • , Maumita Bhattacharya

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

Abstract

Effective management of Low Earth Orbit (LEO) satellite networks depends on data pipelines capable of predicting link-layer quality metrics (e.g., round-trip time (RTT), throughput (TP), and jitter etc.) at timescales suitable for real-time distributed data routing, handover, and congestion management. This study systematically benchmarks sixteen deep learning (DL) architectures representing four inductive-bias families to evaluate two formally stated hypotheses using 30 days of real Starlink telemetry comprising 417 million observations. Firstly, the Spectral Alignment Hypothesis (Research Question (RQ) 1) investigates whether architectures possessing inductive biases that explicitly decompose the quasi-periodic orbital structure of LEO dynamics systematically outperform those processing the telemetry as generic sequential data. Secondly, the Predictability Ceiling Hypothesis (RQ2) posits the existence of an architecture-invariant empirical upper bound on the variance explainable from end-to-end temporal history, quantifying its implications for AI-assisted data management. Empirical analysis demonstrates bounded support for RQ1: frequency-aware models achieve the highest RTT Coefficient of Determination (R2) and optimal Mean Absolute Error (MAE) at sub-100K parameters. However, their absolute predictive advantage over modern Transformers remains marginal (ΔR2 = 0.003 on RTT), with both paradigms proving indistinguishable regarding stochastic jitter. Consequently, RQ2 emerges as the primary contribution: predictive fidelity severely asymptotes at R2 ≈0.54 for propagation metrics (RTT and throughput) and saturates at R2 ≈0.31 for jitter across all sixteen structurally diverse models. This ceiling defines the maximum diagnostic gain achievable by the sequence predictors under investigation, operating exclusively on temporal 30 day LENS telemetry. We conclude that this plateau represents a fundamental informational limit of the data schema itself rather than an algorithmic deficiency, establishing that future AI-augmented edge databases should transition toward multi-modal feature fusion to breach this boundary.

Original languageEnglish
Title of host publicationProceedings of the 9th International Workshop on Artificial Intelligence Techniques for Data Management, aiDM 2026
PublisherAssociation for Computing Machinery (ACM)
Pages41-53
Number of pages13
ISBN (Electronic)9798400727191
DOIs
Publication statusPublished - 16 Jul 2026
Event9th International Workshop on Artificial Intelligence Techniques for Data Management, aiDM 2026 - Bengaluru, India
Duration: 31 May 20265 Jun 2026

Publication series

NameProceedings of the 9th International Workshop on Artificial Intelligence Techniques for Data Management, aiDM 2026

Conference

Conference9th International Workshop on Artificial Intelligence Techniques for Data Management, aiDM 2026
Country/TerritoryIndia
CityBengaluru
Period31/05/265/06/26

Keywords

  • Artificial Intelligence
  • Deep Learning
  • LEO Satellites Telemetry
  • Link Quality Prediction
  • Starlink
  • Time series Analysis

Fingerprint

Dive into the research topics of 'The Data-Schema Bottleneck: Benchmarking 16 Deep Learning Architectures for Real-Time Starlink LEO Telemetry Management'. Together they form a unique fingerprint.

Cite this