Skip to content

Quantifying the Reality Gap for RL-Based UAV Placement at mmWave and Sub-THz

arXiv · arXiv preprint · 2609.11121 · Published 2026-09-10 · 4 authors

Abstract and citation only, verbatim from arXiv; full text lives there. All credit to the authors and arXiv.

Abstract

Reinforcement learning (RL) policies for unmanned aerial vehicle (UAV) placement in mmWave and sub-terahertz networks are typically trained on simplified analytical channels. We quantify the resulting sim-to-real gap on a real urban map of Doha, Qatar, at carriers {28, 140, 183, 300} GHz and altitudes {50, 75, 100, 125} m, evaluating three channel pipelines: an analytical model (FSPL + atmospheric absorption + cuboid LoS), full Monte-Carlo ray tracing in Sionna RT with ITU-R P.676-13 absorption, and a deterministic-LoS hybrid that reuses Sionna's mesh under a closed-form path-gain expression. We formalize the gap on the spatial SNR distribution via four metrics, namely bias, RMSE, Jensen-Shannon divergence, and optimum-deployment displacement. Three findings emerge: at 28/140 GHz, $\sim$70% of the apparent -5.6/-4.8 dB Sionna bias is Monte-Carlo undersampling and shrinks to -1.7/-1.5 dB after mitigation; at 183 GHz a -9.2 dB residual isolates the atmospheric absorption / ITU-R P.676 line-shape disagreement; at 300 GHz the stochastic ray tracer agrees with the analytical model only coincidentally, with a +3.8 dB structural offset exposed by the deterministic-LoS pipeline. Across all carriers the linear-domain regret of the analytical-trained policy stays $\geq$ 0.93, indicating practical near-optimality but with a carrier-resolved SNR bias that warrants explicit reporting.

Authors

  • Abdullateef Almohamad
  • Mostafa Ibrahim
  • Sabit Ekin
  • Khalid Qaraqe

Keywords

  • eess.SY

Citation

Abdullateef Almohamad, Mostafa Ibrahim, Sabit Ekin , et al. (2026). Quantifying the Reality Gap for RL-Based UAV Placement at mmWave and Sub-THz. arXiv ID 2609.11121. https://arxiv.org/abs/2609.11121 ↗