首页 /研究 /Fundamental Limits of Man-in-the-Middle Attack Detection in Model-Free Reinforcement Learning

LEARNING

Fundamental Limits of Man-in-the-Middle Attack Detection in Model-Free Reinforcement Learning

Rishi Rani, Massimo Franceschetti

发表年份: 2026
访问权限: 开放获取

摘要

We consider the problem of learning-based man-in-the-middle (MITM) attacks in cyber-physical systems (CPS), and extend our previously proposed Bellman Deviation Detection (BDD) framework for model-free reinforcement learning (RL). We refine the standard MDP attack model by allowing the reward function to depend on both the current and subsequent states, thereby capturing reward variations induced by errors in the adversary's transition estimate. We also derive an optimal system-identification strategy for the adversary that minimizes detectable value deviations. Further, we prove that the agent's asymptotic learning time required to secure the system scales linearly with the adversary's learning time, and that this matches the optimal lower bound. Hence, the proposed detection scheme is order-optimal in detection efficiency. Finally, we extend the framework to asynchronous and intermittent attack scenarios, where reliable detection is preserved.

关键词

eess.SYcs.LG

Fundamental Limits of Man-in-the-Middle Attack Detection in Model-Free Reinforcement Learning

摘要

关键词

相关论文

The Organization of Behavior

Fractional Brownian Motions, Fractional Noises and Applications

Review of deep learning: concepts, CNN architectures, challenges, applications, future directions

A guide to deep learning in healthcare