גלעד כץ

אקדמי בכיר

Cost effective transfer of reinforcement learning policies

Many challenging real-world problems require the deployment of ensembles – multiple complementary learning models – to reach acceptable performance levels. While effective, applying the entire ensemble to every sample is costly and often unnecessary. Deep Reinforcement Learning (DRL) offers a cost-effective alternative, where detectors are dynamically chosen based on the output of their predecessors, with their usefulness weighted against their computational cost. Despite their potential, DRL-based solutions are not widely used in ensemble management. This can be attributed to the difficulties in configuring the reward function for each new task, the unpredictable reactions of the DRL agent to changes in the data, and the inability to use common performance metrics (e.g., True and False-Positive Rates, TPR/FPR) to guide the DRL model in a multi-objective environment. In this study, we propose methods for fine-tuning and calibrating DRL-based policies to meet multiple performance goals. Moreover, we present a method for transferring effective security policies from one dataset to another. Finally, we demonstrate that our approach is highly robust against adversarial attacks.

שפת פרסום אנגלית
כתב עת Expert Systems with Applications
כרך 237
סטטוס פרסום פורסם - 01.03.2024
121380

Keywords

Deep reinforcement learning
Ensemble learning
Machine learning

ASJC Scopus subject areas

General Engineering
Computer Science Applications
Artificial Intelligence
גישה למסמך
10.1016/j.eswa.2023.121380
קבצים וקישורים אחרים
Link to publication in Scopus