The Proximal Policy Optimization (PPO) and Advantage Actor-Critic (A2C) methods have proven effective in detecting, evaluating, and performing multi-class classification of cyberattacks. This study implements both algorithms within a Reinforcement Learning framework to classify types of attacks based on network traffic features. The datasets used for testing include CIC-IDS2018, CIC-IDS2017, IS…