Skip to content
GCC AI Research

Search

Results for "Proximal Policy Optimization"

Reinforcement Learning-Based Traffic Signal Control for IoT-Enabled Intersections

arXiv ·

Researchers investigated reinforcement learning (RL) for adaptive traffic signal control at an urban intersection in Kuwait, aiming to mitigate urban traffic congestion. They developed a Proximal Policy Optimization (PPO)-based controller that dynamically adjusts green-phase durations using local traffic states in a realistic simulation environment informed by real-world Kuwaiti traffic data. The controller reduced average vehicle delay by 46% relative to fixed-time control and 34% relative to actuated control, while also lowering per-vehicle CO2 emissions by approximately 23%. Why it matters: This demonstrates a practical, learning-based edge traffic signal control solution for IoT-enabled smart city transportation systems, offering significant improvements in traffic flow and environmental impact for car-dependent cities in the Middle East.

Reinforcement learning-based dynamic cleaning scheduling framework for solar energy system

arXiv ·

This study introduces a reinforcement learning (RL) framework using Proximal Policy Optimization (PPO) and Soft Actor-Critic (SAC) to optimize the cleaning schedules of photovoltaic panels in arid regions. Applied to a case study in Abu Dhabi, the PPO-based framework demonstrated up to 13% cost savings compared to simulation optimization methods by dynamically adjusting cleaning intervals based on environmental conditions. The research highlights the potential of RL in enhancing the efficiency and reducing the operational costs of solar power generation.