A traffic-signal simulator and multi-agent RL stack: one small shared policy learns to run every intersection.
Description excerpted from the original listing, which is linked below.