A python script to train a dqn network on a traffic simulator using sumo. The individual traffic lights have their own agent and they learn using dqn. I am planning to either implement QMIX algorithm or some PPO variant.
sdhrt/sumo_marl
Folders and files
| Name | Name | Last commit date | ||
|---|---|---|---|---|