Reinforcement learning produces dominant strategies for the Iterated Prisoner's Dilemma.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 29228001.
- Also identified by DOI 10.1371/journal.pone.0188046 and PMC identifier 5724862.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
We present tournament results and several powerful strategies for the Iterated Prisoner's Dilemma created using reinforcement learning techniques (evolutionary and particle swarm algorithms). These strategies are trained to perform well against a corpus of over 170 distinct opponents, including many well-known and classic strategies. All the trained strategies win standard tournaments against the total collection of other opponents. The trained strategies and one particular human made designed strategy are the top performers in noisy tournaments also.
Medical subject headings
- Learning
- Prisoner Dilemma