Variations on the Reinforcement Learning performance of Blackjack

08/09/2023
by   Avish Buramdoyal, et al.
0

Blackjack or "21" is a popular card-based game of chance and skill. The objective of the game is to win by obtaining a hand total higher than the dealer's without exceeding 21. The ideal blackjack strategy will maximize financial return in the long run while avoiding gambler's ruin. The stochastic environment and inherent reward structure of blackjack presents an appealing problem to better understand reinforcement learning agents in the presence of environment variations. Here we consider a q-learning solution for optimal play and investigate the rate of learning convergence of the algorithm as a function of deck size. A blackjack simulator allowing for universal blackjack rules is also implemented to demonstrate the extent to which a card counter perfectly using the basic strategy and hi-lo system can bring the house to bankruptcy and how environment variations impact this outcome. The novelty of our work is to place this conceptual understanding of the impact of deck size in the context of learning agent convergence.

READ FULL TEXT

page 5

page 6

page 7

page 9

page 10

research
02/15/2020

Deep RL Agent for a Real-Time Action Strategy Game

We introduce a reinforcement learning environment based on Heroic - Magi...
research
03/21/2020

FlapAI Bird: Training an Agent to Play Flappy Bird Using Reinforcement Learning Techniques

Reinforcement learning is one of the most popular approach for automated...
research
08/15/2020

Chrome Dino Run using Reinforcement Learning

Reinforcement Learning is one of the most advanced set of algorithms kno...
research
07/28/2022

Playing a 2D Game Indefinitely using NEAT and Reinforcement Learning

For over a decade now, robotics and the use of artificial agents have be...
research
08/25/2021

Adversary agent reinforcement learning for pursuit-evasion

A reinforcement learning environment with adversary agents is proposed i...
research
01/13/2022

Criticality-Based Varying Step-Number Algorithm for Reinforcement Learning

In the context of reinforcement learning we introduce the concept of cri...
research
05/23/2021

An Efficient Application of Neuroevolution for Competitive Multiagent Learning

Multiagent systems provide an ideal environment for the evaluation and a...

Please sign up or login with your details

Forgot password? Click here to reset