The Game of Hidden Rules: A New Kind of Benchmark Challenge for Machine Learning

07/20/2022
by   Eric Pulick, et al.
0

As machine learning (ML) is more tightly woven into society, it is imperative that we better characterize ML's strengths and limitations if we are to employ it responsibly. Existing benchmark environments for ML, such as board and video games, offer well-defined benchmarks for progress, but constituent tasks are often complex, and it is frequently unclear how task characteristics contribute to overall difficulty for the machine learner. Likewise, without a systematic assessment of how task characteristics influence difficulty, it is challenging to draw meaningful connections between performance in different benchmark environments. We introduce a novel benchmark environment that offers an enormous range of ML challenges and enables precise examination of how task elements influence practical difficulty. The tool frames learning tasks as a "board-clearing game," which we call the Game of Hidden Rules (GOHR). The environment comprises an expressive rule language and a captive server environment that can be installed locally. We propose a set of benchmark rule-learning tasks and plan to support a performance leader-board for researchers interested in attempting to learn our rules. GOHR complements existing environments by allowing fine, controlled modifications to tasks, enabling experimenters to better understand how each facet of a given learning task contributes to its practical difficulty for an arbitrary ML algorithm.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
12/21/2022

NADBenchmarks – a compilation of Benchmark Datasets for Machine Learning Tasks related to Natural Disasters

Climate change has increased the intensity, frequency, and duration of e...
research
10/08/2019

Can We Distinguish Machine Learning from Human Learning?

What makes a task relatively more or less difficult for a machine compar...
research
11/05/2021

Confidential Machine Learning Computation in Untrusted Environments: A Systems Security Perspective

As machine learning (ML) technologies and applications are rapidly chang...
research
06/30/2023

Comparing Reinforcement Learning and Human Learning using the Game of Hidden Rules

Reliable real-world deployment of reinforcement learning (RL) methods re...
research
05/03/2021

Unreasonable Effectiveness of Rule-Based Heuristics in Solving Russian SuperGLUE Tasks

Leader-boards like SuperGLUE are seen as important incentives for active...
research
10/03/2018

Procedural Puzzle Challenge Generation in Fujisan

Challenges for physical solitaire puzzle games are typically designed in...
research
01/22/2014

GGP with Advanced Reasoning and Board Knowledge Discovery

Quality of General Game Playing (GGP) matches suffers from slow state-sw...

Please sign up or login with your details

Forgot password? Click here to reset