Neuro-Symbolic Reinforcement Learning with First-Order Logic

Daiki Kimura,Masaki Ono,Subhajit Chaudhury,Ryosuke Kohita,Akifumi Wachi,Don Joven Agravante,Michiaki Tatsubori,Asim Munawar,Alexander G. Gray

Neuro-Symbolic Reinforcement Learning with First-Order Logic

2021

Daiki Kimura
Masaki Ono
Subhajit Chaudhury
Ryosuke Kohita
Akifumi Wachi
Don Joven Agravante
Michiaki Tatsubori
Asim Munawar
Alexander G. Gray

Deep reinforcement learning (RL) methods often require many trials before convergence, and no direct interpretability of trained policies is provided. In order to achieve fast convergence and interpretability for the policy in RL, we propose a novel RL method for text-based games with a recent neuro-symbolic framework called Logical Neural Network, which can learn symbolic and interpretable rules in their differentiable network. The method is first to extract first-order logical facts from text observation and external word meaning network (ConceptNet), then train a policy in the network with directly interpretable logical operators. Our experimental results show RL training with the proposed method converges significantly faster than other state-of-the-art neuro-symbolic methods in a TextWorld benchmark.

Keywords:

Artificial neural network
Artificial intelligence
Reinforcement learning
Differentiable function
word meaning
Interpretability
Computer science
First-order logic
Benchmark (computing)
Convergence (routing)

Correction
Source
Cite
Save
Machine Reading By IdeaReader

References

Citations