Refutation of Deep Reinforcement Learning in Hol4
Colin James · viXra · 2019
We evaluate rewriting terms as based on axioms in a subset of Robinson arithmetic on which deep reinforcement learning in HOL4 relies: two such axioms as not tautologous. These results form a non tautologous fragment of the universal logic VŁ4.