Math solving pure reinforcement learning Is there a case where a math solver or theorem prover or even just an ai that does basic addition trainedā¦