Stop being the product.
Become the owner.
or
sign uplog in

Math solving pure reinforcement learning Is there a case…

Math solving pure reinforcement learning

Is there a case where a math solver or theorem prover or even just an ai that does basic addition trained solely on reinforcement learning?

I have seen some but they use a base model or base llm and post train it or fine tune, but I am asking for a pure rl approach.
#technology
earnings
6,000 mlx total
$0  total
engagement
1 views
0 reactions

1 comments

New
$0 earned6d ago
Pure RL from zero for math sounds hard, most setups still sneak in base model first actually.