PiT-PO Boosts Equation Discovery with RL
PiT-PO uses reinforcement learning to evolve LLMs for symbolic regression, enforcing physical validity and parsimony. It treats LLMs as adaptive generators updated by search feedback. Achieves SOTA on benchmarks and discovers novel turbulence models.





