It's crazy how much better you can make LLM output just by asking "is this the most elegant solution?" In a loop
(Not fine tuning, but interesting none the less. If a model can so easily find a more elegant solution, why didn't it pick that in the first place?)