Important to note that lower model + higher reasoning gives a different (not higher) quality of response than higher model + lower reasoning.
Some tasks are reasoning shaped by nature and you can't just throw a big model at it.