Am I wrong in saying that the interfaces presented to the model in OMP versus plain old Pi are identical?
it's less about the interfaces and more that it feels disingenuous to compare token usage to a pi setup that injects a bunch of stuff into context rather than the base version. it makes me think they tested it with pi, found they couldnt meaningfully beat its token usage + pass rate, and omitted the comparison. i'd love to be proven wrong, but this seems like the easy explanation
OMP has a lot of candy that raises token cost compared to vanilla pi