Codex reasoning-token clustering at 516 may be leading to degraded performance
github.comwonder if theyre basically doing this llamacpp reasoning trick: https://github.com/ggml-org/llama.cpp/blob/master/common/rea...
id guess its the harness
this is happening on xhigh config!! extremely pissing off