Imagine other people might like to know: if you use an API key to access OpenAI models in the Codex Desktop application, you can't access the faster inference.
I patched my codex application with this code, I would hope such a change could be upstreamed: https://github.com/adenta/codex-fast-mode-patcher