OpenAI Codex, HumanEval, and the Leap from Language Modeling to Code Generation
OpenAI Codex adapted GPT-style language modeling to public source code and introduced HumanEval as a way to measure whether generated functions actually passed tests rather than merely looked plausible.