I agree that the connection to OpenAI is very speculative here. It's true that Pangram is also very unreliable, and the match is pretty weak anyways: "69% came back flagged as fully AI-generated, with another 28% flagged as partially AI-generated".
But the analysis of the internal admin pages for generating articles, the API and the interview structure are pretty damming evidence that this is largely driven by AI.