Doesn’t OpenAI embedding model support 8191/8192 tokens? That aside, declaring a winner by token size is misleading. There are more important factors like cross language support and precision for example
stella_en_1.5B_v5 seems to be an unsung hero model in that regard
plus you may not even want such large token sizes if you just need accurate retrieval of snippets of text (like 1-2 sentences)