Latest GPT-4 training data updated to December 2023
platform.openai.com
platform.openai.com
https://vue.ai/blog/ai-transformation/the-degeneration-of-ge...
We have truly hit the limit of names in this industry.
Ps I also tried gpt-4-turbo but the API rejects it for some reason. Same with the "numbered" ones listed in the documentation, I only seem to have access to the standard gpt-4 somehow.
> when were you last updated?
My last update was in April 2023.
The second step is fine-tuning the model on a much smaller set of annotated data which specify that it should actually "do something" in its responses and what it should do; it "teaches" it that it should actually answer the questions instead of e.g. continuing on with a list of more questions in the same vein, and most such training sets also "teach" it that for certain questions the appropriate response is a refusal.
If you have the original core model (before that instruction tuning) then you can repeat the same process but instead replace the instruction training set with a different one, so you can "instruct" the model to behave differently. Here is a nice and informative article from Eric Hartford about how he did that to make certain 'uncensored' models - https://erichartford.com/uncensored-models
If I ask Gemini, for example "Give me three important events that happened during May 13, year 2023.", then it says that this is "in the future" and responds that it can provide some guesses "Based on publicly available information from early February 2023", so that probably is the cut-off date for the model.
However, I would assume that (just like Bing) for certain questions it can pull in extra information from web searches - the model can be old, but if the system puts some retrieved document(s) in the prompt context so that the model can use them for generating the response, it can use that (limited) fresh information as well.
See this 25 day old screenshot of the docs which says knowledge cutoff April 2023 for the 0125 model
https://www.reddit.com/r/ChatGPT/comments/19fhb6h/new_gpt_4_...
This help article which was updated "this week" also still mentions April 2023
https://help.openai.com/en/articles/8555510-gpt-4-turbo-in-t...
And in the announcement for 0125 they didn't say anything about the knowledge cutoff change
https://openai.com/blog/new-embedding-models-and-api-updates
1) Replaced a pinned version with a new model (problematic for response consistency), or,
2) Decoupled knowledge cutoff from the model (how!?)
This has been a major pain point for me, because GPT-4 constantly uses @validator, "const", "always", etc. features that don't work on v2.
Also it's still more expensive to increase the context length
Both models were not able to have any pydantic v2 training data.