The thing is that the longer the context window is the worse retrieval accuracy is as well. For GPT 5.4 that was released last week, you can see that accuracy is ~97% up to 32k tokens, but drops to ~36% when over 512k tokens. You can fit a million tokens, but GPT will not use it effectively.
0 likes 2 replies
?