this post was submitted on 16 Jan 2025
275 points (93.7% liked)

memes

21084 readers
3296 users here now

Community rules

1. Be civilNo trolling, bigotry or other insulting / annoying behaviour

2. No politicsThis is non-politics community. For political memes please go to !politicalmemes@lemmy.world

3. No recent repostsCheck for reposts when posting a meme, you can only repost after 1 month

4. No botsNo bots without the express approval of the mods or the admins

5. No Spam/Ads/AI SlopNo advertisements or spam. This is an instance rule and the only way to live. We also consider AI slop to be spam in this community and is subject to removal.

A collection of some classic Lemmy memes for your enjoyment

Sister communities

founded 2 years ago
MODERATORS
 
you are viewing a single comment's thread
view the rest of the comments
[–] brucethemoose@lemmy.world 5 points 1 year ago* (last edited 1 year ago)

The context window is indeed the LLM's memory.

...But its also muddy.

Many LLMs get 'dumber' and less attentive as their context windows grow, and OpenAI's models just happen to be one of these. It's awful close to the full 128K, even with the full GPT-4. Mistral models are also really bad at long context understanding while, conversely, I find that Google Gemini and Qwen 2.5 are really good close to their limits.

There are attempts to try and measure this performance objectively, like: https://github.com/NVIDIA/RULER