← LLM context length
What podcasts say about LLM context length
Every statement, with the speaker, the exact quote and the moment it was said.
What experts have said about LLM context length
2 statements · 1 negative · 1 mixed
Longer contexts can shift inference from compute-limited to memory-limited operation.
“as you vary the context length, the KBFetchtime will go up and up, and so that'll cause a transition from compute limited to memory limited”
Open the episode · Reiner Pope – The math behind how LLMs are trained and servedListen at 11:17
LLM context windows may reach 2–5 million tokens in 2026, but not 100 million.
“I would expect it to keep increasing and get to 2 million or 5 million this year. But I don't expect it to go to 100 million.”
Open the episode · #490 – State of AI in 2026: LLMs, Coding, Scaling Laws, China, Agents, GPUs, AGIListen at 2:59:33
Statements are attributed to the speaker as said on the episode and reflect their view at the time, not PodLume's. They are not advice.