LA

Long-context attention compute

Topic

What experts have said about Long-context attention compute

1 statement · 1 negative

  1. Reiner PopeNegativeApr 29, 2026· Dwarkesh Podcast

    Attention’s context-dependent compute cost becomes noticeable at contexts of millions of tokens.

    You start to notice the effect of the quadratic or the linear term up in the millions of tokens or so.

    Listen at 1:52:24

    Open the episode · Reiner Pope – The math behind how LLMs are trained and served

Statements are attributed to the speaker as said on the episode and reflect their view at the time, not PodLume's. They are not advice.

What is PodLume?

PodLume turns podcasts into searchable knowledge. AI-decoded transcripts, identified guests and topics, smart highlights, and cross-show search across the world’s best conversations — all in your pocket.

Long-context attention compute | PodLume