Please wait...

The amount of text (measured in tokens) a model can process in a single interaction. Context window size has exploded from thousands to millions of tokens in 2026, fundamentally changing what's possible. Models can now ingest entire codebases, legal document archives, or hours of transcripts in one go. Llama 4 Scout leads with a 10-million-token window, while Gemini 3 Pro offers 1 million tokens natively integrated into Google Workspace.