The ability of AI to understand and interpret images and videos. Used in India for crop disease detection (Fasal), medical imaging (Niramai), traffic management, and Aadhaar face verification.
The maximum amount of text an LLM can 'remember' in a single conversation, measured in tokens. Leading models now regularly support context windows from a few hundred thousand to over a million tokens: a larger window lets the model reference longer documents and conversations, but check each provider's current published limit rather than relying on old figures.