Streaming
AI in the stackSending a response in pieces as it is produced.
Delivering output progressively instead of waiting for all of it. Does not make anything faster; makes waiting tolerable, which is usually the real problem. Applies to slow page data as much as to model output.