You build an AI chat application.

A user sends:

"Explain how distributed systems work."

Your server calls an LLM API and starts streaming the answer: