/v1/chat/completions endpoint enables your applications to display text tokens as soon as they are produced by the AI model, minimizing perceived latency for chat and terminal interfaces.
Event Stream Format
Each chunk arrives as an SSE event indata: <JSON> format:
Tracking Token Usage During Streams
Passstream_options to receive the final token accounting block in the trailing chunk: