Glossary · AI
Streaming (SSE)
Sending an LLM response incrementally, token by token, as it is generated.
What is Streaming (SSE)?
Streaming delivers a model's output progressively — often over Server-Sent Events (SSE) or a WebSocket — so users see text appear in real time.
Because ad injection operates on this token stream, streaming is where in-conversation monetization happens.
Streaming (SSE) in context
To understand Streaming (SSE) fully, it helps to know the concepts around it. Token Injection, inserting ad content into an LLM's token stream as it is generated. Large Language Model (LLM), a neural network trained on large text corpora to generate and understand language. gRPC, a high-performance RPC framework using protocol buffers, often over HTTP/2.
Together these describe how ai works in practice for an AI app, and where Streaming (SSE) fits among them.
Frequently asked questions
Start monetizing in about 5 minutes
Wrap your existing LLM response stream with the Monetzly SDK and earn on every session, no paywall required.