streaming — Sending a model's response token by token as it is generated, so the reader sees text immediately instead of waiting for the whole answer.
Why it matters
Speaker-attributed transcription that streams rather than batching, with hotword support and an MIT license, is deployable inside a product without a vendor API. Transformers integration is already in place.