Streaming speaker diarization identifies who is speaking during live audio in real time, assigning speaker labels like SPEAKER_A and SPEAKER_B within milliseconds as conversations unfold. Unlike batch processing that waits for complete recordings, this technology enables voice agents and live meeting transcription to act on speaker identity immediately, though it trades some accuracy for speed since decisions become permanent once made.