OptionalDisplayThe participant's display name at capture time, when known.
The LiveKit participant identity this audio came from (the diarization speaker label).
Raw PCM audio bytes for this frame.
OptionalTimestampOptional epoch-ms capture timestamp.
One frame of raw per-participant audio the seam surfaces for diarization + the agent's "hearing". Because LiveKit subscribes tracks per participant, every inbound frame is already attributed to a single speaker — diarization comes free from the SFU.