Clocked Playout Beats Packet-Driven Audio
Network packets arrive when the network allows; speakers need samples when the audio clock demands them.
Network packets arrive when the network allows; speakers need samples when the audio clock demands them.
A low average loss rate can still sound terrible when packets arrive in short bursts separated by long gaps.
Burst timing can hurt a speakerphone even when aggregate RTP loss is close to zero.
A packet can arrive at 09:17:25 and carry an RTP timestamp like 2873419200. Those values describe different things. Confusing them is one of the fastest ways to make real-time audio timing harder than it already is.