redzilla
All tools
Performance

Jitter Buffer

Enter jitter, packetization and latencies and I compute the recommended jitter-buffer depth, packets queued and total mouth-to-ear latency, rated with the G.114 threshold (≤150 ms good).

ms

Variation of the inter-packet delay. Usually reported as mean jitter or P95.

ms

Audio samples per packet. Typical 20 ms for VoIP; 33 ms ≈ 1 frame at 30 fps.

ms

One-way delay (half the RTT). Includes propagation, queuing and serialization.

ms

Algorithmic + processing delay. G.711 ≈ 1 ms; G.729 ≈ 25 ms.

depth = k × jitter minimum 2 packets
Examples
redzilla.cl — jitter
 
Buffer depth
Packets queued
Mouth-to-ear latency
Buffer
Packets
Mouth-to-ear

ITU-T G.114 rating

Mouth-to-ear delay breakdown

Componentms%
How it is computed · buffer, packets and G.114

1. Buffer depth: k × jitter, with a floor of 2 packets (2 × ptime). With k = 2 the buffer absorbs twice the observed jitter.

2. Packets queued: ceil(depth / ptime). The buffer fills with whole frames, so it rounds up.

3. The buffer adds its depth as fixed delay. Mouth-to-ear latency = network one-way + buffer + codec + ptime.

4. ITU-T G.114 (one-way delay): ≤ 150 ms good, 150–400 ms acceptable, > 400 ms poor for most voice applications.

Runs locally in your browser · no sign-up · nothing leaves your browser

Was this tool useful?
Disclaimer We take great care to keep every tool accurate and review it thoroughly; even so, we can't guarantee it is free of errors or take responsibility for how the results are used. We recommend double-checking anything critical.
Found an error? Let us know →