0 / 60 seg.

Well, evolutionarily that might make sense, because we know cognitively what we are going to sound like so maybe we don't need to spend energy analyzing the signal.