Thank you. That last video seems like a pretty good refutation of the hand waving going on.
There are legitimate questions to be answered about audio compression and sampling. Did the video discuss samples of normal voice vs, screaming as well?
The way digital audio is handled could create serious issues, the baud rate of the phone, its compression, the transmission if it was VOIP or over a cordless and the digital storage all raise legitimate questions, they may have answers.
ETA: While one test address a bandwidth compression, it does nothing to address the issues of digital transmission compression, sampling over a cel phone, or the storage compression.
Nor does it show that a yelling voice is at all similar to a normal voice.
As stated before these would all be very easy to verify, very quickly, in two days at most. Given knowledge of the phone used, the placement of the phone vs. the yelling, the transmission method and the storage method.
It would really only take two days to run tests, have different people in that apartment complex yell, capture the voice recordings at the site using a similar phone at a similar distance, then run samples vs. their speaking voices. Given say twenty volunteers randomly chosen, a variety of yelling and screaming samples (ten at least per individual), then we could find error bars very quickly.