If I have to take the risk of simplifying,
1. We humans have managed to take huge amount of information and compress it using a loss function containing some bias we have about the information.
2. We now ask ourselves to decompress the same information with some additional cross-entropy. As a side effect of this process we sometimes spurt out information that may or may not have any meaning since the compression was lossy.
3. Now, we ask ourselves to present this some-what newly decompressed information with brevity in order to understand what we've learned from it.
Knowing that this process is happening on a larger scale, this resurfaces the argument if meaning can be reduced to computation only.
Although some might favor this argument but we are at the risk of anthropomorphizing this process.
The idea presented in the post itself is perspicuous (in Grant Sanderson own words) as he always does.