excellent article, I will happily follow it but the intro caught me well, it reminded me a bit of Shannon's Bandwagon
https://reimbar.org/papers/bandwagon/ ; in computational neuroscience the problem setup is quite diverse around literature, with some "big" normative theories kind of accepted, but very far imho from that concepts and derivations like the channel capacity theorem become actually useful in terms of expalantory power or technology, they remain some niche work or some left away approaches in general. Reminded me also of last Tishby's work about explanations of deep learning training dynamics using the information bottleneck method. I ignore though and wonder if approaches like this are used as train targets on LLMs or something or they are just used to philosophize around (which I actually enjoy)