Addressing a fundamental limitation in deep vision models: lack of spatial attention | Read Paper on Bytez