ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference | Read Paper on Bytez