Castling-ViT: Compressing Self-Attention Via Switching Towards Linear-Angular Attention at Vision Transformer Inference | AMiner