GaVA-CLIP:Refining Multimodal Representations With Clinical Knowledge and Numerical Parameters for Gait Video