SourceThe "Law'' of the Unconscious Contrastive Learner: Probabilistic Alignment of Unpaired Modalities
The paper tests assumption 3 (that contrastive representations are uniform on the unit hypersphere) on real-world data. It computes language representations for each model over the AudioSet ontology and performs a two-sample Kolmogorov-Smirnov test against a uniform hypersphere distribution. CLIP yields a p-value of 0.0877 and CLAP a p-value of 0.1788, both above the conventional 0.05 significance threshold, so neither model's representations show significant deviation from uniformity.