Pith. sign in

REVIEW 3 cited by

Semantics at an Angle: When Cosine Similarity Works Until It Doesn't

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2504.16318 v2 pith:V7XL7TQM submitted 2025-04-22 cs.LG

Semantics at an Angle: When Cosine Similarity Works Until It Doesn't

classification cs.LG
keywords cosinesimilarityembeddingslimitationswhenaddressadoptionalignment
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Cosine similarity has become a standard metric for comparing embeddings in modern machine learning. Its scale-invariance and alignment with model training objectives have contributed to its widespread adoption. However, recent studies have revealed important limitations, particularly when embedding norms carry meaningful semantic information. This informal article offers a reflective and selective examination of the evolution, strengths, and limitations of cosine similarity. We highlight why it performs well in many settings, where it tends to break down, and how emerging alternatives are beginning to address its blind spots. We hope to offer a mix of conceptual clarity and practical perspective, especially for quantitative scientists who think about embeddings not just as vectors, but as geometric and philosophical objects.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. HistoRAG: Embedding Historical Methodology in Retrieval-Augmented Generation Through Critical Technical Practice

    cs.CL 2026-06 unverdicted novelty 6.0

    HistoRAG embeds historiographical principles into RAG via temporal windowing, decoupled retrieval, and contestable LLM relevance judgments, evaluated on 102k Der Spiegel articles from 1950-1979.

  2. Same Patient, Different Words, Different Diagnosis? Evaluating Semantic Stability in Clinical LLMs

    cs.CL 2026-05 unverdicted novelty 6.0

    Domain specialization does not consistently improve clinical LLM robustness to meaning-preserving prompt variations, as shown by new sensitivity metrics on DiagnosisQA and MedQA.

  3. A Dual Cross-Attention Graph Learning Framework For Multimodal MRI-Based Major Depressive Disorder Detection

    cs.CV 2026-04 unverdicted novelty 4.0

    Dual cross-attention fusion of sMRI and rs-fMRI data achieves 84.71% accuracy in MDD detection on the REST-meta-MDD dataset, outperforming concatenation on functional atlases.