Revisiting Cosine Similarity via Normalized ICA-transformed Embeddings
Fuente:
arXiv
Saved in:
| Main Authors: | Yamagiwa, Hiroaki, Oyama, Momose, Shimodaira, Hidetoshi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Understanding Higher-Order Correlations Among Semantic Components in Embeddings
by: Oyama, Momose, et al.
Published: (2024)
by: Oyama, Momose, et al.
Published: (2024)
Axis Tour: Word Tour Determines the Order of Axes in ICA-transformed Embeddings
by: Yamagiwa, Hiroaki, et al.
Published: (2024)
by: Yamagiwa, Hiroaki, et al.
Published: (2024)
Mapping 1,000+ Language Models via the Log-Likelihood Vector
by: Oyama, Momose, et al.
Published: (2025)
by: Oyama, Momose, et al.
Published: (2025)
Likelihood Variance as Text Importance for Resampling Texts to Map Language Models
by: Oyama, Momose, et al.
Published: (2025)
by: Oyama, Momose, et al.
Published: (2025)
Domain Mixture Design via Log-Likelihood Differences for Aligning Language Models with a Target Model
by: Kishino, Ryo, et al.
Published: (2026)
by: Kishino, Ryo, et al.
Published: (2026)
Norm of Mean Contextualized Embeddings Determines their Variance
by: Yamagiwa, Hiroaki, et al.
Published: (2024)
by: Yamagiwa, Hiroaki, et al.
Published: (2024)
Establishing a Scale for Kullback-Leibler Divergence in Language Models Across Various Settings
by: Kishino, Ryo, et al.
Published: (2025)
by: Kishino, Ryo, et al.
Published: (2025)
Language Model Maps for Prompt-Response Distributions via Log-Likelihood Vectors
by: Takase, Yusuke, et al.
Published: (2026)
by: Takase, Yusuke, et al.
Published: (2026)
Measuring Affinity between Attention-Head Weight Subspaces via the Projection Kernel
by: Yamagiwa, Hiroaki, et al.
Published: (2026)
by: Yamagiwa, Hiroaki, et al.
Published: (2026)
Shimo Lab at "Discharge Me!": Discharge Summarization by Prompt-Driven Concatenation of Electronic Health Record Sections
by: He, Yunzhen, et al.
Published: (2024)
by: He, Yunzhen, et al.
Published: (2024)
Predicting drug-gene relations via analogy tasks with word embeddings
by: Yamagiwa, Hiroaki, et al.
Published: (2024)
by: Yamagiwa, Hiroaki, et al.
Published: (2024)
Quantifying Lexical Semantic Shift via Unbalanced Optimal Transport
by: Kishino, Ryo, et al.
Published: (2024)
by: Kishino, Ryo, et al.
Published: (2024)
3D Rotation and Translation for Hyperbolic Knowledge Graph Embedding
by: Zhu, Yihua, et al.
Published: (2023)
by: Zhu, Yihua, et al.
Published: (2023)
Block-Diagonal Orthogonal Relation and Matrix Entity for Knowledge Graph Embedding
by: Zhu, Yihua, et al.
Published: (2024)
by: Zhu, Yihua, et al.
Published: (2024)
Knowledge Sanitization of Large Language Models
by: Ishibashi, Yoichi, et al.
Published: (2023)
by: Ishibashi, Yoichi, et al.
Published: (2023)
Zipfian Whitening
by: Yokoi, Sho, et al.
Published: (2024)
by: Yokoi, Sho, et al.
Published: (2024)
DeLTa: A Decoding Strategy based on Logit Trajectory Prediction Improves Factuality and Reasoning Ability
by: He, Yunzhen, et al.
Published: (2025)
by: He, Yunzhen, et al.
Published: (2025)
Beyond Chains: Bridging Large Language Models and Knowledge Bases in Complex Question Answering
by: Zhu, Yihua, et al.
Published: (2025)
by: Zhu, Yihua, et al.
Published: (2025)
Draft on the Fly: Adaptive Self-Speculative Decoding using Cosine Similarity
by: Metel, Michael R., et al.
Published: (2024)
by: Metel, Michael R., et al.
Published: (2024)
Combining Discrete Wavelet and Cosine Transforms for Efficient Sentence Embedding
by: Salama, Rana, et al.
Published: (2025)
by: Salama, Rana, et al.
Published: (2025)
Exploring Intra and Inter-language Consistency in Embeddings with ICA
by: Li, Rongzhi, et al.
Published: (2024)
by: Li, Rongzhi, et al.
Published: (2024)
Rethinking Layer Relevance in Large Language Models Beyond Cosine Similarity
by: Hinostroza, Cristian, et al.
Published: (2026)
by: Hinostroza, Cristian, et al.
Published: (2026)
Cosine-Similarity Routing with Semantic Anchors for Interpretable Mixture-of-Experts Language Models
by: Ternovtsii, Ivan, et al.
Published: (2025)
by: Ternovtsii, Ivan, et al.
Published: (2025)
Beyond Cosine Similarity: Zero-Initialized Residual Complex Projection for Aspect-Based Sentiment Analysis
by: Wang, Yijin, et al.
Published: (2026)
by: Wang, Yijin, et al.
Published: (2026)
Memorization, Emergence, and Explaining Reversal Failures: A Controlled Study of Relational Semantics in LLMs
by: Zhu, Yihua, et al.
Published: (2026)
by: Zhu, Yihua, et al.
Published: (2026)
Reasoning Depth and Environment Complexity: A Controlled Study of RLVR Data Allocation across Logical Reasoning Tasks
by: Zhu, Yihua, et al.
Published: (2026)
by: Zhu, Yihua, et al.
Published: (2026)
Efficient Prompt Caching via Embedding Similarity
by: Zhu, Hanlin, et al.
Published: (2024)
by: Zhu, Hanlin, et al.
Published: (2024)
Semantic Textual Similarity Assessment in Chest X-ray Reports Using a Domain-Specific Cosine-Based Metric
by: Picha, Sayeh Gholipour, et al.
Published: (2024)
by: Picha, Sayeh Gholipour, et al.
Published: (2024)
Beyond Cosine Similarity: Taming Semantic Drift and Antonym Intrusion in a 15-Million Node Turkish Synonym Graph
by: Tosun, Ebubekir, et al.
Published: (2026)
by: Tosun, Ebubekir, et al.
Published: (2026)
Revisiting Word Embeddings in the LLM Era
by: Mahajan, Yash, et al.
Published: (2025)
by: Mahajan, Yash, et al.
Published: (2025)
Improving Similar Case Retrieval Ranking Performance By Revisiting RankSVM
by: Liu, Yuqi, et al.
Published: (2025)
by: Liu, Yuqi, et al.
Published: (2025)
Revisiting Code Similarity Evaluation with Abstract Syntax Tree Edit Distance
by: Song, Yewei, et al.
Published: (2024)
by: Song, Yewei, et al.
Published: (2024)
SimReg: Achieving Higher Performance in the Pretraining via Embedding Similarity Regularization
by: Sun, Yan, et al.
Published: (2026)
by: Sun, Yan, et al.
Published: (2026)
Is Cosine-Similarity of Embeddings Really About Similarity?
by: Steck, Harald, et al.
Published: (2024)
by: Steck, Harald, et al.
Published: (2024)
Are ELECTRA's Sentence Embeddings Beyond Repair? The Case of Semantic Textual Similarity
by: Rep, Ivan, et al.
Published: (2024)
by: Rep, Ivan, et al.
Published: (2024)
Enhancing Semantic Similarity Understanding in Arabic NLP with Nested Embedding Learning
by: Nacar, Omer, et al.
Published: (2024)
by: Nacar, Omer, et al.
Published: (2024)
Mean-Pooled Cosine Similarity is Not Length-Invariant: Theory and Cross-Domain Evidence for a Length-Invariant Alternative
by: Mitra, Sibayan, et al.
Published: (2026)
by: Mitra, Sibayan, et al.
Published: (2026)
Estimating Text Similarity based on Semantic Concept Embeddings
by: der Brück, Tim vor, et al.
Published: (2024)
by: der Brück, Tim vor, et al.
Published: (2024)
Claim2Vec: Embedding Fact-Check Claims for Multilingual Similarity and Clustering
by: Panchendrarajan, Rrubaa, et al.
Published: (2026)
by: Panchendrarajan, Rrubaa, et al.
Published: (2026)
CASE -- Condition-Aware Sentence Embeddings for Conditional Semantic Textual Similarity Measurement
by: Zhang, Gaifan, et al.
Published: (2025)
by: Zhang, Gaifan, et al.
Published: (2025)
Similar Items
-
Understanding Higher-Order Correlations Among Semantic Components in Embeddings
by: Oyama, Momose, et al.
Published: (2024) -
Axis Tour: Word Tour Determines the Order of Axes in ICA-transformed Embeddings
by: Yamagiwa, Hiroaki, et al.
Published: (2024) -
Mapping 1,000+ Language Models via the Log-Likelihood Vector
by: Oyama, Momose, et al.
Published: (2025) -
Likelihood Variance as Text Importance for Resampling Texts to Map Language Models
by: Oyama, Momose, et al.
Published: (2025) -
Domain Mixture Design via Log-Likelihood Differences for Aligning Language Models with a Target Model
by: Kishino, Ryo, et al.
Published: (2026)