Semantic Convergence: Investigating Shared Representations Across Scaled LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Son, Daniel, Rathore, Sanjana, Rufail, Andrew, Simon, Adrian, Zhang, Daniel, Dave, Soham, Blondin, Cole, Zhu, Kevin, O'Brien, Sean |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Causal Dimensionality of Transformer Representations: Measurement, Scaling, and Layer Structure
von: Sarkar, Nilesh, et al.
Veröffentlicht: (2026)
von: Sarkar, Nilesh, et al.
Veröffentlicht: (2026)
Do LLMs Truly Understand When a Precedent Is Overruled?
von: Zhang, Li, et al.
Veröffentlicht: (2025)
von: Zhang, Li, et al.
Veröffentlicht: (2025)
Shift-Reduce Task-Oriented Semantic Parsing with Stack-Transformers
von: Fernández-González, Daniel
Veröffentlicht: (2022)
von: Fernández-González, Daniel
Veröffentlicht: (2022)
Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free
von: Zhang, Li, et al.
Veröffentlicht: (2026)
von: Zhang, Li, et al.
Veröffentlicht: (2026)
Thinking Longer, Not Always Smarter: Evaluating LLM Capabilities in Hierarchical Legal Reasoning
von: Zhang, Li, et al.
Veröffentlicht: (2025)
von: Zhang, Li, et al.
Veröffentlicht: (2025)
T-Norm Operators for EU AI Act Compliance Classification: An Empirical Comparison of Lukasiewicz, Product, and Gödel Semantics in a Neuro-Symbolic Reasoning System
von: Laabs, Adam
Veröffentlicht: (2026)
von: Laabs, Adam
Veröffentlicht: (2026)
Approaches to Semantic Textual Similarity in Slovak Language: From Algorithms to Transformers
von: Radosky, Lukas, et al.
Veröffentlicht: (2026)
von: Radosky, Lukas, et al.
Veröffentlicht: (2026)
How much do LLMs learn from negative examples?
von: Hamdan, Shadi, et al.
Veröffentlicht: (2025)
von: Hamdan, Shadi, et al.
Veröffentlicht: (2025)
Communicative Agents for Slideshow Storytelling Video Generation based on LLMs
von: Fan, Jingxing, et al.
Veröffentlicht: (2025)
von: Fan, Jingxing, et al.
Veröffentlicht: (2025)
Towards Ontology-Enhanced Representation Learning for Large Language Models
von: Ronzano, Francesco, et al.
Veröffentlicht: (2024)
von: Ronzano, Francesco, et al.
Veröffentlicht: (2024)
PairCFR: Enhancing Model Training on Paired Counterfactually Augmented Data through Contrastive Learning
von: Qiu, Xiaoqi, et al.
Veröffentlicht: (2024)
von: Qiu, Xiaoqi, et al.
Veröffentlicht: (2024)
Causally Grounded Mechanistic Interpretability for LLMs with Faithful Natural-Language Explanations
von: Mahale, Ajay Pravin
Veröffentlicht: (2026)
von: Mahale, Ajay Pravin
Veröffentlicht: (2026)
Beyond Accuracy: Decomposing the Reasoning Efficiency of LLMs
von: Kaiser, Daniel, et al.
Veröffentlicht: (2026)
von: Kaiser, Daniel, et al.
Veröffentlicht: (2026)
Reference-Guided Verdict: LLMs-as-Judges in Automatic Evaluation of Free-Form QA
von: Badshah, Sher, et al.
Veröffentlicht: (2024)
von: Badshah, Sher, et al.
Veröffentlicht: (2024)
When Does Content-Based Routing Work? Representation Requirements for Selective Attention in Hybrid Sequence Models
von: Basu, Abhinaba
Veröffentlicht: (2026)
von: Basu, Abhinaba
Veröffentlicht: (2026)
Word Overuse and Alignment in Large Language Models: The Influence of Learning from Human Feedback
von: Juzek, Tom S., et al.
Veröffentlicht: (2025)
von: Juzek, Tom S., et al.
Veröffentlicht: (2025)
Detecting Sleeper Agents in Large Language Models via Semantic Drift Analysis
von: Zanbaghi, Shahin, et al.
Veröffentlicht: (2025)
von: Zanbaghi, Shahin, et al.
Veröffentlicht: (2025)
Multilingual Multi-Label Emotion Classification at Scale with Synthetic Data
von: Borisov, Vadim
Veröffentlicht: (2026)
von: Borisov, Vadim
Veröffentlicht: (2026)
IFMTBench: A Comprehensive Benchmark for Multilingual Translation Instruction Following
von: Sun, Mingrui, et al.
Veröffentlicht: (2026)
von: Sun, Mingrui, et al.
Veröffentlicht: (2026)
Carefully Structured Compression: Efficiently Managing StarCraft II Data
von: Ferenczi, Bryce, et al.
Veröffentlicht: (2024)
von: Ferenczi, Bryce, et al.
Veröffentlicht: (2024)
Fine-tuning of Large Language Models for Constituency Parsing Using a Sequence to Sequence Approach
von: Delgado, Francisco Jose Cortes, et al.
Veröffentlicht: (2025)
von: Delgado, Francisco Jose Cortes, et al.
Veröffentlicht: (2025)
Doğal Dil İşlemede Tokenizasyon Standartları ve Ölçümü: Türkçe Üzerinden Büyük Dil Modellerinin Karşılaştırmalı Analizi
von: Bayram, M. Ali, et al.
Veröffentlicht: (2025)
von: Bayram, M. Ali, et al.
Veröffentlicht: (2025)
Büyük Dil Modelleri için TR-MMLU Benchmarkı: Performans Değerlendirmesi, Zorluklar ve İyileştirme Fırsatları
von: Bayram, M. Ali, et al.
Veröffentlicht: (2025)
von: Bayram, M. Ali, et al.
Veröffentlicht: (2025)
Targeted Lexical Injection: Unlocking Latent Cross-Lingual Alignment in Lugha-Llama via Early-Layer LoRA Fine-Tuning
von: Ngugi, Stanley
Veröffentlicht: (2025)
von: Ngugi, Stanley
Veröffentlicht: (2025)
How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework
von: Nieth, Björn, et al.
Veröffentlicht: (2026)
von: Nieth, Björn, et al.
Veröffentlicht: (2026)
Beyond Subtokens: A Rich Character Embedding for Low-resource and Morphologically Complex Languages
von: Schneider, Felix, et al.
Veröffentlicht: (2026)
von: Schneider, Felix, et al.
Veröffentlicht: (2026)
A Study into Investigating Temporal Robustness of LLMs
von: Wallat, Jonas, et al.
Veröffentlicht: (2025)
von: Wallat, Jonas, et al.
Veröffentlicht: (2025)
The Concept Allocation Zone: Tracking How Concepts Form Across Transformer Depth
von: Henry, James
Veröffentlicht: (2026)
von: Henry, James
Veröffentlicht: (2026)
Enhancing Mathematical Problem Solving in LLMs through Execution-Driven Reasoning Augmentation
von: Basarkar, Aditya, et al.
Veröffentlicht: (2026)
von: Basarkar, Aditya, et al.
Veröffentlicht: (2026)
Constitution or Collapse? Exploring Constitutional AI with Llama 3-8B
von: Zhang, Xue
Veröffentlicht: (2025)
von: Zhang, Xue
Veröffentlicht: (2025)
Learning What Matters: Probabilistic Task Selection via Mutual Information for Model Finetuning
von: Chanda, Prateek, et al.
Veröffentlicht: (2025)
von: Chanda, Prateek, et al.
Veröffentlicht: (2025)
XAutoLM: Efficient Fine-Tuning of Language Models via Meta-Learning and AutoML
von: Estevanell-Valladares, Ernesto L., et al.
Veröffentlicht: (2025)
von: Estevanell-Valladares, Ernesto L., et al.
Veröffentlicht: (2025)
Efficiently Scanning and Resampling Spatio-Temporal Tasks with Irregular Observations
von: Ferenczi, Bryce, et al.
Veröffentlicht: (2024)
von: Ferenczi, Bryce, et al.
Veröffentlicht: (2024)
PCA- and SVM-Grad-CAM for Convolutional Neural Networks: Closed-form Jacobian Expression
von: Omae, Yuto
Veröffentlicht: (2025)
von: Omae, Yuto
Veröffentlicht: (2025)
Transactional Attention: Semantic Sponsorship for KV-Cache Retention
von: Basu, Abhinaba
Veröffentlicht: (2026)
von: Basu, Abhinaba
Veröffentlicht: (2026)
GATE: Graph-based Adaptive Tool Evolution Across Diverse Tasks
von: Luo, Jianwen, et al.
Veröffentlicht: (2025)
von: Luo, Jianwen, et al.
Veröffentlicht: (2025)
Diverse LLMs or Diverse Question Interpretations? That is the Ensembling Question
von: Rosales, Rafael, et al.
Veröffentlicht: (2025)
von: Rosales, Rafael, et al.
Veröffentlicht: (2025)
Mubeen AI: A Specialized Arabic Language Model for Heritage Preservation and User Intent Understanding
von: Aljafari, Mohammed, et al.
Veröffentlicht: (2025)
von: Aljafari, Mohammed, et al.
Veröffentlicht: (2025)
Give it Space! Explicit Disentangling of Positional and Semantic Representations in Encoders
von: Lequeu, Pierre-Antoine, et al.
Veröffentlicht: (2026)
von: Lequeu, Pierre-Antoine, et al.
Veröffentlicht: (2026)
Can LLMs Compute with Reasons?
von: Sandilya, Harshit, et al.
Veröffentlicht: (2024)
von: Sandilya, Harshit, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Causal Dimensionality of Transformer Representations: Measurement, Scaling, and Layer Structure
von: Sarkar, Nilesh, et al.
Veröffentlicht: (2026) -
Do LLMs Truly Understand When a Precedent Is Overruled?
von: Zhang, Li, et al.
Veröffentlicht: (2025) -
Shift-Reduce Task-Oriented Semantic Parsing with Stack-Transformers
von: Fernández-González, Daniel
Veröffentlicht: (2022) -
Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free
von: Zhang, Li, et al.
Veröffentlicht: (2026) -
Thinking Longer, Not Always Smarter: Evaluating LLM Capabilities in Hierarchical Legal Reasoning
von: Zhang, Li, et al.
Veröffentlicht: (2025)