SIGMA: Scalable Spectral Insights for LLM Model Collapse
Fuente:
arXiv
Salvato in:
| Autori principali: | Gu, Yi, Pang, Lingyou, Ye, Xiangkun, Wang, Tianyu, Lin, Jianyu, Priebe, Carey E., Aue, Alexander |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Linguistic Collapse: Neural Collapse in (Large) Language Models
di: Wu, Robert, et al.
Pubblicazione: (2024)
di: Wu, Robert, et al.
Pubblicazione: (2024)
MORQA: Benchmarking Evaluation Metrics for Medical Open-Ended Question Answering
di: Yim, Wen-wai, et al.
Pubblicazione: (2025)
di: Yim, Wen-wai, et al.
Pubblicazione: (2025)
AI in Investment Analysis: LLMs for Equity Stock Ratings
di: Papasotiriou, Kassiani, et al.
Pubblicazione: (2024)
di: Papasotiriou, Kassiani, et al.
Pubblicazione: (2024)
Model Surgery: Modulating LLM's Behavior Via Simple Parameter Editing
di: Wang, Huanqian, et al.
Pubblicazione: (2024)
di: Wang, Huanqian, et al.
Pubblicazione: (2024)
Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?
di: Pan, Leyi, et al.
Pubblicazione: (2025)
di: Pan, Leyi, et al.
Pubblicazione: (2025)
Contrasting Linguistic Patterns in Human and LLM-Generated News Text
di: Muñoz-Ortiz, Alberto, et al.
Pubblicazione: (2023)
di: Muñoz-Ortiz, Alberto, et al.
Pubblicazione: (2023)
A Survey on Collaborating Small and Large Language Models for Performance, Cost-effectiveness, Cloud-edge Privacy, and Trustworthiness
di: Wang, Fali, et al.
Pubblicazione: (2025)
di: Wang, Fali, et al.
Pubblicazione: (2025)
Compression Method Matters: Benchmark-Dependent Output Dynamics in LLM Prompt Compression
di: Johnson, Warren
Pubblicazione: (2026)
di: Johnson, Warren
Pubblicazione: (2026)
RoleRAG: Enhancing LLM Role-Playing via Graph Guided Retrieval
di: Wang, Yongjie, et al.
Pubblicazione: (2025)
di: Wang, Yongjie, et al.
Pubblicazione: (2025)
Co-NAML-LSTUR: A Combined Model with Attentive Multi-View Learning and Long- and Short-term User Representations for News Recommendation
di: Nguyen, Minh Hoang, et al.
Pubblicazione: (2025)
di: Nguyen, Minh Hoang, et al.
Pubblicazione: (2025)
PaperAudit-Bench: Benchmarking Error Detection in Research Papers for Critical Automated Peer Review
di: Tu, Songjun, et al.
Pubblicazione: (2026)
di: Tu, Songjun, et al.
Pubblicazione: (2026)
Exploring and Reshaping the Weight Distribution in LLM
di: Ye, Chunming, et al.
Pubblicazione: (2025)
di: Ye, Chunming, et al.
Pubblicazione: (2025)
An NLP-Driven Framework for Curriculum-Labor Market Alignment: Schema-Constrained LLM Extraction, ESCO-Anchored Semantic Matching, and Multi-Dimensional Gap Quantification
di: Turaev, Sherzod, et al.
Pubblicazione: (2026)
di: Turaev, Sherzod, et al.
Pubblicazione: (2026)
Fast Quiet-STaR: Thinking Without Thought Tokens
di: Huang, Wei, et al.
Pubblicazione: (2025)
di: Huang, Wei, et al.
Pubblicazione: (2025)
Make Literature-Based Discovery Great Again through Reproducible Pipelines
di: Cestnik, Bojan, et al.
Pubblicazione: (2025)
di: Cestnik, Bojan, et al.
Pubblicazione: (2025)
Data and AI governance: Promoting equity, ethics, and fairness in large language models
di: Abhishek, Alok, et al.
Pubblicazione: (2025)
di: Abhishek, Alok, et al.
Pubblicazione: (2025)
Do Reasoning Models Enhance Embedding Models?
di: Chan, Wun Yu, et al.
Pubblicazione: (2026)
di: Chan, Wun Yu, et al.
Pubblicazione: (2026)
SHARP: Social Harm Analysis via Risk Profiles for Measuring Inequities in Large Language Models
di: Abhishek, Alok, et al.
Pubblicazione: (2026)
di: Abhishek, Alok, et al.
Pubblicazione: (2026)
Judgment2vec: Apply Graph Analytics to Searching and Recommendation of Similar Judgments
di: Shao, Hsuan-Lei
Pubblicazione: (2024)
di: Shao, Hsuan-Lei
Pubblicazione: (2024)
BEATS: Bias Evaluation and Assessment Test Suite for Large Language Models
di: Abhishek, Alok, et al.
Pubblicazione: (2025)
di: Abhishek, Alok, et al.
Pubblicazione: (2025)
Research on a hybrid LSTM-CNN-Attention model for text-based web content classification
di: Kuz, Mykola, et al.
Pubblicazione: (2025)
di: Kuz, Mykola, et al.
Pubblicazione: (2025)
Reasoning Promotes Robustness in Theory of Mind Tasks
di: de Haan, Ian B., et al.
Pubblicazione: (2026)
di: de Haan, Ian B., et al.
Pubblicazione: (2026)
MedMemoryBench: Benchmarking Agent Memory in Personalized Healthcare
di: Wang, Yihao, et al.
Pubblicazione: (2026)
di: Wang, Yihao, et al.
Pubblicazione: (2026)
Rule Extraction in Machine Learning: Chat Incremental Pattern Constructor
di: Nwokocha, Caleb Princewill
Pubblicazione: (2022)
di: Nwokocha, Caleb Princewill
Pubblicazione: (2022)
Constitution or Collapse? Exploring Constitutional AI with Llama 3-8B
di: Zhang, Xue
Pubblicazione: (2025)
di: Zhang, Xue
Pubblicazione: (2025)
Large Language Models are Inconsistent and Biased Evaluators
di: Stureborg, Rickard, et al.
Pubblicazione: (2024)
di: Stureborg, Rickard, et al.
Pubblicazione: (2024)
TRAWL: Tensor Reduced and Approximated Weights for Large Language Models
di: Luo, Yiran, et al.
Pubblicazione: (2024)
di: Luo, Yiran, et al.
Pubblicazione: (2024)
From Noise to Diversity: Random Embedding Injection in LLM Reasoning
di: Kim, Heejun, et al.
Pubblicazione: (2026)
di: Kim, Heejun, et al.
Pubblicazione: (2026)
Emotional Sequential Influence Modeling on False Information
di: Naskar, Debashis, et al.
Pubblicazione: (2024)
di: Naskar, Debashis, et al.
Pubblicazione: (2024)
D-SMART: Enhancing LLM Dialogue Consistency via Dynamic Structured Memory And Reasoning Tree
di: Lei, Xiang, et al.
Pubblicazione: (2025)
di: Lei, Xiang, et al.
Pubblicazione: (2025)
Thinking Longer, Not Always Smarter: Evaluating LLM Capabilities in Hierarchical Legal Reasoning
di: Zhang, Li, et al.
Pubblicazione: (2025)
di: Zhang, Li, et al.
Pubblicazione: (2025)
A transfer learning approach for automatic conflicts detection in software requirement sentence pairs based on dual encoders
di: Wang, Yizheng, et al.
Pubblicazione: (2025)
di: Wang, Yizheng, et al.
Pubblicazione: (2025)
MarkLLM: An Open-Source Toolkit for LLM Watermarking
di: Pan, Leyi, et al.
Pubblicazione: (2024)
di: Pan, Leyi, et al.
Pubblicazione: (2024)
The Knesset Corpus: An Annotated Corpus of Hebrew Parliamentary Proceedings
di: Goldin, Gili, et al.
Pubblicazione: (2024)
di: Goldin, Gili, et al.
Pubblicazione: (2024)
Math Natural Language Inference: this should be easy!
di: de Paiva, Valeria, et al.
Pubblicazione: (2025)
di: de Paiva, Valeria, et al.
Pubblicazione: (2025)
New Skills or Sharper Primitives? A Probabilistic Perspective on the Emergence of Reasoning in RLVR
di: Wang, Zhilin, et al.
Pubblicazione: (2026)
di: Wang, Zhilin, et al.
Pubblicazione: (2026)
Pitfalls in Evaluating Interpretability Agents
di: Haklay, Tal, et al.
Pubblicazione: (2026)
di: Haklay, Tal, et al.
Pubblicazione: (2026)
d-TreeRPO: Towards More Reliable Policy Optimization for Diffusion Language Models
di: Pan, Leyi, et al.
Pubblicazione: (2025)
di: Pan, Leyi, et al.
Pubblicazione: (2025)
Towards Effective and Efficient Continual Pre-training of Large Language Models
di: Chen, Jie, et al.
Pubblicazione: (2024)
di: Chen, Jie, et al.
Pubblicazione: (2024)
An Unforgeable Publicly Verifiable Watermark for Large Language Models
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
di: Liu, Aiwei, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Linguistic Collapse: Neural Collapse in (Large) Language Models
di: Wu, Robert, et al.
Pubblicazione: (2024) -
MORQA: Benchmarking Evaluation Metrics for Medical Open-Ended Question Answering
di: Yim, Wen-wai, et al.
Pubblicazione: (2025) -
AI in Investment Analysis: LLMs for Equity Stock Ratings
di: Papasotiriou, Kassiani, et al.
Pubblicazione: (2024) -
Model Surgery: Modulating LLM's Behavior Via Simple Parameter Editing
di: Wang, Huanqian, et al.
Pubblicazione: (2024) -
Can LLM Watermarks Robustly Prevent Unauthorized Knowledge Distillation?
di: Pan, Leyi, et al.
Pubblicazione: (2025)