Guardado en:
| Autores principales: | Chakrabarti, Lawhori, Johnson-Leung, Jennifer, Baumgaertner, Bert, Vakanski, Aleksandar, Xian, Min, Zhang, Boyu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2605.21391 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GCSAM: Gradient Centralized Sharpness Aware Minimization
por: Hassan, Mohamed, et al.
Publicado: (2025)
por: Hassan, Mohamed, et al.
Publicado: (2025)
A2DMN: Anatomy-Aware Dilated Multiscale Network for Breast Ultrasound Semantic Segmentation
por: Lucke, Kyle, et al.
Publicado: (2024)
por: Lucke, Kyle, et al.
Publicado: (2024)
Do Sharpness-based Optimizers Improve Generalization in Medical Image Analysis?
por: Hassan, Mohamed, et al.
Publicado: (2024)
por: Hassan, Mohamed, et al.
Publicado: (2024)
Encoder-Decoder or Decoder-Only? Revisiting Encoder-Decoder Large Language Model
por: Zhang, Biao, et al.
Publicado: (2025)
por: Zhang, Biao, et al.
Publicado: (2025)
Predictive Modeling and Uncertainty Quantification of Fatigue Life in Metal Alloys using Machine Learning
por: Chang, Jiang, et al.
Publicado: (2025)
por: Chang, Jiang, et al.
Publicado: (2025)
Density Matrices for Metaphor Understanding
por: Owers, Jay, et al.
Publicado: (2024)
por: Owers, Jay, et al.
Publicado: (2024)
Decomposition-Enhanced Training for Post-Hoc Attributions In Language Models
por: Balasubramanian, Sriram, et al.
Publicado: (2025)
por: Balasubramanian, Sriram, et al.
Publicado: (2025)
Exploring Communicative Participation in Care Homes
por: Katharina Giordano, et al.
Publicado: (2025)
por: Katharina Giordano, et al.
Publicado: (2025)
You Only Cache Once: Decoder-Decoder Architectures for Language Models
por: Sun, Yutao, et al.
Publicado: (2024)
por: Sun, Yutao, et al.
Publicado: (2024)
Comparative Analysis of Multi-Omics Integration Using Advanced Graph Neural Networks for Cancer Classification
por: Alharbi, Fadi, et al.
Publicado: (2024)
por: Alharbi, Fadi, et al.
Publicado: (2024)
Verifying Claims About Metaphors with Large-Scale Automatic Metaphor Identification
por: Aono, Kotaro, et al.
Publicado: (2024)
por: Aono, Kotaro, et al.
Publicado: (2024)
How Good is Post-Hoc Watermarking With Language Model Rephrasing?
por: Fernandez, Pierre, et al.
Publicado: (2025)
por: Fernandez, Pierre, et al.
Publicado: (2025)
Na'vi or Knave: Jailbreaking Language Models via Metaphorical Avatars
por: Yan, Yu, et al.
Publicado: (2024)
por: Yan, Yu, et al.
Publicado: (2024)
Not Eliminate but Aggregate: Post-Hoc Control over Mixture-of-Experts to Address Shortcut Shifts in Natural Language Understanding
por: Honda, Ukyo, et al.
Publicado: (2024)
por: Honda, Ukyo, et al.
Publicado: (2024)
Metaphor Understanding Challenge Dataset for LLMs
por: Tong, Xiaoyu, et al.
Publicado: (2024)
por: Tong, Xiaoyu, et al.
Publicado: (2024)
Towards Multimodal Metaphor Understanding: A Chinese Dataset and Model for Metaphor Mapping Identification
por: Zhang, Dongyu, et al.
Publicado: (2025)
por: Zhang, Dongyu, et al.
Publicado: (2025)
Uncertainty Quantification in Multivariable Regression for Material Property Prediction with Bayesian Neural Networks
por: Li, Longze, et al.
Publicado: (2023)
por: Li, Longze, et al.
Publicado: (2023)
Ethical Implications of Training Deceptive AI
por: Starace, Jason, et al.
Publicado: (2026)
por: Starace, Jason, et al.
Publicado: (2026)
Smoothie-Qwen: Post-Hoc Smoothing to Reduce Language Bias in Multilingual LLMs
por: Ji, SeungWon, et al.
Publicado: (2025)
por: Ji, SeungWon, et al.
Publicado: (2025)
Decentralized Distributed Proximal Policy Optimization (DD-PPO) for High Performance Computing Scheduling on Multi-User Systems
por: Sgambati, Matthew, et al.
Publicado: (2025)
por: Sgambati, Matthew, et al.
Publicado: (2025)
Scaling Laws of Decoder-Only Models on the Multilingual Machine Translation Task
por: Caillaut, Gaëtan, et al.
Publicado: (2024)
por: Caillaut, Gaëtan, et al.
Publicado: (2024)
Self-AMPLIFY: Improving Small Language Models with Self Post Hoc Explanations
por: Bhan, Milan, et al.
Publicado: (2024)
por: Bhan, Milan, et al.
Publicado: (2024)
Machine Translation with Large Language Models: Decoder Only vs. Encoder-Decoder
por: M., Abhinav P., et al.
Publicado: (2024)
por: M., Abhinav P., et al.
Publicado: (2024)
from Benign import Toxic: Jailbreaking the Language Model via Adversarial Metaphors
por: Yan, Yu, et al.
Publicado: (2025)
por: Yan, Yu, et al.
Publicado: (2025)
From Metaphor to Mechanism: How LLMs Decode Traditional Chinese Medicine Symbolic Language for Modern Clinical Relevance
por: Tang, Jiacheng, et al.
Publicado: (2025)
por: Tang, Jiacheng, et al.
Publicado: (2025)
Evaluating Small Decoder-Only Language Models for Grammar Correction and Text Simplification
por: Lamelas, Anthony
Publicado: (2026)
por: Lamelas, Anthony
Publicado: (2026)
On The Adaptation of Unlimiformer for Decoder-Only Transformers
por: Ahrabian, Kian, et al.
Publicado: (2024)
por: Ahrabian, Kian, et al.
Publicado: (2024)
Enhancing Post-Hoc Attributions in Long Document Comprehension via Coarse Grained Answer Decomposition
por: Ramu, Pritika, et al.
Publicado: (2024)
por: Ramu, Pritika, et al.
Publicado: (2024)
The Role of Syntactic Span Preferences in Post-Hoc Explanation Disagreement
por: Kamp, Jonathan, et al.
Publicado: (2024)
por: Kamp, Jonathan, et al.
Publicado: (2024)
Meanings are like Onions: a Layered Approach to Metaphor Processing
por: Cappa, Silvia, et al.
Publicado: (2025)
por: Cappa, Silvia, et al.
Publicado: (2025)
Metaphor and Large Language Models: When Surface Features Matter More than Deep Understanding
por: Sanchez-Bayona, Elisa, et al.
Publicado: (2025)
por: Sanchez-Bayona, Elisa, et al.
Publicado: (2025)
MzansiText and MzansiLM: An Open Corpus and Decoder-Only Language Model for South African Languages
por: Lombard, Anri, et al.
Publicado: (2026)
por: Lombard, Anri, et al.
Publicado: (2026)
EntropyCache: Decoded Token Entropy Guided KV Caching for Diffusion Language Models
por: Cheong, Minsoo, et al.
Publicado: (2026)
por: Cheong, Minsoo, et al.
Publicado: (2026)
Is More Data Worth the Cost? Dataset Scaling Laws in a Tiny Attention-Only Decoder
por: Wiegand, Götz-Henrik, et al.
Publicado: (2026)
por: Wiegand, Götz-Henrik, et al.
Publicado: (2026)
Cost-Performance Optimization for Processing Low-Resource Language Tasks Using Commercial LLMs
por: Nag, Arijit, et al.
Publicado: (2024)
por: Nag, Arijit, et al.
Publicado: (2024)
Understanding Performance Collapse in Layer-Pruned Large Language Models via Decision Representation Transitions
por: Shi, Boyu, et al.
Publicado: (2026)
por: Shi, Boyu, et al.
Publicado: (2026)
Entropy-Based Decoding for Retrieval-Augmented Large Language Models
por: Qiu, Zexuan, et al.
Publicado: (2024)
por: Qiu, Zexuan, et al.
Publicado: (2024)
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models
por: Wu, Di, et al.
Publicado: (2024)
por: Wu, Di, et al.
Publicado: (2024)
The Frequency Confound in Language-Model Surprisal and Metaphor Novelty
por: Momen, Omar, et al.
Publicado: (2026)
por: Momen, Omar, et al.
Publicado: (2026)
Finding Challenging Metaphors that Confuse Pretrained Language Models
por: Li, Yucheng, et al.
Publicado: (2024)
por: Li, Yucheng, et al.
Publicado: (2024)
Ejemplares similares
-
GCSAM: Gradient Centralized Sharpness Aware Minimization
por: Hassan, Mohamed, et al.
Publicado: (2025) -
A2DMN: Anatomy-Aware Dilated Multiscale Network for Breast Ultrasound Semantic Segmentation
por: Lucke, Kyle, et al.
Publicado: (2024) -
Do Sharpness-based Optimizers Improve Generalization in Medical Image Analysis?
por: Hassan, Mohamed, et al.
Publicado: (2024) -
Encoder-Decoder or Decoder-Only? Revisiting Encoder-Decoder Large Language Model
por: Zhang, Biao, et al.
Publicado: (2025) -
Predictive Modeling and Uncertainty Quantification of Fatigue Life in Metal Alloys using Machine Learning
por: Chang, Jiang, et al.
Publicado: (2025)