Mimicking How Humans Interpret Out-of-Context Sentences Through Controlled Toxicity Decoding
Fuente:
arXiv
Saved in:
| Main Authors: | Trusca, Maria Mihaela, Allein, Liesbeth |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Assessing LLM Reasoning Through Implicit Causal Chain Discovery in Climate Discourse
by: Allein, Liesbeth, et al.
Published: (2025)
by: Allein, Liesbeth, et al.
Published: (2025)
ClimateCause: Complex and Implicit Causal Structures in Climate Reports
by: Allein, Liesbeth, et al.
Published: (2026)
by: Allein, Liesbeth, et al.
Published: (2026)
Sequence-to-Sequence Spanish Pre-trained Language Models
by: Araujo, Vladimir, et al.
Published: (2023)
by: Araujo, Vladimir, et al.
Published: (2023)
Edit-Constrained Decoding for Sentence Simplification
by: Zetsu, Tatsuya, et al.
Published: (2024)
by: Zetsu, Tatsuya, et al.
Published: (2024)
When Smiley Turns Hostile: Interpreting How Emojis Trigger LLMs' Toxicity
by: Cui, Shiyao, et al.
Published: (2025)
by: Cui, Shiyao, et al.
Published: (2025)
Action-based image editing guided by human instructions
by: Trusca, Maria Mihaela, et al.
Published: (2024)
by: Trusca, Maria Mihaela, et al.
Published: (2024)
DM-Align: Leveraging the Power of Natural Language Instructions to Make Changes to Images
by: Trusca, Maria Mihaela, et al.
Published: (2024)
by: Trusca, Maria Mihaela, et al.
Published: (2024)
Efficient Zero-Shot Long Document Classification by Reducing Context Through Sentence Ranking
by: Kokate, Prathamesh, et al.
Published: (2025)
by: Kokate, Prathamesh, et al.
Published: (2025)
Computational Sentence-level Metrics Predicting Human Sentence Comprehension
by: Sun, Kun, et al.
Published: (2024)
by: Sun, Kun, et al.
Published: (2024)
Sentence-Anchored Gist Compression for Long-Context LLMs
by: Tarasov, Dmitrii, et al.
Published: (2025)
by: Tarasov, Dmitrii, et al.
Published: (2025)
Predicting Sentence Acceptability Judgments in Multimodal Contexts
by: Jang, Hyewon, et al.
Published: (2026)
by: Jang, Hyewon, et al.
Published: (2026)
XL-DURel: Finetuning Sentence Transformers for Ordinal Word-in-Context Classification
by: Yadav, Sachin, et al.
Published: (2025)
by: Yadav, Sachin, et al.
Published: (2025)
Cross-Preference Learning for Sentence-Level and Context-Aware Machine Translation
by: Li, Ying, et al.
Published: (2026)
by: Li, Ying, et al.
Published: (2026)
Interpretable Detection of Out-of-Context Misinformation with Neural-Symbolic-Enhanced Large Multimodal Model
by: Zhang, Yizhou, et al.
Published: (2023)
by: Zhang, Yizhou, et al.
Published: (2023)
DecoderLens: Layerwise Interpretation of Encoder-Decoder Transformers
by: Langedijk, Anna, et al.
Published: (2023)
by: Langedijk, Anna, et al.
Published: (2023)
Concept-Based Interpretability for Toxicity Detection
by: Garg, Samarth, et al.
Published: (2025)
by: Garg, Samarth, et al.
Published: (2025)
Interpretable Stylistic Variation in Human and LLM Writing Across Genres, Models, and Decoding Strategies
by: Rallapalli, Swati, et al.
Published: (2026)
by: Rallapalli, Swati, et al.
Published: (2026)
ProTrix: Building Models for Planning and Reasoning over Tables with Sentence Context
by: Wu, Zirui, et al.
Published: (2024)
by: Wu, Zirui, et al.
Published: (2024)
Sentence Smith: Controllable Edits for Evaluating Text Embeddings
by: Li, Hongji, et al.
Published: (2025)
by: Li, Hongji, et al.
Published: (2025)
Analysing Zero-Shot Readability-Controlled Sentence Simplification
by: Barayan, Abdullah, et al.
Published: (2024)
by: Barayan, Abdullah, et al.
Published: (2024)
How Many Bytes Can You Take Out Of Brain-To-Text Decoding?
by: Antonello, Richard, et al.
Published: (2024)
by: Antonello, Richard, et al.
Published: (2024)
Mitigating Out-of-Entity Errors in Named Entity Recognition: A Sentence-Level Strategy
by: Jiang, Guochao, et al.
Published: (2024)
by: Jiang, Guochao, et al.
Published: (2024)
Stephanie: Step-by-Step Dialogues for Mimicking Human Interactions in Social Conversations
by: Yang, Hao, et al.
Published: (2024)
by: Yang, Hao, et al.
Published: (2024)
Redefining Experts: Interpretable Decomposition of Language Models for Toxicity Mitigation
by: Shaik, Zuhair Hasan, et al.
Published: (2025)
by: Shaik, Zuhair Hasan, et al.
Published: (2025)
Human-Interpretable Adversarial Prompt Attack on Large Language Models with Situational Context
by: Das, Nilanjana, et al.
Published: (2024)
by: Das, Nilanjana, et al.
Published: (2024)
Dual Alignment Between Language Model Layers and Human Sentence Processing
by: Kuribayashi, Tatsuki, et al.
Published: (2026)
by: Kuribayashi, Tatsuki, et al.
Published: (2026)
ParetoRAG: Leveraging Sentence-Context Attention for Robust and Efficient Retrieval-Augmented Generation
by: Yao, Ruobing, et al.
Published: (2025)
by: Yao, Ruobing, et al.
Published: (2025)
Do Large Language Models Understand Logic or Just Mimick Context?
by: Yan, Junbing, et al.
Published: (2024)
by: Yan, Junbing, et al.
Published: (2024)
Wide-In, Narrow-Out: Revokable Decoding for Efficient and Effective DLLMs
by: Hong, Feng, et al.
Published: (2025)
by: Hong, Feng, et al.
Published: (2025)
Multilingual Sentence-T5: Scalable Sentence Encoders for Multilingual Applications
by: Yano, Chihiro, et al.
Published: (2024)
by: Yano, Chihiro, et al.
Published: (2024)
Timeline-based Sentence Decomposition with In-Context Learning for Temporal Fact Extraction
by: Chen, Jianhao, et al.
Published: (2024)
by: Chen, Jianhao, et al.
Published: (2024)
Out-of-Vocabulary Sampling Boosts Speculative Decoding
by: Timor, Nadav, et al.
Published: (2025)
by: Timor, Nadav, et al.
Published: (2025)
Out of Sight Out of Mind, Out of Sight Out of Mind: Measuring Bias in Language Models Against Overlooked Marginalized Groups in Regional Contexts
by: Elsafoury, Fatma, et al.
Published: (2025)
by: Elsafoury, Fatma, et al.
Published: (2025)
How Effectively Can BERT Models Interpret Context and Detect Bengali Communal Violent Text?
by: Khondoker, Abdullah, et al.
Published: (2025)
by: Khondoker, Abdullah, et al.
Published: (2025)
Modeling Human Sentence Processing with Left-Corner Recurrent Neural Network Grammars
by: Yoshida, Ryo, et al.
Published: (2021)
by: Yoshida, Ryo, et al.
Published: (2021)
Task-agnostic Prompt Compression with Context-aware Sentence Embedding and Reward-guided Task Descriptor
by: Liskavets, Barys, et al.
Published: (2025)
by: Liskavets, Barys, et al.
Published: (2025)
Uncovering Gaps in How Humans and LLMs Interpret Subjective Language
by: Jones, Erik, et al.
Published: (2025)
by: Jones, Erik, et al.
Published: (2025)
SCI-IDEA: Context-Aware Scientific Ideation Using Token and Sentence Embeddings
by: Keya, Farhana, et al.
Published: (2025)
by: Keya, Farhana, et al.
Published: (2025)
Prompt Compression with Context-Aware Sentence Encoding for Fast and Improved LLM Inference
by: Liskavets, Barys, et al.
Published: (2024)
by: Liskavets, Barys, et al.
Published: (2024)
Contextual Clarity: Generating Sentences with Transformer Models using Context-Reverso Data
by: Musaev, Ruslan
Published: (2024)
by: Musaev, Ruslan
Published: (2024)
Similar Items
-
Assessing LLM Reasoning Through Implicit Causal Chain Discovery in Climate Discourse
by: Allein, Liesbeth, et al.
Published: (2025) -
ClimateCause: Complex and Implicit Causal Structures in Climate Reports
by: Allein, Liesbeth, et al.
Published: (2026) -
Sequence-to-Sequence Spanish Pre-trained Language Models
by: Araujo, Vladimir, et al.
Published: (2023) -
Edit-Constrained Decoding for Sentence Simplification
by: Zetsu, Tatsuya, et al.
Published: (2024) -
When Smiley Turns Hostile: Interpreting How Emojis Trigger LLMs' Toxicity
by: Cui, Shiyao, et al.
Published: (2025)