Saved in:
| Main Authors: | Datta, Ayan, Marreddy, Mounika, Mehler, Alexander, Zhao, Zhixue, Mamidi, Radhika |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.00778 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Large Language Models Decide Early and Explain Later
by: Datta, Ayan, et al.
Published: (2026)
by: Datta, Ayan, et al.
Published: (2026)
SemEval-2024 Task 8: Weighted Layer Averaging RoBERTa for Black-Box Machine-Generated Text Detection
by: Datta, Ayan, et al.
Published: (2024)
by: Datta, Ayan, et al.
Published: (2024)
IndicSentEval: How Effectively do Multilingual Transformer Models encode Linguistic Properties for Indic Languages?
by: Aravapalli, Akhilesh, et al.
Published: (2024)
by: Aravapalli, Akhilesh, et al.
Published: (2024)
Analyzing Biases in Political Dialogue: Tagging U.S. Presidential Debates with an Extended DAMSL Framework
by: Prahallad, Lavanya, et al.
Published: (2025)
by: Prahallad, Lavanya, et al.
Published: (2025)
Significance of Chain of Thought in Gender Bias Mitigation for English-Dravidian Machine Translation
by: Prahallad, Lavanya, et al.
Published: (2024)
by: Prahallad, Lavanya, et al.
Published: (2024)
Enhancing Hate Speech Detection on Social Media: A Comparative Analysis of Machine Learning Models and Text Transformation Approaches
by: Mishra, Saurabh, et al.
Published: (2026)
by: Mishra, Saurabh, et al.
Published: (2026)
Zero-Shot Multi-task Hallucination Detection
by: Bhamidipati, Patanjali, et al.
Published: (2024)
by: Bhamidipati, Patanjali, et al.
Published: (2024)
USDC: A Dataset of $\underline{U}$ser $\underline{S}$tance and $\underline{D}$ogmatism in Long $\underline{C}$onversations
by: Marreddy, Mounika, et al.
Published: (2024)
by: Marreddy, Mounika, et al.
Published: (2024)
Mast Kalandar at SemEval-2024 Task 8: On the Trail of Textual Origins: RoBERTa-BiLSTM Approach to Detect AI-Generated Text
by: Bafna, Jainit Sushil, et al.
Published: (2024)
by: Bafna, Jainit Sushil, et al.
Published: (2024)
Multi-modal brain encoding models for multi-modal stimuli
by: Oota, Subba Reddy, et al.
Published: (2025)
by: Oota, Subba Reddy, et al.
Published: (2025)
Explanation Generation for Contradiction Reconciliation with LLMs
by: Chan, Jason, et al.
Published: (2026)
by: Chan, Jason, et al.
Published: (2026)
Position: On the Methodological Pitfalls of Evaluating Base LLMs for Reasoning
by: Chan, Jason, et al.
Published: (2025)
by: Chan, Jason, et al.
Published: (2025)
Position: Logical Soundness is not a Reliable Criterion for Neurosymbolic Fact-Checking with LLMs
by: Chan, Jason, et al.
Published: (2026)
by: Chan, Jason, et al.
Published: (2026)
RULEBREAKERS: Challenging LLMs at the Crossroads between Formal Logic and Human-like Reasoning
by: Chan, Jason, et al.
Published: (2024)
by: Chan, Jason, et al.
Published: (2024)
It's All About In-Context Learning! Teaching Extremely Low-Resource Languages to LLMs
by: Li, Yue, et al.
Published: (2025)
by: Li, Yue, et al.
Published: (2025)
Tracing and Reversing Edits in LLMs
by: Youssef, Paul, et al.
Published: (2025)
by: Youssef, Paul, et al.
Published: (2025)
How to Make LLMs Forget: On Reversing In-Context Knowledge Edits
by: Youssef, Paul, et al.
Published: (2024)
by: Youssef, Paul, et al.
Published: (2024)
Can Confidence Estimates Decide When Chain-of-Thought Is Necessary for LLMs?
by: Lewis-Lim, Samuel, et al.
Published: (2025)
by: Lewis-Lim, Samuel, et al.
Published: (2025)
The Spatial Semantics of Iconic Gesture
by: Lücking, Andy, et al.
Published: (2024)
by: Lücking, Andy, et al.
Published: (2024)
Comparing Explanation Faithfulness between Multilingual and Monolingual Fine-tuned Language Models
by: Zhao, Zhixue, et al.
Published: (2024)
by: Zhao, Zhixue, et al.
Published: (2024)
Survey-to-Behavior: Downstream Alignment of Human Values in LLMs via Survey Questions
by: Nie, Shangrui, et al.
Published: (2025)
by: Nie, Shangrui, et al.
Published: (2025)
Incorporating Attribution Importance for Improving Faithfulness Metrics
by: Zhao, Zhixue, et al.
Published: (2023)
by: Zhao, Zhixue, et al.
Published: (2023)
ReAGent: A Model-agnostic Feature Attribution Method for Generative Language Models
by: Zhao, Zhixue, et al.
Published: (2024)
by: Zhao, Zhixue, et al.
Published: (2024)
Label Set Optimization via Activation Distribution Kurtosis for Zero-shot Classification with Generative Models
by: Li, Yue, et al.
Published: (2024)
by: Li, Yue, et al.
Published: (2024)
Every Character Counts: From Vulnerability to Defense in Phishing Detection
by: Chiper, Maria, et al.
Published: (2025)
by: Chiper, Maria, et al.
Published: (2025)
Late-to-Early Training: LET LLMs Learn Earlier, So Faster and Better
by: Zhao, Ji, et al.
Published: (2026)
by: Zhao, Ji, et al.
Published: (2026)
LLMs Encode Harmfulness and Refusal Separately
by: Zhao, Jiachen, et al.
Published: (2025)
by: Zhao, Jiachen, et al.
Published: (2025)
Exploring Vision Language Models for Multimodal and Multilingual Stance Detection
by: Vasilakes, Jake, et al.
Published: (2025)
by: Vasilakes, Jake, et al.
Published: (2025)
Mechanistic Interpretability of Large-Scale Counting in LLMs through a System-2 Strategy
by: Hasani, Hosein, et al.
Published: (2026)
by: Hasani, Hosein, et al.
Published: (2026)
Do LLMs Provide Consistent Answers to Health-Related Questions across Languages?
by: Schlicht, Ipek Baris, et al.
Published: (2025)
by: Schlicht, Ipek Baris, et al.
Published: (2025)
Efficient Pruning of Text-to-Image Models: Insights from Pruning Stable Diffusion
by: Ramesh, Samarth N, et al.
Published: (2024)
by: Ramesh, Samarth N, et al.
Published: (2024)
Exploring the Limits of Model Compression in LLMs: A Knowledge Distillation Study on QA Tasks
by: Datta, Joyeeta, et al.
Published: (2025)
by: Datta, Joyeeta, et al.
Published: (2025)
SCRum-9: Multilingual Stance Classification over Rumours on Social Media
by: Li, Yue, et al.
Published: (2025)
by: Li, Yue, et al.
Published: (2025)
Position: Enough of Scaling LLMs! Lets Focus on Downscaling
by: Goel, Yash, et al.
Published: (2025)
by: Goel, Yash, et al.
Published: (2025)
Value-Guided KV Compression for LLMs via Approximated CUR Decomposition
by: Sengupta, Ayan, et al.
Published: (2025)
by: Sengupta, Ayan, et al.
Published: (2025)
Understanding the Physics of Key-Value Cache Compression for LLMs through Attention Dynamics
by: Ananthanarayanan, Samhruth, et al.
Published: (2026)
by: Ananthanarayanan, Samhruth, et al.
Published: (2026)
Contextual Position Encoding: Learning to Count What's Important
by: Golovneva, Olga, et al.
Published: (2024)
by: Golovneva, Olga, et al.
Published: (2024)
Understanding the Ability of LLMs to Handle Character-Level Perturbation
by: Zhuo, Anyuan, et al.
Published: (2025)
by: Zhuo, Anyuan, et al.
Published: (2025)
Enhancing Character-Level Understanding in LLMs through Token Internal Structure Learning
by: Xu, Zhu, et al.
Published: (2024)
by: Xu, Zhu, et al.
Published: (2024)
You Shall Know a Tool by the Traces it Leaves: The Predictability of Sentiment Analysis Tools
by: Baumartz, Daniel, et al.
Published: (2024)
by: Baumartz, Daniel, et al.
Published: (2024)
Similar Items
-
Large Language Models Decide Early and Explain Later
by: Datta, Ayan, et al.
Published: (2026) -
SemEval-2024 Task 8: Weighted Layer Averaging RoBERTa for Black-Box Machine-Generated Text Detection
by: Datta, Ayan, et al.
Published: (2024) -
IndicSentEval: How Effectively do Multilingual Transformer Models encode Linguistic Properties for Indic Languages?
by: Aravapalli, Akhilesh, et al.
Published: (2024) -
Analyzing Biases in Political Dialogue: Tagging U.S. Presidential Debates with an Extended DAMSL Framework
by: Prahallad, Lavanya, et al.
Published: (2025) -
Significance of Chain of Thought in Gender Bias Mitigation for English-Dravidian Machine Translation
by: Prahallad, Lavanya, et al.
Published: (2024)