Gespeichert in:
| Hauptverfasser: | Chytas, Sotirios Panagiotis, Singh, Vikas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2601.11575 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FoGE: Fock Space inspired encoding for graph prompting
von: Chytas, Sotirios Panagiotis, et al.
Veröffentlicht: (2025)
von: Chytas, Sotirios Panagiotis, et al.
Veröffentlicht: (2025)
ReCo: Reminder Composition Mitigates Hallucinations in Vision-Language Models
von: Chytas, Sotirios Panagiotis, et al.
Veröffentlicht: (2025)
von: Chytas, Sotirios Panagiotis, et al.
Veröffentlicht: (2025)
Pooling Image Datasets With Multiple Covariate Shift and Imbalance
von: Chytas, Sotirios Panagiotis, et al.
Veröffentlicht: (2024)
von: Chytas, Sotirios Panagiotis, et al.
Veröffentlicht: (2024)
Evaluating the Efficacy of AI Techniques in Textual Anonymization: A Comparative Study
von: Asimopoulos, Dimitris, et al.
Veröffentlicht: (2024)
von: Asimopoulos, Dimitris, et al.
Veröffentlicht: (2024)
Benchmarking Concept-Spilling Across Languages in LLMs
von: Badanin, Ilia, et al.
Veröffentlicht: (2026)
von: Badanin, Ilia, et al.
Veröffentlicht: (2026)
Error Taxonomy-Guided Prompt Optimization
von: Singh, Mayank, et al.
Veröffentlicht: (2026)
von: Singh, Mayank, et al.
Veröffentlicht: (2026)
Systematic Evaluation of Long-Context LLMs on Financial Concepts
von: Gupta, Lavanya, et al.
Veröffentlicht: (2024)
von: Gupta, Lavanya, et al.
Veröffentlicht: (2024)
Concept-Based Interpretability for Toxicity Detection
von: Garg, Samarth, et al.
Veröffentlicht: (2025)
von: Garg, Samarth, et al.
Veröffentlicht: (2025)
Domain Knowledge-Enhanced LLMs for Fraud and Concept Drift Detection
von: Şenol, Ali, et al.
Veröffentlicht: (2025)
von: Şenol, Ali, et al.
Veröffentlicht: (2025)
CRCE: Coreference-Retention Concept Erasure in Text-to-Image Diffusion Models
von: Xue, Yuyang, et al.
Veröffentlicht: (2025)
von: Xue, Yuyang, et al.
Veröffentlicht: (2025)
Soft Thinking: Unlocking the Reasoning Potential of LLMs in Continuous Concept Space
von: Zhang, Zhen, et al.
Veröffentlicht: (2025)
von: Zhang, Zhen, et al.
Veröffentlicht: (2025)
Beyond Single Concept Vector: Modeling Concept Subspace in LLMs with Gaussian Distribution
von: Zhao, Haiyan, et al.
Veröffentlicht: (2024)
von: Zhao, Haiyan, et al.
Veröffentlicht: (2024)
Funny or Persuasive, but Not Both: Evaluating Fine-Grained Multi-Concept Control in LLMs
von: Labroo, Arya, et al.
Veröffentlicht: (2026)
von: Labroo, Arya, et al.
Veröffentlicht: (2026)
Exploiting LLMs' Reasoning Capability to Infer Implicit Concepts in Legal Information Retrieval
von: Nguyen, Hai-Long, et al.
Veröffentlicht: (2024)
von: Nguyen, Hai-Long, et al.
Veröffentlicht: (2024)
NEAT: Concept driven Neuron Attribution in LLMs
von: Kavuri, Vivek Hruday, et al.
Veröffentlicht: (2025)
von: Kavuri, Vivek Hruday, et al.
Veröffentlicht: (2025)
Can LLMs Capture Human Preferences?
von: Goli, Ali, et al.
Veröffentlicht: (2023)
von: Goli, Ali, et al.
Veröffentlicht: (2023)
Are Today's LLMs Ready to Explain Well-Being Concepts?
von: Jiang, Bohan, et al.
Veröffentlicht: (2025)
von: Jiang, Bohan, et al.
Veröffentlicht: (2025)
Evaluate Bias without Manual Test Sets: A Concept Representation Perspective for LLMs
von: Gao, Lang, et al.
Veröffentlicht: (2025)
von: Gao, Lang, et al.
Veröffentlicht: (2025)
LIBERTy: A Causal Framework for Benchmarking Concept-Based Explanations of LLMs with Structural Counterfactuals
von: Toker, Gilat, et al.
Veröffentlicht: (2026)
von: Toker, Gilat, et al.
Veröffentlicht: (2026)
The Grounding Gap: How LLMs Anchor the Meaning of Abstract Concepts Differently from Humans
von: Chlapanis, Odysseas S., et al.
Veröffentlicht: (2026)
von: Chlapanis, Odysseas S., et al.
Veröffentlicht: (2026)
From Concrete to Abstract: A Multimodal Generative Approach to Abstract Concept Learning
von: Xie, Haodong, et al.
Veröffentlicht: (2024)
von: Xie, Haodong, et al.
Veröffentlicht: (2024)
SLIDE: Reference-free Evaluation for Machine Translation using a Sliding Document Window
von: Raunak, Vikas, et al.
Veröffentlicht: (2023)
von: Raunak, Vikas, et al.
Veröffentlicht: (2023)
Benchmarking Advanced Text Anonymisation Methods: A Comparative Study on Novel and Traditional Approaches
von: Asimopoulos, Dimitris, et al.
Veröffentlicht: (2024)
von: Asimopoulos, Dimitris, et al.
Veröffentlicht: (2024)
Solve the Loop: Attractor Models for Language and Reasoning
von: Fein-Ashley, Jacob, et al.
Veröffentlicht: (2026)
von: Fein-Ashley, Jacob, et al.
Veröffentlicht: (2026)
PythonSaga: Redefining the Benchmark to Evaluate Code Generating LLMs
von: Yadav, Ankit, et al.
Veröffentlicht: (2024)
von: Yadav, Ankit, et al.
Veröffentlicht: (2024)
Position: Avoid Overstretching LLMs for every Enterprise Task
von: Singh, Kuldeep, et al.
Veröffentlicht: (2026)
von: Singh, Kuldeep, et al.
Veröffentlicht: (2026)
Prompting with Phonemes: Enhancing LLMs' Multilinguality for Non-Latin Script Languages
von: Nguyen, Hoang H, et al.
Veröffentlicht: (2024)
von: Nguyen, Hoang H, et al.
Veröffentlicht: (2024)
On Instruction-Finetuning Neural Machine Translation Models
von: Raunak, Vikas, et al.
Veröffentlicht: (2024)
von: Raunak, Vikas, et al.
Veröffentlicht: (2024)
A Toolbox, Not a Hammer -- Multi-TAG: Scaling Math Reasoning with Multi-Tool Aggregation
von: Yao, Bohan, et al.
Veröffentlicht: (2025)
von: Yao, Bohan, et al.
Veröffentlicht: (2025)
Reasoning about concepts with LLMs: Inconsistencies abound
von: Uceda-Sosa, Rosario, et al.
Veröffentlicht: (2024)
von: Uceda-Sosa, Rosario, et al.
Veröffentlicht: (2024)
Hallucination as Trajectory Commitment: Causal Evidence for Asymmetric Attractor Dynamics in Transformer Generation
von: Akarlar, G. Aytug
Veröffentlicht: (2026)
von: Akarlar, G. Aytug
Veröffentlicht: (2026)
Towards Reliable Evaluation of Behavior Steering Interventions in LLMs
von: Pres, Itamar, et al.
Veröffentlicht: (2024)
von: Pres, Itamar, et al.
Veröffentlicht: (2024)
Interpreting the Effects of Quantization on LLMs
von: Singh, Manpreet, et al.
Veröffentlicht: (2025)
von: Singh, Manpreet, et al.
Veröffentlicht: (2025)
Estimation of Concept Explanations Should be Uncertainty Aware
von: Piratla, Vihari, et al.
Veröffentlicht: (2023)
von: Piratla, Vihari, et al.
Veröffentlicht: (2023)
EtiCor++: Towards Understanding Etiquettical Bias in LLMs
von: Dwivedi, Ashutosh, et al.
Veröffentlicht: (2025)
von: Dwivedi, Ashutosh, et al.
Veröffentlicht: (2025)
XplainLLM: A Knowledge-Augmented Dataset for Reliable Grounded Explanations in LLMs
von: Chen, Zichen, et al.
Veröffentlicht: (2023)
von: Chen, Zichen, et al.
Veröffentlicht: (2023)
Addressing Bias in LLMs: Strategies and Application to Fair AI-based Recruitment
von: Peña, Alejandro, et al.
Veröffentlicht: (2025)
von: Peña, Alejandro, et al.
Veröffentlicht: (2025)
Judging the Judges: Evaluating Alignment and Vulnerabilities in LLMs-as-Judges
von: Thakur, Aman Singh, et al.
Veröffentlicht: (2024)
von: Thakur, Aman Singh, et al.
Veröffentlicht: (2024)
Benchmarking LLMs for Pairwise Causal Discovery in Biomedical and Multi-Domain Contexts
von: Anuyah, Sydney, et al.
Veröffentlicht: (2026)
von: Anuyah, Sydney, et al.
Veröffentlicht: (2026)
Futureproof Static Memory Planning
von: Lamprakos, Christos, et al.
Veröffentlicht: (2025)
von: Lamprakos, Christos, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FoGE: Fock Space inspired encoding for graph prompting
von: Chytas, Sotirios Panagiotis, et al.
Veröffentlicht: (2025) -
ReCo: Reminder Composition Mitigates Hallucinations in Vision-Language Models
von: Chytas, Sotirios Panagiotis, et al.
Veröffentlicht: (2025) -
Pooling Image Datasets With Multiple Covariate Shift and Imbalance
von: Chytas, Sotirios Panagiotis, et al.
Veröffentlicht: (2024) -
Evaluating the Efficacy of AI Techniques in Textual Anonymization: A Comparative Study
von: Asimopoulos, Dimitris, et al.
Veröffentlicht: (2024) -
Benchmarking Concept-Spilling Across Languages in LLMs
von: Badanin, Ilia, et al.
Veröffentlicht: (2026)