X-SHIELD: Regularization for eXplainable Artificial Intelligence
Fuente:
arXiv
Guardado en:
| Autores principales: | Sevillano-García, Iván, Luengo, Julián, Herrera, Francisco |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Novel Approaches to Artificial Intelligence Development Based on the Nearest Neighbor Method
por: Priezzhev, I. I., et al.
Publicado: (2025)
por: Priezzhev, I. I., et al.
Publicado: (2025)
Data Ethics in the Era of Healthcare Artificial Intelligence in Africa: An Ubuntu Philosophy Perspective
por: Mahamadou, Abdoul Jalil Djiberou, et al.
Publicado: (2024)
por: Mahamadou, Abdoul Jalil Djiberou, et al.
Publicado: (2024)
Foundational Requirements for Artificial General Intelligence: A Falsifiable Framework Based on Signal Prediction
por: Šprogar, Matej
Publicado: (2025)
por: Šprogar, Matej
Publicado: (2025)
Intervention Complexity as a Canonical Reward and a Measure of Intelligence
por: McCane, Brendan
Publicado: (2026)
por: McCane, Brendan
Publicado: (2026)
A Study on the Application of Artificial Intelligence in Ecological Design
por: Zhao, Hengyue
Publicado: (2025)
por: Zhao, Hengyue
Publicado: (2025)
Deep Memory Search: A Metaheuristic Approach for Optimizing Heuristic Search
por: Hedar, Abdel-Rahman, et al.
Publicado: (2024)
por: Hedar, Abdel-Rahman, et al.
Publicado: (2024)
The Stochastic Gap: A Markovian Framework for Pre-Deployment Reliability and Oversight-Cost Auditing in Agentic Artificial Intelligence
por: Pal, Biplab, et al.
Publicado: (2026)
por: Pal, Biplab, et al.
Publicado: (2026)
Embedded Universal Predictive Intelligence: a coherent framework for multi-agent learning
por: Meulemans, Alexander, et al.
Publicado: (2025)
por: Meulemans, Alexander, et al.
Publicado: (2025)
Large Language Models as Attribution Regularizers for Efficient Model Training
por: Vukadin, Davor, et al.
Publicado: (2025)
por: Vukadin, Davor, et al.
Publicado: (2025)
Relevance-driven Input Dropout: an Explanation-guided Regularization Technique
por: Gururaj, Shreyas, et al.
Publicado: (2025)
por: Gururaj, Shreyas, et al.
Publicado: (2025)
Training Artificial Neural Networks by Coordinate Search Algorithm
por: Rokhsatyazdi, Ehsan, et al.
Publicado: (2024)
por: Rokhsatyazdi, Ehsan, et al.
Publicado: (2024)
Conservative Bias in Multi-Teacher Learning: Why Agents Prefer Low-Reward Advisors
por: Mesto, Maher, et al.
Publicado: (2025)
por: Mesto, Maher, et al.
Publicado: (2025)
Teacher-Student Guided Inverse Modeling for Steel Final Hardness Estimation
por: Alsheikh, Ahmad, et al.
Publicado: (2025)
por: Alsheikh, Ahmad, et al.
Publicado: (2025)
Transformer Mechanisms Mimic Frontostriatal Gating Operations When Trained on Human Working Memory Tasks
por: Traylor, Aaron, et al.
Publicado: (2024)
por: Traylor, Aaron, et al.
Publicado: (2024)
Exploring the Use of ChatGPT for a Systematic Literature Review: a Design-Based Research
por: Huang, Qian, et al.
Publicado: (2024)
por: Huang, Qian, et al.
Publicado: (2024)
Inference Time Causal Probing in LLMs
por: Khorasani, Sadegh, et al.
Publicado: (2026)
por: Khorasani, Sadegh, et al.
Publicado: (2026)
Comprehensive Metapath-based Heterogeneous Graph Transformer for Gene-Disease Association Prediction
por: Cui, Wentao, et al.
Publicado: (2025)
por: Cui, Wentao, et al.
Publicado: (2025)
Scalable and Robust LLM Unlearning by Correcting Responses with Retrieved Exclusions
por: Kim, Junbeom, et al.
Publicado: (2025)
por: Kim, Junbeom, et al.
Publicado: (2025)
Collaborative Multi-Agent Scripts Generation for Enhancing Imperfect-Information Reasoning in Murder Mystery Games
por: Zhong, Keyang, et al.
Publicado: (2026)
por: Zhong, Keyang, et al.
Publicado: (2026)
AI Benchmark Democratization and Carpentry
por: von Laszewski, Gregor, et al.
Publicado: (2025)
por: von Laszewski, Gregor, et al.
Publicado: (2025)
Embedded Safety-Aligned Intelligence via Differentiable Internal Alignment Embeddings
por: Rathva, Harsh, et al.
Publicado: (2025)
por: Rathva, Harsh, et al.
Publicado: (2025)
OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models
por: Cotti, Luca, et al.
Publicado: (2025)
por: Cotti, Luca, et al.
Publicado: (2025)
Training Language Models to Win Debates with Self-Play Improves Judge Accuracy
por: Arnesen, Samuel, et al.
Publicado: (2024)
por: Arnesen, Samuel, et al.
Publicado: (2024)
Sketch Decompositions for Classical Planning via Deep Reinforcement Learning
por: Aichmüller, Michael, et al.
Publicado: (2024)
por: Aichmüller, Michael, et al.
Publicado: (2024)
Planning vs Reasoning: Ablations to Test Capabilities of LoRA layers
por: Redkar, Neel
Publicado: (2024)
por: Redkar, Neel
Publicado: (2024)
Improving Agent Behaviors with RL Fine-tuning for Autonomous Driving
por: Peng, Zhenghao, et al.
Publicado: (2024)
por: Peng, Zhenghao, et al.
Publicado: (2024)
Learning to Select Goals in Automated Planning with Deep-Q Learning
por: Núñez-Molina, Carlos, et al.
Publicado: (2024)
por: Núñez-Molina, Carlos, et al.
Publicado: (2024)
LeanProgress: Guiding Search for Neural Theorem Proving via Proof Progress Prediction
por: George, Robert Joseph, et al.
Publicado: (2025)
por: George, Robert Joseph, et al.
Publicado: (2025)
GammaZero: Learning To Guide POMDP Belief Space Search With Graph Representations
por: Mangannavar, Rajesh, et al.
Publicado: (2025)
por: Mangannavar, Rajesh, et al.
Publicado: (2025)
Automated CAD Modeling Sequence Generation from Text Descriptions via Transformer-Based Large Language Models
por: Liao, Jianxing, et al.
Publicado: (2025)
por: Liao, Jianxing, et al.
Publicado: (2025)
JURY-RL: Votes Propose, Proofs Dispose for Label-Free RLVR
por: Chen, Xinjie, et al.
Publicado: (2026)
por: Chen, Xinjie, et al.
Publicado: (2026)
OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind
por: Srishty, Sharmin Sultana, et al.
Publicado: (2026)
por: Srishty, Sharmin Sultana, et al.
Publicado: (2026)
A Domain-Independent Agent Architecture for Adaptive Operation in Evolving Open Worlds
por: Mohan, Shiwali, et al.
Publicado: (2023)
por: Mohan, Shiwali, et al.
Publicado: (2023)
ARCTraj: A Dataset and Benchmark of Human Reasoning Trajectories for Abstract Problem Solving
por: Kim, Sejin, et al.
Publicado: (2025)
por: Kim, Sejin, et al.
Publicado: (2025)
BitCal-TTS: Bit-Calibrated Test-Time Scaling for Quantized Reasoning Models
por: Patarlapalli, Sai Babu, et al.
Publicado: (2026)
por: Patarlapalli, Sai Babu, et al.
Publicado: (2026)
EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique
por: Zhu, Chenglin, et al.
Publicado: (2025)
por: Zhu, Chenglin, et al.
Publicado: (2025)
A Parallel Hybrid Action Space Reinforcement Learning Model for Real-world Adaptive Traffic Signal Control
por: Wang, Yuxuan, et al.
Publicado: (2025)
por: Wang, Yuxuan, et al.
Publicado: (2025)
Position Paper: Bounded Alignment: What (Not) To Expect From AGI Agents
por: Minai, Ali A.
Publicado: (2025)
por: Minai, Ali A.
Publicado: (2025)
Robust and Diverse Multi-Agent Learning via Rational Policy Gradient
por: Lauffer, Niklas, et al.
Publicado: (2025)
por: Lauffer, Niklas, et al.
Publicado: (2025)
Beyond Mimicry: Preference Coherence in LLMs
por: Mikaelson, Luhan, et al.
Publicado: (2025)
por: Mikaelson, Luhan, et al.
Publicado: (2025)
Ejemplares similares
-
Novel Approaches to Artificial Intelligence Development Based on the Nearest Neighbor Method
por: Priezzhev, I. I., et al.
Publicado: (2025) -
Data Ethics in the Era of Healthcare Artificial Intelligence in Africa: An Ubuntu Philosophy Perspective
por: Mahamadou, Abdoul Jalil Djiberou, et al.
Publicado: (2024) -
Foundational Requirements for Artificial General Intelligence: A Falsifiable Framework Based on Signal Prediction
por: Šprogar, Matej
Publicado: (2025) -
Intervention Complexity as a Canonical Reward and a Measure of Intelligence
por: McCane, Brendan
Publicado: (2026) -
A Study on the Application of Artificial Intelligence in Ecological Design
por: Zhao, Hengyue
Publicado: (2025)