Taming the Centaur(s) with LAPITHS: a framework for a theoretically grounded interpretation of AI performances
Fuente:
arXiv
Salvato in:
| Autori principali: | Da Pelo, Matteo, Donvito, Alessio, Frongia, Claudio, Salis, Pietro, Lieto, Antonio |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Structural Ranking of the Cognitive Plausibility of Computational Models of Analogy and Metaphors with the Minimal Cognitive Grid
di: Donvito, Alessio, et al.
Pubblicazione: (2026)
di: Donvito, Alessio, et al.
Pubblicazione: (2026)
Effective Generative AI: The Human-Algorithm Centaur
di: Saghafian, Soroush, et al.
Pubblicazione: (2024)
di: Saghafian, Soroush, et al.
Pubblicazione: (2024)
Not Yet AlphaFold for the Mind: Evaluating Centaur as a Synthetic Participant
di: Namazova, Sabrina, et al.
Pubblicazione: (2025)
di: Namazova, Sabrina, et al.
Pubblicazione: (2025)
FAIRGAME: a Framework for AI Agents Bias Recognition using Game Theory
di: Buscemi, Alessio, et al.
Pubblicazione: (2025)
di: Buscemi, Alessio, et al.
Pubblicazione: (2025)
zIA: a GenAI-powered local auntie assists tourists in Italy
di: Cassani, Alexio, et al.
Pubblicazione: (2024)
di: Cassani, Alexio, et al.
Pubblicazione: (2024)
Automated Road Safety: Enhancing Sign and Surface Damage Detection with AI
di: Merolla, Davide, et al.
Pubblicazione: (2024)
di: Merolla, Davide, et al.
Pubblicazione: (2024)
CentaurEval: Benchmarking Human-in-the-Loop Value in Agentic Coding
di: Luo, Hanjun, et al.
Pubblicazione: (2025)
di: Luo, Hanjun, et al.
Pubblicazione: (2025)
Automating the Correctness Assessment of AI-generated Code for Security Contexts
di: Cotroneo, Domenico, et al.
Pubblicazione: (2023)
di: Cotroneo, Domenico, et al.
Pubblicazione: (2023)
Standing on FURM ground -- A framework for evaluating Fair, Useful, and Reliable AI Models in healthcare systems
di: Callahan, Alison, et al.
Pubblicazione: (2024)
di: Callahan, Alison, et al.
Pubblicazione: (2024)
Can LLMs effectively provide game-theoretic-based scenarios for cybersecurity?
di: Proverbio, Daniele, et al.
Pubblicazione: (2025)
di: Proverbio, Daniele, et al.
Pubblicazione: (2025)
Centaur: Robust End-to-End Autonomous Driving with Test-Time Training
di: Sima, Chonghao, et al.
Pubblicazione: (2025)
di: Sima, Chonghao, et al.
Pubblicazione: (2025)
Pure and Physics-Guided Deep Learning Solutions for Spatio-Temporal Groundwater Level Prediction at Arbitrary Locations
di: Salis, Matteo, et al.
Pubblicazione: (2026)
di: Salis, Matteo, et al.
Pubblicazione: (2026)
Aetheria: A multimodal interpretable content safety framework based on multi-agent debate and collaboration
di: He, Yuxiang, et al.
Pubblicazione: (2025)
di: He, Yuxiang, et al.
Pubblicazione: (2025)
A theoretical framework for self-supervised contrastive learning for continuous dependent data
di: Marusov, Alexander, et al.
Pubblicazione: (2025)
di: Marusov, Alexander, et al.
Pubblicazione: (2025)
Legal interpretation and AI: from expert systems to argumentation and LLMs
di: Janeček, Václav, et al.
Pubblicazione: (2026)
di: Janeček, Václav, et al.
Pubblicazione: (2026)
Media and responsible AI governance: a game-theoretic and LLM analysis
di: Balabanova, Nataliya, et al.
Pubblicazione: (2025)
di: Balabanova, Nataliya, et al.
Pubblicazione: (2025)
Taming Uncertainty via Automation: Observing, Analyzing, and Optimizing Agentic AI Systems
di: Moshkovich, Dany, et al.
Pubblicazione: (2025)
di: Moshkovich, Dany, et al.
Pubblicazione: (2025)
WEITS: A Wavelet-enhanced residual framework for interpretable time series forecasting
di: Guo, Ziyou, et al.
Pubblicazione: (2024)
di: Guo, Ziyou, et al.
Pubblicazione: (2024)
An interpretable framework using foundation models for fish sex identification
di: Miao, Zheng, et al.
Pubblicazione: (2026)
di: Miao, Zheng, et al.
Pubblicazione: (2026)
Time Distributed Deep Learning Models for Purely Exogenous Forecasting: Application to Water Table Depth Prediction using Weather Image Time Series
di: Salis, Matteo, et al.
Pubblicazione: (2024)
di: Salis, Matteo, et al.
Pubblicazione: (2024)
Playing games with knowledge: AI-Induced delusions need game theoretic interventions
di: Beaumaster, Will, et al.
Pubblicazione: (2026)
di: Beaumaster, Will, et al.
Pubblicazione: (2026)
ReclAIm: A multi-agent framework for degradation-aware performance tuning of medical imaging AI
di: Tzanis, Eleftherios, et al.
Pubblicazione: (2025)
di: Tzanis, Eleftherios, et al.
Pubblicazione: (2025)
Polynomial Neural Sheaf Diffusion: A Spectral Filtering Approach on Cellular Sheaves
di: Borgi, Alessio, et al.
Pubblicazione: (2025)
di: Borgi, Alessio, et al.
Pubblicazione: (2025)
Sustaining AI safety: Control-theoretic external impossibility, intrinsic necessity, and structural requirements
di: Mazzu, James M.
Pubblicazione: (2026)
di: Mazzu, James M.
Pubblicazione: (2026)
RAE-AR: Taming Autoregressive Models with Representation Autoencoders
di: Yu, Hu, et al.
Pubblicazione: (2026)
di: Yu, Hu, et al.
Pubblicazione: (2026)
AI in a vat: Fundamental limits of efficient world modelling for agent sandboxing and interpretability
di: Rosas, Fernando, et al.
Pubblicazione: (2025)
di: Rosas, Fernando, et al.
Pubblicazione: (2025)
CBEval: A framework for evaluating and interpreting cognitive biases in LLMs
di: Shaikh, Ammar, et al.
Pubblicazione: (2024)
di: Shaikh, Ammar, et al.
Pubblicazione: (2024)
PRISM: Phase-enhanced Radial-based Image Signature Mapping framework for fingerprinting AI-generated images
di: Ricco, Emanuele, et al.
Pubblicazione: (2025)
di: Ricco, Emanuele, et al.
Pubblicazione: (2025)
AI Harmonics: a human-centric and harms severity-adaptive AI risk assessment framework
di: Vei, Sofia, et al.
Pubblicazione: (2025)
di: Vei, Sofia, et al.
Pubblicazione: (2025)
SUDO: a framework for evaluating clinical artificial intelligence systems without ground-truth annotations
di: Kiyasseh, Dani, et al.
Pubblicazione: (2024)
di: Kiyasseh, Dani, et al.
Pubblicazione: (2024)
Human-Centered AI and Autonomy in Robotics: Insights from a Bibliometric Study
di: Casini, Simona, et al.
Pubblicazione: (2025)
di: Casini, Simona, et al.
Pubblicazione: (2025)
Technology as uncharted territory: Contextual integrity and the notion of AI as new ethical ground
di: Mussgnug, Alexander Martin
Pubblicazione: (2024)
di: Mussgnug, Alexander Martin
Pubblicazione: (2024)
Deep Learning tools to support deforestation monitoring in the Ivory Coast using SAR and Optical satellite imagery
di: Sartor, Gabriele, et al.
Pubblicazione: (2024)
di: Sartor, Gabriele, et al.
Pubblicazione: (2024)
Human-Centered Evaluation of RAG outputs: a framework and questionnaire for human-AI collaboration
di: Mangold, Aline, et al.
Pubblicazione: (2025)
di: Mangold, Aline, et al.
Pubblicazione: (2025)
Taming Data Challenges in ML-based Security Tasks Using Generative AI
di: Kanchi, Shravya, et al.
Pubblicazione: (2025)
di: Kanchi, Shravya, et al.
Pubblicazione: (2025)
KAPSO: A Knowledge-grounded framework for Autonomous Program Synthesis and Optimization
di: Nadafian, Alireza, et al.
Pubblicazione: (2026)
di: Nadafian, Alireza, et al.
Pubblicazione: (2026)
An interpretable generative multimodal neuroimaging-genomics framework for decoding Alzheimer's disease
di: Dolci, Giorgio, et al.
Pubblicazione: (2024)
di: Dolci, Giorgio, et al.
Pubblicazione: (2024)
Spiker+: a framework for the generation of efficient Spiking Neural Networks FPGA accelerators for inference at the edge
di: Carpegna, Alessio, et al.
Pubblicazione: (2024)
di: Carpegna, Alessio, et al.
Pubblicazione: (2024)
xAI-Drop: Don't Use What You Cannot Explain
di: De Luca, Vincenzo Marco, et al.
Pubblicazione: (2024)
di: De Luca, Vincenzo Marco, et al.
Pubblicazione: (2024)
A cybersecurity AI agent selection and decision support framework
di: Malatji, Masike
Pubblicazione: (2025)
di: Malatji, Masike
Pubblicazione: (2025)
Documenti analoghi
-
Structural Ranking of the Cognitive Plausibility of Computational Models of Analogy and Metaphors with the Minimal Cognitive Grid
di: Donvito, Alessio, et al.
Pubblicazione: (2026) -
Effective Generative AI: The Human-Algorithm Centaur
di: Saghafian, Soroush, et al.
Pubblicazione: (2024) -
Not Yet AlphaFold for the Mind: Evaluating Centaur as a Synthetic Participant
di: Namazova, Sabrina, et al.
Pubblicazione: (2025) -
FAIRGAME: a Framework for AI Agents Bias Recognition using Game Theory
di: Buscemi, Alessio, et al.
Pubblicazione: (2025) -
zIA: a GenAI-powered local auntie assists tourists in Italy
di: Cassani, Alexio, et al.
Pubblicazione: (2024)