Salvato in:
| Autore principale: | Tlaie, Alejandro |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2410.19749 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Large language models in medicine: the potentials and pitfalls
di: Omiye, Jesutofunmi A., et al.
Pubblicazione: (2023)
di: Omiye, Jesutofunmi A., et al.
Pubblicazione: (2023)
Exploring and steering the moral compass of Large Language Models
di: Tlaie, Alejandro
Pubblicazione: (2024)
di: Tlaie, Alejandro
Pubblicazione: (2024)
Promises and pitfalls of artificial intelligence for legal applications
di: Kapoor, Sayash, et al.
Pubblicazione: (2024)
di: Kapoor, Sayash, et al.
Pubblicazione: (2024)
Rules, Cases, and Reasoning: Positivist Legal Theory as a Framework for Pluralistic AI Alignment
di: Caputo, Nicholas A.
Pubblicazione: (2024)
di: Caputo, Nicholas A.
Pubblicazione: (2024)
Securing External Deeper-than-black-box GPAI Evaluations
di: Tlaie, Alejandro, et al.
Pubblicazione: (2025)
di: Tlaie, Alejandro, et al.
Pubblicazione: (2025)
The AI Alignment Paradox
di: West, Robert, et al.
Pubblicazione: (2024)
di: West, Robert, et al.
Pubblicazione: (2024)
Rethinking AI Cultural Alignment
di: Bravansky, Michal, et al.
Pubblicazione: (2025)
di: Bravansky, Michal, et al.
Pubblicazione: (2025)
Justifications for Democratizing AI Alignment and Their Prospects
di: Steingrüber, André, et al.
Pubblicazione: (2025)
di: Steingrüber, André, et al.
Pubblicazione: (2025)
AI and Social Theory
di: Mokander, Jakob, et al.
Pubblicazione: (2024)
di: Mokander, Jakob, et al.
Pubblicazione: (2024)
Understanding the Process of Human-AI Value Alignment
di: McKinlay, Jack, et al.
Pubblicazione: (2025)
di: McKinlay, Jack, et al.
Pubblicazione: (2025)
Privacy Ethics Alignment in AI: A Stakeholder-Centric Framework for Ethical AI
di: Barthwal, Ankur, et al.
Pubblicazione: (2025)
di: Barthwal, Ankur, et al.
Pubblicazione: (2025)
Whose Truth? Pluralistic Geo-Alignment for (Agentic) AI
di: Janowicz, Krzysztof, et al.
Pubblicazione: (2025)
di: Janowicz, Krzysztof, et al.
Pubblicazione: (2025)
Addressing the regulatory gap: moving towards an EU AI audit ecosystem beyond the AI Act by including civil society
di: Hartmann, David, et al.
Pubblicazione: (2024)
di: Hartmann, David, et al.
Pubblicazione: (2024)
Security, privacy, and agentic AI in a regulatory view: From definitions and distinctions to provisions and reflections
di: Zhang, Shiliang, et al.
Pubblicazione: (2026)
di: Zhang, Shiliang, et al.
Pubblicazione: (2026)
Building the ethical AI framework of the future: from philosophy to practice
di: Catapang, Jasper Kyle
Pubblicazione: (2026)
di: Catapang, Jasper Kyle
Pubblicazione: (2026)
The Value of Gen-AI Conversations: A bottom-up Framework for AI Value Alignment
di: Motnikar, Lenart, et al.
Pubblicazione: (2025)
di: Motnikar, Lenart, et al.
Pubblicazione: (2025)
AI and Human Oversight: A Risk-Based Framework for Alignment
di: Kandikatla, Laxmiraju, et al.
Pubblicazione: (2025)
di: Kandikatla, Laxmiraju, et al.
Pubblicazione: (2025)
AI threats to national security can be countered through an incident regime
di: Ortega, Alejandro
Pubblicazione: (2025)
di: Ortega, Alejandro
Pubblicazione: (2025)
AI Alignment at Your Discretion
di: Buyl, Maarten, et al.
Pubblicazione: (2025)
di: Buyl, Maarten, et al.
Pubblicazione: (2025)
Alignment, Agency and Autonomy in Frontier AI: A Systems Engineering Perspective
di: Tallam, Krti
Pubblicazione: (2025)
di: Tallam, Krti
Pubblicazione: (2025)
Representative Social Choice: From Learning Theory to AI Alignment
di: Qiu, Tianyi
Pubblicazione: (2024)
di: Qiu, Tianyi
Pubblicazione: (2024)
Characterizing AI Agents for Alignment and Governance
di: Kasirzadeh, Atoosa, et al.
Pubblicazione: (2025)
di: Kasirzadeh, Atoosa, et al.
Pubblicazione: (2025)
AI for bureaucratic productivity: Measuring the potential of AI to help automate 143 million UK government transactions
di: Straub, Vincent J., et al.
Pubblicazione: (2024)
di: Straub, Vincent J., et al.
Pubblicazione: (2024)
Alignment Is Not Enough: A Relational Framework for Moral Standing in Human-AI Interaction
di: Pasandi, Faezeh B., et al.
Pubblicazione: (2026)
di: Pasandi, Faezeh B., et al.
Pubblicazione: (2026)
Wide Reflective Equilibrium in LLM Alignment: Bridging Moral Epistemology and AI Safety
di: Brophy, Matthew
Pubblicazione: (2025)
di: Brophy, Matthew
Pubblicazione: (2025)
Large language models eroding science understanding: an experimental study
di: Collins, Harry, et al.
Pubblicazione: (2026)
di: Collins, Harry, et al.
Pubblicazione: (2026)
ELEPHANT: Measuring and understanding social sycophancy in LLMs
di: Cheng, Myra, et al.
Pubblicazione: (2025)
di: Cheng, Myra, et al.
Pubblicazione: (2025)
Teacher agency in the age of generative AI: towards a framework of hybrid intelligence for learning design
di: Frøsig, Thomas B, et al.
Pubblicazione: (2024)
di: Frøsig, Thomas B, et al.
Pubblicazione: (2024)
AI-Educational Development Loop (AI-EDL): A Conceptual Framework to Bridge AI Capabilities with Classical Educational Theories
di: Yu, Ning, et al.
Pubblicazione: (2025)
di: Yu, Ning, et al.
Pubblicazione: (2025)
The Coming Crisis of Multi-Agent Misalignment: AI Alignment Must Be a Dynamic and Social Process
di: Carichon, Florian, et al.
Pubblicazione: (2025)
di: Carichon, Florian, et al.
Pubblicazione: (2025)
Alignment Debt: The Hidden Work of Making AI Usable
di: Oyemike, Cumi, et al.
Pubblicazione: (2025)
di: Oyemike, Cumi, et al.
Pubblicazione: (2025)
From tools to thieves: Measuring and understanding public perceptions of AI through crowdsourced metaphors
di: Cheng, Myra, et al.
Pubblicazione: (2025)
di: Cheng, Myra, et al.
Pubblicazione: (2025)
Standing on FURM ground -- A framework for evaluating Fair, Useful, and Reliable AI Models in healthcare systems
di: Callahan, Alison, et al.
Pubblicazione: (2024)
di: Callahan, Alison, et al.
Pubblicazione: (2024)
AI Thinking: A framework for rethinking artificial intelligence in practice
di: Newman-Griffis, Denis
Pubblicazione: (2024)
di: Newman-Griffis, Denis
Pubblicazione: (2024)
Balancing Safety and Helpfulness in Healthcare AI Assistants through Iterative Preference Alignment
di: Nghiem, Huy, et al.
Pubblicazione: (2025)
di: Nghiem, Huy, et al.
Pubblicazione: (2025)
Deconstructing Student Perceptions of Generative AI (GenAI) through an Expectancy Value Theory (EVT)-based Instrument
di: Chan, Cecilia Ka Yuk, et al.
Pubblicazione: (2023)
di: Chan, Cecilia Ka Yuk, et al.
Pubblicazione: (2023)
Why AI Alignment Failure Is Structural: Learned Human Interaction Structures and AGI as an Endogenous Evolutionary Shock
di: Sornette, Didier, et al.
Pubblicazione: (2026)
di: Sornette, Didier, et al.
Pubblicazione: (2026)
Risk Alignment in Agentic AI Systems
di: Clatterbuck, Hayley, et al.
Pubblicazione: (2024)
di: Clatterbuck, Hayley, et al.
Pubblicazione: (2024)
Commercial Persuasion in AI-Mediated Conversations
di: Salvi, Francesco, et al.
Pubblicazione: (2026)
di: Salvi, Francesco, et al.
Pubblicazione: (2026)
Bidirectional Human-AI Alignment in Education for Trustworthy Learning Environments
di: Shen, Hua
Pubblicazione: (2025)
di: Shen, Hua
Pubblicazione: (2025)
Documenti analoghi
-
Large language models in medicine: the potentials and pitfalls
di: Omiye, Jesutofunmi A., et al.
Pubblicazione: (2023) -
Exploring and steering the moral compass of Large Language Models
di: Tlaie, Alejandro
Pubblicazione: (2024) -
Promises and pitfalls of artificial intelligence for legal applications
di: Kapoor, Sayash, et al.
Pubblicazione: (2024) -
Rules, Cases, and Reasoning: Positivist Legal Theory as a Framework for Pluralistic AI Alignment
di: Caputo, Nicholas A.
Pubblicazione: (2024) -
Securing External Deeper-than-black-box GPAI Evaluations
di: Tlaie, Alejandro, et al.
Pubblicazione: (2025)