A Checks-and-Balances Framework for Context-Aware Ethical AI Alignment
Fuente:
arXiv
Guardado en:
| Autor principal: | Chang, Edward Y. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
MACI: Multi-Agent Collaborative Intelligence for Adaptive Reasoning and Temporal Planning
por: Chang, Edward Y.
Publicado: (2025)
por: Chang, Edward Y.
Publicado: (2025)
The Empowerment of Science of Science by Large Language Models: New Tools and Methods
por: Liang, Guoqiang, et al.
Publicado: (2025)
por: Liang, Guoqiang, et al.
Publicado: (2025)
The impact of postediting on AI generative translation in Yemeni context: Translating literary prose by ChatGPT
por: Al-wagieh, Nasim, et al.
Publicado: (2026)
por: Al-wagieh, Nasim, et al.
Publicado: (2026)
RepoAgent: An LLM-Powered Open-Source Framework for Repository-level Code Documentation Generation
por: Luo, Qinyu, et al.
Publicado: (2024)
por: Luo, Qinyu, et al.
Publicado: (2024)
Blockchain As a Platform For Artificial Intelligence (AI) Transparency
por: Akther, Afroja, et al.
Publicado: (2025)
por: Akther, Afroja, et al.
Publicado: (2025)
HySem: A context length optimized LLM pipeline for unstructured tabular extraction
por: PP, Narayanan, et al.
Publicado: (2024)
por: PP, Narayanan, et al.
Publicado: (2024)
LLMeBench: A Flexible Framework for Accelerating LLMs Benchmarking
por: Dalvi, Fahim, et al.
Publicado: (2023)
por: Dalvi, Fahim, et al.
Publicado: (2023)
GenAI Content Detection Task 2: AI vs. Human -- Academic Essay Authenticity Challenge
por: Chowdhury, Shammur Absar, et al.
Publicado: (2024)
por: Chowdhury, Shammur Absar, et al.
Publicado: (2024)
LAraBench: Benchmarking Arabic AI with Large Language Models
por: Abdelali, Ahmed, et al.
Publicado: (2023)
por: Abdelali, Ahmed, et al.
Publicado: (2023)
CultranAI at PalmX 2025: Data Augmentation for Cultural Knowledge Representation
por: Bhatti, Hunzalah Hassan, et al.
Publicado: (2025)
por: Bhatti, Hunzalah Hassan, et al.
Publicado: (2025)
NativQA Framework: Enabling LLMs and VLMs with Native, Local, and Everyday Knowledge
por: Alam, Firoj, et al.
Publicado: (2025)
por: Alam, Firoj, et al.
Publicado: (2025)
Evaluating the efficacy of LLM Safety Solutions : The Palit Benchmark Dataset
por: Palit, Sayon, et al.
Publicado: (2025)
por: Palit, Sayon, et al.
Publicado: (2025)
TituLLMs: A Family of Bangla LLMs with Comprehensive Benchmarking
por: Nahin, Shahriar Kabir, et al.
Publicado: (2025)
por: Nahin, Shahriar Kabir, et al.
Publicado: (2025)
Native vs Non-Native Language Prompting: A Comparative Analysis
por: Kmainasi, Mohamed Bayan, et al.
Publicado: (2024)
por: Kmainasi, Mohamed Bayan, et al.
Publicado: (2024)
OASIS: A Multilingual and Multimodal Dataset for Culturally Grounded Spoken Visual QA
por: Alam, Firoj, et al.
Publicado: (2025)
por: Alam, Firoj, et al.
Publicado: (2025)
Propaganda to Hate: A Multimodal Analysis of Arabic Memes with Multi-Agent LLMs
por: Alam, Firoj, et al.
Publicado: (2024)
por: Alam, Firoj, et al.
Publicado: (2024)
A Multiple-Fill-in-the-Blank Exam Approach for Enhancing Zero-Resource Hallucination Detection in Large Language Models
por: Munakata, Satoshi, et al.
Publicado: (2024)
por: Munakata, Satoshi, et al.
Publicado: (2024)
Improved IR-based Bug Localization with Intelligent Relevance Feedback
por: Samir, Asif Mohammed, et al.
Publicado: (2025)
por: Samir, Asif Mohammed, et al.
Publicado: (2025)
Beyond MCQ: An Open-Ended Arabic Cultural QA Benchmark with Dialect Variants
por: Bhatti, Hunzalah Hassan, et al.
Publicado: (2025)
por: Bhatti, Hunzalah Hassan, et al.
Publicado: (2025)
Transcribing Bengali Text with Regional Dialects to IPA using District Guided Tokens
por: Islam, S M Jishanul, et al.
Publicado: (2024)
por: Islam, S M Jishanul, et al.
Publicado: (2024)
LayerRoute: Input-Conditioned Adaptive Layer Skipping via LoRA Fine-Tuning for Agentic Language Models
por: Sikdar, Prateek Kumar
Publicado: (2026)
por: Sikdar, Prateek Kumar
Publicado: (2026)
AraDiCE: Benchmarks for Dialectal and Cultural Capabilities in LLMs
por: Mousi, Basel, et al.
Publicado: (2024)
por: Mousi, Basel, et al.
Publicado: (2024)
Towards robust long-context understanding of large language model via active recap learning
por: Hui, Chenyu
Publicado: (2026)
por: Hui, Chenyu
Publicado: (2026)
LlamaLens: Specialized Multilingual LLM for Analyzing News and Social Media Content
por: Kmainasi, Mohamed Bayan, et al.
Publicado: (2024)
por: Kmainasi, Mohamed Bayan, et al.
Publicado: (2024)
NativQA: Multilingual Culturally-Aligned Natural Query for LLMs
por: Hasan, Md. Arid, et al.
Publicado: (2024)
por: Hasan, Md. Arid, et al.
Publicado: (2024)
ThatiAR: Subjectivity Detection in Arabic News Sentences
por: Suwaileh, Reem, et al.
Publicado: (2024)
por: Suwaileh, Reem, et al.
Publicado: (2024)
Grammatically-Guided Sparse Attention for Efficient and Interpretable Transformers
por: Pratyush, Spandan
Publicado: (2026)
por: Pratyush, Spandan
Publicado: (2026)
Scalability Optimization in Cloud-Based AI Inference Services: Strategies for Real-Time Load Balancing and Automated Scaling
por: Jin, Yihong, et al.
Publicado: (2025)
por: Jin, Yihong, et al.
Publicado: (2025)
On The Role of Intentionality in Knowledge Representation: Analyzing Scene Context for Cognitive Agents with a Tiny Language Model
por: Burgess, Mark
Publicado: (2025)
por: Burgess, Mark
Publicado: (2025)
A Fine-Grained Complexity View on Propositional Abduction -- Algorithms and Lower Bounds
por: Lagerkvist, Victor, et al.
Publicado: (2025)
por: Lagerkvist, Victor, et al.
Publicado: (2025)
FedPF: Accurate Target Privacy Preserving Federated Learning Balancing Fairness and Utility
por: Sun, Kangkang, et al.
Publicado: (2025)
por: Sun, Kangkang, et al.
Publicado: (2025)
Integrating Emotional and Linguistic Models for Ethical Compliance in Large Language Models
por: Chang, Edward Y.
Publicado: (2024)
por: Chang, Edward Y.
Publicado: (2024)
Haptic Repurposing with GenAI
por: Wang, Haoyu
Publicado: (2024)
por: Wang, Haoyu
Publicado: (2024)
StarCraft+: Benchmarking Multi-agent Algorithms in Adversary Paradigm
por: Li, Yadong, et al.
Publicado: (2025)
por: Li, Yadong, et al.
Publicado: (2025)
Advanced Game-Theoretic Frameworks for Multi-Agent AI Challenges: A 2025 Outlook
por: Malinovskiy, Pavel
Publicado: (2025)
por: Malinovskiy, Pavel
Publicado: (2025)
The acquisition of English irregular inflections by Yemeni L1 Arabic learners: A Universal Grammar approach
por: Alsawsh, Muneef Y., et al.
Publicado: (2026)
por: Alsawsh, Muneef Y., et al.
Publicado: (2026)
Infinity Parser: Layout Aware Reinforcement Learning for Scanned Document Parsing
por: Wang, Baode, et al.
Publicado: (2025)
por: Wang, Baode, et al.
Publicado: (2025)
TRIM: Achieving Extreme Sparsity with Targeted Row-wise Iterative Metric-driven Pruning
por: Beck, Florentin, et al.
Publicado: (2025)
por: Beck, Florentin, et al.
Publicado: (2025)
The Architecture of Errors: From Universal Impossibility to Patch-Local LLM Reliability
por: Arbuzov, Mikhail L., et al.
Publicado: (2026)
por: Arbuzov, Mikhail L., et al.
Publicado: (2026)
Detecting Check-Worthy Claims in Political Debates, Speeches, and Interviews Using Audio Data
por: Ivanov, Petar, et al.
Publicado: (2023)
por: Ivanov, Petar, et al.
Publicado: (2023)
Ejemplares similares
-
MACI: Multi-Agent Collaborative Intelligence for Adaptive Reasoning and Temporal Planning
por: Chang, Edward Y.
Publicado: (2025) -
The Empowerment of Science of Science by Large Language Models: New Tools and Methods
por: Liang, Guoqiang, et al.
Publicado: (2025) -
The impact of postediting on AI generative translation in Yemeni context: Translating literary prose by ChatGPT
por: Al-wagieh, Nasim, et al.
Publicado: (2026) -
RepoAgent: An LLM-Powered Open-Source Framework for Repository-level Code Documentation Generation
por: Luo, Qinyu, et al.
Publicado: (2024) -
Blockchain As a Platform For Artificial Intelligence (AI) Transparency
por: Akther, Afroja, et al.
Publicado: (2025)