Salvato in:
| Autori principali: | Guo, Wenqi Marshall, Du, Yiyang, Tworek, Heidi J. S., Du, Shan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2509.08833 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Position: Universal Aesthetic Alignment Narrows Artistic Expression
di: Guo, Wenqi Marshall, et al.
Pubblicazione: (2025)
di: Guo, Wenqi Marshall, et al.
Pubblicazione: (2025)
LangGas: Introducing Language in Selective Zero-Shot Background Subtraction for Semi-Transparent Gas Leak Detection with a New Dataset
di: Guo, Wenqi, et al.
Pubblicazione: (2025)
di: Guo, Wenqi, et al.
Pubblicazione: (2025)
Pluralistic Alignment Over Time
di: Klassen, Toryn Q., et al.
Pubblicazione: (2024)
di: Klassen, Toryn Q., et al.
Pubblicazione: (2024)
Mind the Gap: Pitfalls of LLM Alignment with Asian Public Opinion
di: Shankar, Hari, et al.
Pubblicazione: (2026)
di: Shankar, Hari, et al.
Pubblicazione: (2026)
Take Caution in Using LLMs as Human Surrogates: Scylla Ex Machina
di: Gao, Yuan, et al.
Pubblicazione: (2024)
di: Gao, Yuan, et al.
Pubblicazione: (2024)
Evaluating the Promise and Pitfalls of LLMs in Hiring Decisions
di: Anzenberg, Eitan, et al.
Pubblicazione: (2025)
di: Anzenberg, Eitan, et al.
Pubblicazione: (2025)
AI Identity, Empowerment, and Mindfulness in Mitigating Unethical AI Use
di: Shaayesteh, Mayssam Tarighi, et al.
Pubblicazione: (2025)
di: Shaayesteh, Mayssam Tarighi, et al.
Pubblicazione: (2025)
Safe in the Future, Dangerous in the Past: Dissecting Temporal and Linguistic Vulnerabilities in LLMs
di: Said, Muhammad Abdullahi, et al.
Pubblicazione: (2025)
di: Said, Muhammad Abdullahi, et al.
Pubblicazione: (2025)
Navigating Pitfalls: Evaluating LLMs in Machine Learning Programming Education
di: Kumar, Smitha, et al.
Pubblicazione: (2025)
di: Kumar, Smitha, et al.
Pubblicazione: (2025)
The AI Risk Spectrum: From Dangerous Capabilities to Existential Threats
di: Grey, Markov, et al.
Pubblicazione: (2025)
di: Grey, Markov, et al.
Pubblicazione: (2025)
FaceLinkGen: Rethinking Identity Leakage in Privacy-Preserving Face Recognition with Identity Extraction
di: Guo, Wenqi, et al.
Pubblicazione: (2026)
di: Guo, Wenqi, et al.
Pubblicazione: (2026)
TikTok Engagement Traces Over Time and Health Risky Behaviors: Combining Data Linkage and Computational Methods
di: Zhao, Xinyan, et al.
Pubblicazione: (2024)
di: Zhao, Xinyan, et al.
Pubblicazione: (2024)
VSF: Simple, Efficient, and Effective Negative Guidance in Few-Step Image Generation Models By Value Sign Flip
di: Guo, Wenqi, et al.
Pubblicazione: (2025)
di: Guo, Wenqi, et al.
Pubblicazione: (2025)
Toward Responsible and Beneficial AI: Comparing Regulatory and Guidance-Based Approaches -A Comprehensive Comparative Analysis of Artificial Intelligence Governance Frameworks across the European Union, United States, China, and IEEE
di: Du, Jian
Pubblicazione: (2025)
di: Du, Jian
Pubblicazione: (2025)
No Size Fits All: The Perils and Pitfalls of Leveraging LLMs Vary with Company Size
di: Urlana, Ashok, et al.
Pubblicazione: (2024)
di: Urlana, Ashok, et al.
Pubblicazione: (2024)
Dynamic Bayesian Item Response Model with Decomposition (D-BIRD): Modeling Cohort and Individual Learning Over Time
di: Lee, Hansol, et al.
Pubblicazione: (2025)
di: Lee, Hansol, et al.
Pubblicazione: (2025)
Boosting Fairness and Robustness in Over-the-Air Federated Learning
di: Oksuz, Halil Yigit, et al.
Pubblicazione: (2024)
di: Oksuz, Halil Yigit, et al.
Pubblicazione: (2024)
Pitfalls of Evidence-Based AI Policy
di: Casper, Stephen, et al.
Pubblicazione: (2025)
di: Casper, Stephen, et al.
Pubblicazione: (2025)
Exploring Safety Alignment Evaluation of LLMs in Chinese Mental Health Dialogues via LLM-as-Judge
di: Cai, Yunna, et al.
Pubblicazione: (2025)
di: Cai, Yunna, et al.
Pubblicazione: (2025)
Verify with Caution: The Pitfalls of Relying on Imperfect Factuality Metrics
di: Godbole, Ameya, et al.
Pubblicazione: (2025)
di: Godbole, Ameya, et al.
Pubblicazione: (2025)
Faults and Pitfalls in Implementing the Right to be Forgotten
di: Sun, Chen, et al.
Pubblicazione: (2026)
di: Sun, Chen, et al.
Pubblicazione: (2026)
Technical Requirements for Halting Dangerous AI Activities
di: Barnett, Peter, et al.
Pubblicazione: (2025)
di: Barnett, Peter, et al.
Pubblicazione: (2025)
Risks of AI Scientists: Prioritizing Safeguarding Over Autonomy
di: Tang, Xiangru, et al.
Pubblicazione: (2024)
di: Tang, Xiangru, et al.
Pubblicazione: (2024)
Access Over Deception: Fighting Deceptive Patterns through Accessibility
di: Pellkvist, Tobias, et al.
Pubblicazione: (2026)
di: Pellkvist, Tobias, et al.
Pubblicazione: (2026)
Recommendation Fairness in Social Networks Over Time
di: Cao, Meng, et al.
Pubblicazione: (2024)
di: Cao, Meng, et al.
Pubblicazione: (2024)
Large Language Models as Students Who Think Aloud: Overly Coherent, Verbose, and Confident
di: Borchers, Conrad, et al.
Pubblicazione: (2026)
di: Borchers, Conrad, et al.
Pubblicazione: (2026)
The Disintegration of Free Speech
di: Mei, Yiyang
Pubblicazione: (2026)
di: Mei, Yiyang
Pubblicazione: (2026)
Personalized Parsons Puzzles as Scaffolding Enhance Practice Engagement Over Just Showing LLM-Powered Solutions
di: Hou, Xinying, et al.
Pubblicazione: (2025)
di: Hou, Xinying, et al.
Pubblicazione: (2025)
Out of the Loop Again: How Dangerous is Weaponizing Automated Nuclear Systems?
di: Schwartz, Joshua A., et al.
Pubblicazione: (2025)
di: Schwartz, Joshua A., et al.
Pubblicazione: (2025)
Impact of AI Tools on Learning Outcomes: Decreasing Knowledge and Over-Reliance
di: Benedek, Márton, et al.
Pubblicazione: (2025)
di: Benedek, Márton, et al.
Pubblicazione: (2025)
Expanding External Access To Frontier AI Models For Dangerous Capability Evaluations
di: Charnock, Jacob, et al.
Pubblicazione: (2026)
di: Charnock, Jacob, et al.
Pubblicazione: (2026)
Fluent but Foreign: Even Regional LLMs Lack Cultural Alignment
di: Agarwal, Dhruv, et al.
Pubblicazione: (2025)
di: Agarwal, Dhruv, et al.
Pubblicazione: (2025)
Measuring Compliance with the California Consumer Privacy Act Over Space and Time
di: Tran, Van, et al.
Pubblicazione: (2024)
di: Tran, Van, et al.
Pubblicazione: (2024)
Reclaiming Constitutional Authority of Algorithmic Power
di: Mei, Yiyang, et al.
Pubblicazione: (2025)
di: Mei, Yiyang, et al.
Pubblicazione: (2025)
Optimizing Mastery Learning by Fast-Forwarding Over-Practice Steps
di: Xia, Meng, et al.
Pubblicazione: (2025)
di: Xia, Meng, et al.
Pubblicazione: (2025)
Evaluating the Clinical Safety of LLMs in Response to High-Risk Mental Health Disclosures
di: Shah, Siddharth, et al.
Pubblicazione: (2025)
di: Shah, Siddharth, et al.
Pubblicazione: (2025)
Randomness, Not Representation: The Unreliability of Evaluating Cultural Alignment in LLMs
di: Khan, Ariba, et al.
Pubblicazione: (2025)
di: Khan, Ariba, et al.
Pubblicazione: (2025)
ZS-VCOS: Zero-Shot Video Camouflaged Object Segmentation By Optical Flow and Open Vocabulary Object Detection
di: Guo, Wenqi, et al.
Pubblicazione: (2025)
di: Guo, Wenqi, et al.
Pubblicazione: (2025)
An FDA for AI? Pitfalls and Plausibility of Approval Regulation for Frontier Artificial Intelligence
di: Carpenter, Daniel, et al.
Pubblicazione: (2024)
di: Carpenter, Daniel, et al.
Pubblicazione: (2024)
The Fair Game: Auditing & Debiasing AI Algorithms Over Time
di: Basu, Debabrota, et al.
Pubblicazione: (2025)
di: Basu, Debabrota, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Position: Universal Aesthetic Alignment Narrows Artistic Expression
di: Guo, Wenqi Marshall, et al.
Pubblicazione: (2025) -
LangGas: Introducing Language in Selective Zero-Shot Background Subtraction for Semi-Transparent Gas Leak Detection with a New Dataset
di: Guo, Wenqi, et al.
Pubblicazione: (2025) -
Pluralistic Alignment Over Time
di: Klassen, Toryn Q., et al.
Pubblicazione: (2024) -
Mind the Gap: Pitfalls of LLM Alignment with Asian Public Opinion
di: Shankar, Hari, et al.
Pubblicazione: (2026) -
Take Caution in Using LLMs as Human Surrogates: Scylla Ex Machina
di: Gao, Yuan, et al.
Pubblicazione: (2024)