Saved in:
| Main Author: | Ollagnier, Anaïs |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2507.20614 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Addressing Antisocial Behavior in Multi-Party Dialogs Through Multimodal Representation Learning
by: Bakarou, Hajar, et al.
Published: (2025)
by: Bakarou, Hajar, et al.
Published: (2025)
DataLens: Enhancing Dataset Discovery via Network Topologies
by: Ollagnier, Anaïs, et al.
Published: (2025)
by: Ollagnier, Anaïs, et al.
Published: (2025)
"Dark Triad" Model Organisms of Misalignment: Narrow Fine-Tuning Mirrors Human Antisocial Behavior
by: Lulla, Roshni, et al.
Published: (2026)
by: Lulla, Roshni, et al.
Published: (2026)
From Data to Behavior: Predicting Unintended Model Behaviors Before Training
by: Wang, Mengru, et al.
Published: (2026)
by: Wang, Mengru, et al.
Published: (2026)
Outraged AI: Large language models prioritise emotion over cost in fairness enforcement
by: Liu, Hao, et al.
Published: (2025)
by: Liu, Hao, et al.
Published: (2025)
Moral Outrage Shapes Commitments Beyond Attention: Multimodal Moral Emotions on YouTube in Korea and the US
by: Park, Seongchan, et al.
Published: (2026)
by: Park, Seongchan, et al.
Published: (2026)
Understanding Fanchuan in Livestreaming Platforms: A New Form of Online Antisocial Behavior
by: Wei, Yiluo, et al.
Published: (2025)
by: Wei, Yiluo, et al.
Published: (2025)
Predicting Movie Hits Before They Happen with LLMs
by: Agah, Shaghayegh, et al.
Published: (2025)
by: Agah, Shaghayegh, et al.
Published: (2025)
Exploitation Without Deception: Dark Triad Feature Steering Reveals Separable Antisocial Circuits in Language Models
by: Berg, Cameron, et al.
Published: (2026)
by: Berg, Cameron, et al.
Published: (2026)
When Safety Fails Before the Answer: Benchmarking Harmful Behavior Detection in Reasoning Chains
by: Kakkar, Ishita, et al.
Published: (2026)
by: Kakkar, Ishita, et al.
Published: (2026)
Online-PVLM: Advancing Personalized VLMs with Online Concept Learning
by: Bai, Huiyu, et al.
Published: (2025)
by: Bai, Huiyu, et al.
Published: (2025)
Think Before Refusal : Triggering Safety Reflection in LLMs to Mitigate False Refusal Behavior
by: Si, Shengyun, et al.
Published: (2025)
by: Si, Shengyun, et al.
Published: (2025)
Before It's Too Late: A State Space Model for the Early Prediction of Misinformation and Disinformation Engagement
by: Tian, Lin, et al.
Published: (2025)
by: Tian, Lin, et al.
Published: (2025)
From Input Perception to Predictive Insight: Modeling Model Blind Spots Before They Become Errors
by: Mi, Maggie, et al.
Published: (2025)
by: Mi, Maggie, et al.
Published: (2025)
Knowing Before Saying: LLM Representations Encode Information About Chain-of-Thought Success Before Completion
by: Afzal, Anum, et al.
Published: (2025)
by: Afzal, Anum, et al.
Published: (2025)
Simulated Ignorance Fails: A Systematic Study of LLM Behaviors on Forecasting Problems Before Model Knowledge Cutoff
by: Li, Zehan, et al.
Published: (2026)
by: Li, Zehan, et al.
Published: (2026)
Can We Predict Before Executing Machine Learning Agents?
by: Zheng, Jingsheng, et al.
Published: (2026)
by: Zheng, Jingsheng, et al.
Published: (2026)
The Technology of Outrage: Bias in Artificial Intelligence
by: Bridewell, Will, et al.
Published: (2024)
by: Bridewell, Will, et al.
Published: (2024)
Measuring and Forecasting Conversation Incivility: the Role of Antisocial and Prosocial Behaviors
by: Yu, Xinchen, et al.
Published: (2024)
by: Yu, Xinchen, et al.
Published: (2024)
Can LLM Agents Simulate Multi-Turn Human Behavior? Evidence from Real Online Customer Behavior Data
by: Lu, Yuxuan, et al.
Published: (2025)
by: Lu, Yuxuan, et al.
Published: (2025)
Leveraging Large Language Models for NLG Evaluation: Advances and Challenges
by: Li, Zhen, et al.
Published: (2024)
by: Li, Zhen, et al.
Published: (2024)
O1 Embedder: Let Retrievers Think Before Action
by: Yan, Ruiran, et al.
Published: (2025)
by: Yan, Ruiran, et al.
Published: (2025)
Falcon-UI: Understanding GUI Before Following User Instructions
by: Shen, Huawen, et al.
Published: (2024)
by: Shen, Huawen, et al.
Published: (2024)
Outrage
Published: (2019)
Published: (2019)
A Survey on Online User Aggression: Content Detection and Behavioral Analysis on Social Media
by: Mane, Swapnil, et al.
Published: (2023)
by: Mane, Swapnil, et al.
Published: (2023)
Machine Behavior in Relational Moral Dilemmas: Moral Rightness, Predicted Human Behavior, and Model Decisions
by: Kim, Jiseon, et al.
Published: (2026)
by: Kim, Jiseon, et al.
Published: (2026)
Temporal Dynamics of Coordinated Online Behavior: Stability, Archetypes, and Influence
by: Tardelli, Serena, et al.
Published: (2023)
by: Tardelli, Serena, et al.
Published: (2023)
D.Va: Validate Your Demonstration First Before You Use It
by: Zhang, Qi, et al.
Published: (2025)
by: Zhang, Qi, et al.
Published: (2025)
Negative Before Positive: Asymmetric Valence Processing in Large Language Models
by: Venkatesh, Sohan
Published: (2026)
by: Venkatesh, Sohan
Published: (2026)
Before and After Temperature: A Distributional View of Creative LLM Generation
by: Parupudi, V. S. Raghu, et al.
Published: (2026)
by: Parupudi, V. S. Raghu, et al.
Published: (2026)
The Real, the Better: Aligning Large Language Models with Online Human Behaviors
by: Jiang, Guanying, et al.
Published: (2024)
by: Jiang, Guanying, et al.
Published: (2024)
Antisocial behavior towards large language model users: experimental evidence
by: Niszczota, Paweł, et al.
Published: (2026)
by: Niszczota, Paweł, et al.
Published: (2026)
Beyond Idealized Patients: Evaluating LLMs under Challenging Patient Behaviors in Medical Consultations
by: Li, Yahan, et al.
Published: (2026)
by: Li, Yahan, et al.
Published: (2026)
Can We Predict Alignment Before Models Finish Thinking? Towards Monitoring Misaligned Reasoning Models
by: Chan, Yik Siu, et al.
Published: (2025)
by: Chan, Yik Siu, et al.
Published: (2025)
The Branch Not Taken: Predicting Branching in Online Conversations
by: Meital, Shai, et al.
Published: (2024)
by: Meital, Shai, et al.
Published: (2024)
Diffusion Language Models Know the Answer Before Decoding
by: Li, Pengxiang, et al.
Published: (2025)
by: Li, Pengxiang, et al.
Published: (2025)
Chip-Tuning: Classify Before Language Models Say
by: Zhu, Fangwei, et al.
Published: (2024)
by: Zhu, Fangwei, et al.
Published: (2024)
Thinking Before You Speak: A Proactive Test-time Scaling Approach
by: Liu, Cong, et al.
Published: (2025)
by: Liu, Cong, et al.
Published: (2025)
Walk Before You Run! Concise LLM Reasoning via Reinforcement Learning
by: Song, Mingyang, et al.
Published: (2025)
by: Song, Mingyang, et al.
Published: (2025)
Understanding Before Reasoning: Enhancing Chain-of-Thought with Iterative Summarization Pre-Prompting
by: Zhu, Dong-Hai, et al.
Published: (2025)
by: Zhu, Dong-Hai, et al.
Published: (2025)
Similar Items
-
Addressing Antisocial Behavior in Multi-Party Dialogs Through Multimodal Representation Learning
by: Bakarou, Hajar, et al.
Published: (2025) -
DataLens: Enhancing Dataset Discovery via Network Topologies
by: Ollagnier, Anaïs, et al.
Published: (2025) -
"Dark Triad" Model Organisms of Misalignment: Narrow Fine-Tuning Mirrors Human Antisocial Behavior
by: Lulla, Roshni, et al.
Published: (2026) -
From Data to Behavior: Predicting Unintended Model Behaviors Before Training
by: Wang, Mengru, et al.
Published: (2026) -
Outraged AI: Large language models prioritise emotion over cost in fairness enforcement
by: Liu, Hao, et al.
Published: (2025)