Mecha-nudges for Machines
Fuente:
arXiv
Saved in:
| Main Authors: | Frey, Giulio, Ethayarajh, Kawin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Humanline: Online Alignment as Perceptual Loss
by: Liu, Sijia, et al.
Published: (2025)
by: Liu, Sijia, et al.
Published: (2025)
Understanding Dataset Difficulty with $\mathcal{V}$-Usable Information
by: Ethayarajh, Kawin, et al.
Published: (2021)
by: Ethayarajh, Kawin, et al.
Published: (2021)
KTO: Model Alignment as Prospect Theoretic Optimization
by: Ethayarajh, Kawin, et al.
Published: (2024)
by: Ethayarajh, Kawin, et al.
Published: (2024)
Spot the BlindSpots: Systematic Identification and Quantification of Fine-Grained LLM Biases in Contact Center Summaries
by: Mayilvaghanan, Kawin, et al.
Published: (2025)
by: Mayilvaghanan, Kawin, et al.
Published: (2025)
Anchor Points: Benchmarking Models with Much Fewer Examples
by: Vivek, Rajan, et al.
Published: (2023)
by: Vivek, Rajan, et al.
Published: (2023)
Potential of large language model-powered nudges for promoting daily water and energy conservation
by: Li, Zonghan, et al.
Published: (2025)
by: Li, Zonghan, et al.
Published: (2025)
Data Checklist: On Unit-Testing Datasets with Usable Information
by: Zhang, Heidi C., et al.
Published: (2024)
by: Zhang, Heidi C., et al.
Published: (2024)
Exploiting Edge Features for Transferable Adversarial Attacks in Distributed Machine Learning
by: Rossolini, Giulio, et al.
Published: (2025)
by: Rossolini, Giulio, et al.
Published: (2025)
Assessing the Impact of Packing on Machine Learning-Based Malware Detection and Classification Systems
by: Gibert, Daniel, et al.
Published: (2024)
by: Gibert, Daniel, et al.
Published: (2024)
Finch: Prompt-guided Key-Value Cache Compression
by: Corallo, Giulio, et al.
Published: (2024)
by: Corallo, Giulio, et al.
Published: (2024)
How Worst-Case Are Adversarial Attacks? Linking Adversarial and Perturbation Robustness
by: Rossolini, Giulio
Published: (2026)
by: Rossolini, Giulio
Published: (2026)
Enhancing behavioral nudges with large language model-based iterative personalization: A field experiment on electricity and hot-water conservation
by: Li, Zonghan, et al.
Published: (2026)
by: Li, Zonghan, et al.
Published: (2026)
A Survey of Machine Learning Models and Datasets for the Multi-label Classification of Textual Hate Speech in English
by: Bäumler, Julian, et al.
Published: (2025)
by: Bäumler, Julian, et al.
Published: (2025)
Certified Adversarial Robustness of Machine Learning-based Malware Detectors via (De)Randomized Smoothing
by: Gibert, Daniel, et al.
Published: (2024)
by: Gibert, Daniel, et al.
Published: (2024)
Mixtures of Neural Cellular Automata: A Stochastic Framework for Growth Modelling and Self-Organization
by: Milite, Salvatore, et al.
Published: (2025)
by: Milite, Salvatore, et al.
Published: (2025)
Considerations Influencing Offense-Defense Dynamics From Artificial Intelligence
by: Corsi, Giulio, et al.
Published: (2024)
by: Corsi, Giulio, et al.
Published: (2024)
Parallel Context-of-Experts Decoding for Retrieval Augmented Generation
by: Corallo, Giulio, et al.
Published: (2026)
by: Corallo, Giulio, et al.
Published: (2026)
Prompt-Based Value Steering of Large Language Models
by: Abbo, Giulio Antonio, et al.
Published: (2025)
by: Abbo, Giulio Antonio, et al.
Published: (2025)
Recursive Symbolic Consciousness: A Formal Model of Emergent Intelligence Across Minds and Machines
by: Goudy, Anastasia
Published: (2025)
by: Goudy, Anastasia
Published: (2025)
Compositional Symmetry as Compression: Lie Pseudogroup Structure in Algorithmic Agents
by: Ruffini, Giulio
Published: (2025)
by: Ruffini, Giulio
Published: (2025)
Blue Teaming Function-Calling Agents
by: Dolcetti, Greta, et al.
Published: (2026)
by: Dolcetti, Greta, et al.
Published: (2026)
Towards a Practical Defense against Adversarial Attacks on Deep Learning-based Malware Detectors via Randomized Smoothing
by: Gibert, Daniel, et al.
Published: (2023)
by: Gibert, Daniel, et al.
Published: (2023)
A word association network methodology for evaluating implicit biases in LLMs compared to humans
by: Abramski, Katherine, et al.
Published: (2025)
by: Abramski, Katherine, et al.
Published: (2025)
Increasing the Confidence of Deep Neural Networks by Coverage Analysis
by: Rossolini, Giulio, et al.
Published: (2021)
by: Rossolini, Giulio, et al.
Published: (2021)
Feedback-MPPI: Fast Sampling-Based MPC via Rollout Differentiation -- Adios low-level controllers
by: Belvedere, Tommaso, et al.
Published: (2025)
by: Belvedere, Tommaso, et al.
Published: (2025)
Artificial Intelligence in Brazilian News: A Mixed-Methods Analysis
by: Hernandes, Raphael, et al.
Published: (2024)
by: Hernandes, Raphael, et al.
Published: (2024)
LLMs left, right, and center: Assessing GPT's capabilities to label political bias from web domains
by: Hernandes, Raphael, et al.
Published: (2024)
by: Hernandes, Raphael, et al.
Published: (2024)
Auditing Google's Search Algorithm: Measuring News Diversity Across Brazil, the UK, and the US
by: Hernandes, Raphael, et al.
Published: (2024)
by: Hernandes, Raphael, et al.
Published: (2024)
A Human Behavioral Baseline for Collective Governance in Software Projects
by: Noori, Mobina, et al.
Published: (2025)
by: Noori, Mobina, et al.
Published: (2025)
Towards Assurance of LLM Adversarial Robustness using Ontology-Driven Argumentation
by: Momcilovic, Tomas Bueno, et al.
Published: (2024)
by: Momcilovic, Tomas Bueno, et al.
Published: (2024)
Offline vs. Online Learning in Model-based RL: Lessons for Data Collection Strategies
by: Chen, Jiaqi, et al.
Published: (2025)
by: Chen, Jiaqi, et al.
Published: (2025)
Tabular Data: Is Deep Learning all you need?
by: Zabërgja, Guri, et al.
Published: (2024)
by: Zabërgja, Guri, et al.
Published: (2024)
Leveraging small language models for Text2SPARQL tasks to improve the resilience of AI assistance
by: Brei, Felix, et al.
Published: (2024)
by: Brei, Felix, et al.
Published: (2024)
I Was Blind but Now I See: Implementing Vision-Enabled Dialogue in Social Robots
by: Abbo, Giulio Antonio, et al.
Published: (2023)
by: Abbo, Giulio Antonio, et al.
Published: (2023)
A Robust Defense against Adversarial Attacks on Deep Learning-based Malware Detectors via (De)Randomized Smoothing
by: Gibert, Daniel, et al.
Published: (2024)
by: Gibert, Daniel, et al.
Published: (2024)
The "LLM World of Words" English free association norms generated by large language models
by: Abramski, Katherine, et al.
Published: (2024)
by: Abramski, Katherine, et al.
Published: (2024)
KGQuest: Template-Driven QA Generation from Knowledge Graphs with LLM-Based Refinement
by: Nayab, Sania, et al.
Published: (2025)
by: Nayab, Sania, et al.
Published: (2025)
The Algorithmic Regulator
by: Ruffini, Giulio
Published: (2025)
by: Ruffini, Giulio
Published: (2025)
How do Scaling Laws Apply to Knowledge Graph Engineering Tasks? The Impact of Model Size on Large Language Model Performance
by: Heim, Desiree, et al.
Published: (2025)
by: Heim, Desiree, et al.
Published: (2025)
The DeepLog Neurosymbolic Machine
by: Derkinderen, Vincent, et al.
Published: (2025)
by: Derkinderen, Vincent, et al.
Published: (2025)
Similar Items
-
Humanline: Online Alignment as Perceptual Loss
by: Liu, Sijia, et al.
Published: (2025) -
Understanding Dataset Difficulty with $\mathcal{V}$-Usable Information
by: Ethayarajh, Kawin, et al.
Published: (2021) -
KTO: Model Alignment as Prospect Theoretic Optimization
by: Ethayarajh, Kawin, et al.
Published: (2024) -
Spot the BlindSpots: Systematic Identification and Quantification of Fine-Grained LLM Biases in Contact Center Summaries
by: Mayilvaghanan, Kawin, et al.
Published: (2025) -
Anchor Points: Benchmarking Models with Much Fewer Examples
by: Vivek, Rajan, et al.
Published: (2023)