Saved in:
| Main Authors: | Rosenfeld, Katherine A., Sonnewald, Maike, Jindal, Sonia J., McCarthy, Kevin A., Proctor, Joshua L. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2407.12812 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Importance of Architecture Choice in Deep Learning for Climate Applications
by: Dräger, Simon, et al.
Published: (2024)
by: Dräger, Simon, et al.
Published: (2024)
A Moonshot for AI Oracles in the Sciences
by: Kaiser, Bryan, et al.
Published: (2024)
by: Kaiser, Bryan, et al.
Published: (2024)
PAME-AI: Patient Messaging Creation and Optimization using Agentic AI
by: Luo, Junjie, et al.
Published: (2025)
by: Luo, Junjie, et al.
Published: (2025)
REVEAL: Multi-turn Evaluation of Image-Input Harms for Vision LLM
by: Jindal, Madhur, et al.
Published: (2025)
by: Jindal, Madhur, et al.
Published: (2025)
Towards Understanding Specification Gaming in Reasoning Models
by: Nishimura-Gasparian, Kei, et al.
Published: (2026)
by: Nishimura-Gasparian, Kei, et al.
Published: (2026)
Why Chain of Thought Fails in Clinical Text Understanding
by: Wu, Jiageng, et al.
Published: (2025)
by: Wu, Jiageng, et al.
Published: (2025)
Positive and Risky Message Assessment for Music Products
by: Zhang, Yigeng, et al.
Published: (2023)
by: Zhang, Yigeng, et al.
Published: (2023)
Applying Natural Language Processing and Hierarchical Machine Learning Approaches to Text Difficulty Classification
by: Balyan, Renu, et al.
Published: (2020)
by: Balyan, Renu, et al.
Published: (2020)
Applying Natural Language Processing and Hierarchical Machine Learning Approaches to Text Difficulty Classification
by: Balyan, Renu, et al.
Published: (2020)
by: Balyan, Renu, et al.
Published: (2020)
PolicyBank: Evolving Policy Understanding for LLM Agents
by: Choi, Jihye, et al.
Published: (2026)
by: Choi, Jihye, et al.
Published: (2026)
AI Does Not Alter Perceptions of Text Messages
by: Diamond, N'yoma
Published: (2024)
by: Diamond, N'yoma
Published: (2024)
Early Signs of Steganographic Capabilities in Frontier LLMs
by: Zolkowski, Artur, et al.
Published: (2025)
by: Zolkowski, Artur, et al.
Published: (2025)
Cause and Effect: Can Large Language Models Truly Understand Causality?
by: Ashwani, Swagata, et al.
Published: (2024)
by: Ashwani, Swagata, et al.
Published: (2024)
The Use of AI Tools to Develop and Validate Q-Matrices
by: Fan, Kevin, et al.
Published: (2026)
by: Fan, Kevin, et al.
Published: (2026)
A Framework for Human Evaluation of Large Language Models in Healthcare Derived from Literature Review
by: Tam, Thomas Yu Chow, et al.
Published: (2024)
by: Tam, Thomas Yu Chow, et al.
Published: (2024)
DaVinci at SemEval-2024 Task 9: Few-shot prompting GPT-3.5 for Unconventional Reasoning
by: Mathur, Suyash Vardhan, et al.
Published: (2024)
by: Mathur, Suyash Vardhan, et al.
Published: (2024)
Generating Effective Ensembles for Sentiment Analysis
by: Etelis, Itay, et al.
Published: (2024)
by: Etelis, Itay, et al.
Published: (2024)
Language models for longitudinal analysis of abusive content in Billboard Music Charts
by: Chandra, Rohitash, et al.
Published: (2025)
by: Chandra, Rohitash, et al.
Published: (2025)
Script-Based Dialog Policy Planning for LLM-Powered Conversational Agents: A Basic Architecture for an "AI Therapist"
by: Wasenmüller, Robert, et al.
Published: (2024)
by: Wasenmüller, Robert, et al.
Published: (2024)
Shoot First, Ask Questions Later? Building Rational Agents that Explore and Act Like People
by: Grand, Gabriel, et al.
Published: (2025)
by: Grand, Gabriel, et al.
Published: (2025)
Building Trust in Conversational AI: A Comprehensive Review and Solution Architecture for Explainable, Privacy-Aware Systems using LLMs and Knowledge Graph
by: Zafar, Ahtsham, et al.
Published: (2023)
by: Zafar, Ahtsham, et al.
Published: (2023)
MuSaG: A Multimodal German Sarcasm Dataset with Full-Modal Annotations
by: Scott, Aaron, et al.
Published: (2025)
by: Scott, Aaron, et al.
Published: (2025)
Multimodal Proposal for an AI-Based Tool to Increase Cross-Assessment of Messages
by: Castro, Alejandro Álvarez, et al.
Published: (2025)
by: Castro, Alejandro Álvarez, et al.
Published: (2025)
Building AI Agents to Improve Job Referral Requests to Strangers
by: Chu, Ross, et al.
Published: (2025)
by: Chu, Ross, et al.
Published: (2025)
The Path of Least Resistance: Guiding LLM Reasoning Trajectories with Prefix Consensus
by: Jindal, Ishan, et al.
Published: (2026)
by: Jindal, Ishan, et al.
Published: (2026)
Human-AI Interaction and User Satisfaction: Empirical Evidence from Online Reviews of AI Products
by: Pasch, Stefan, et al.
Published: (2025)
by: Pasch, Stefan, et al.
Published: (2025)
WorldCoder, a Model-Based LLM Agent: Building World Models by Writing Code and Interacting with the Environment
by: Tang, Hao, et al.
Published: (2024)
by: Tang, Hao, et al.
Published: (2024)
Understanding Understanding: A Pragmatic Framework Motivated by Large Language Models
by: Leyton-Brown, Kevin, et al.
Published: (2024)
by: Leyton-Brown, Kevin, et al.
Published: (2024)
Policy-as-Prompt: Turning AI Governance Rules into Guardrails for AI Agents
by: Kholkar, Gauri, et al.
Published: (2025)
by: Kholkar, Gauri, et al.
Published: (2025)
Deep sequence models tend to memorize geometrically; it is unclear why
by: Noroozizadeh, Shahriar, et al.
Published: (2025)
by: Noroozizadeh, Shahriar, et al.
Published: (2025)
The Double Contingency Problem: AI Recursion and the Limits of Interspecies Understanding
by: Bishop, Graham L.
Published: (2025)
by: Bishop, Graham L.
Published: (2025)
Building Production-Ready Probes For Gemini
by: Kramár, János, et al.
Published: (2026)
by: Kramár, János, et al.
Published: (2026)
Evidence-Augmented Policy Optimization with Reward Co-Evolution for Long-Context Reasoning
by: Guan, Xin, et al.
Published: (2026)
by: Guan, Xin, et al.
Published: (2026)
ReviewEval: An Evaluation Framework for AI-Generated Reviews
by: Garg, Madhav Krishan, et al.
Published: (2025)
by: Garg, Madhav Krishan, et al.
Published: (2025)
Provable Coordination for LLM Agents via Message Sequence Charts
by: Bollig, Benedikt, et al.
Published: (2026)
by: Bollig, Benedikt, et al.
Published: (2026)
Citation: A Key to Building Responsible and Accountable Large Language Models
by: Huang, Jie, et al.
Published: (2023)
by: Huang, Jie, et al.
Published: (2023)
LP-LM: No Hallucinations in Question Answering with Logic Programming
by: Wu, Katherine, et al.
Published: (2025)
by: Wu, Katherine, et al.
Published: (2025)
mrCAD: Multimodal Refinement of Computer-aided Designs
by: McCarthy, William P., et al.
Published: (2025)
by: McCarthy, William P., et al.
Published: (2025)
Optimizing Large Language Models for Detecting Symptoms of Comorbid Depression or Anxiety in Chronic Diseases: Insights from Patient Messages
by: Kim, Jiyeong, et al.
Published: (2025)
by: Kim, Jiyeong, et al.
Published: (2025)
"Understanding AI": Semantic Grounding in Large Language Models
by: Lyre, Holger
Published: (2024)
by: Lyre, Holger
Published: (2024)
Similar Items
-
The Importance of Architecture Choice in Deep Learning for Climate Applications
by: Dräger, Simon, et al.
Published: (2024) -
A Moonshot for AI Oracles in the Sciences
by: Kaiser, Bryan, et al.
Published: (2024) -
PAME-AI: Patient Messaging Creation and Optimization using Agentic AI
by: Luo, Junjie, et al.
Published: (2025) -
REVEAL: Multi-turn Evaluation of Image-Input Harms for Vision LLM
by: Jindal, Madhur, et al.
Published: (2025) -
Towards Understanding Specification Gaming in Reasoning Models
by: Nishimura-Gasparian, Kei, et al.
Published: (2026)