Enregistré dans:
| Auteurs principaux: | Havin, Miriam, Kleinman, Timna Wharton, Koren, Moran, Dover, Yaniv, Goldstein, Ariel |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2503.01844 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Systematic Biases in LLM Simulations of Debates
par: Taubenfeld, Amir, et autres
Publié: (2024)
par: Taubenfeld, Amir, et autres
Publié: (2024)
A closer look at how large language models trust humans: patterns and biases
par: Lerman, Valeria, et autres
Publié: (2025)
par: Lerman, Valeria, et autres
Publié: (2025)
Do Zombies Understand? A Choose-Your-Own-Adventure Exploration of Machine Cognition
par: Goldstein, Ariel, et autres
Publié: (2024)
par: Goldstein, Ariel, et autres
Publié: (2024)
Confident-Knowledge Diversity Drives Human-Human and Human-AI Free Discussion Synergy and Reveals Pure-AI Discussion Shortfalls
par: Sheffer, Tom, et autres
Publié: (2025)
par: Sheffer, Tom, et autres
Publié: (2025)
Tell me who its founders are and I'll tell you what your online community looks like: Online community founders' personality and community attributes
par: Dover, Yaniv, et autres
Publié: (2025)
par: Dover, Yaniv, et autres
Publié: (2025)
Can LLMs Learn Macroeconomic Narratives from Social Media?
par: Gueta, Almog, et autres
Publié: (2024)
par: Gueta, Almog, et autres
Publié: (2024)
Does ChatGPT Have a Mind?
par: Goldstein, Simon, et autres
Publié: (2024)
par: Goldstein, Simon, et autres
Publié: (2024)
Distributional reasoning in LLMs: Parallel reasoning processes in multi-hop reasoning
par: Shalev, Yuval, et autres
Publié: (2024)
par: Shalev, Yuval, et autres
Publié: (2024)
Mind Your Theory: Theory of Mind Goes Deeper Than Reasoning
par: Wagner, Eitan, et autres
Publié: (2024)
par: Wagner, Eitan, et autres
Publié: (2024)
SEPSIS: I Can Catch Your Lies -- A New Paradigm for Deception Detection
par: Rani, Anku, et autres
Publié: (2023)
par: Rani, Anku, et autres
Publié: (2023)
Motivation in Large Language Models
par: Nahum, Omer, et autres
Publié: (2026)
par: Nahum, Omer, et autres
Publié: (2026)
Indications of Belief-Guided Agency and Meta-Cognitive Monitoring in Large Language Models
par: Yalon, Noam Steinmetz, et autres
Publié: (2026)
par: Yalon, Noam Steinmetz, et autres
Publié: (2026)
We Should Separate Memorization from Copyright
par: Haviv, Adi, et autres
Publié: (2026)
par: Haviv, Adi, et autres
Publié: (2026)
Looking Beyond The Top-1: Transformers Determine Top Tokens In Order
par: Lioubashevski, Daria, et autres
Publié: (2024)
par: Lioubashevski, Daria, et autres
Publié: (2024)
Mind Your Format: Towards Consistent Evaluation of In-Context Learning Improvements
par: Voronov, Anton, et autres
Publié: (2024)
par: Voronov, Anton, et autres
Publié: (2024)
Read Your Own Mind: Reasoning Helps Surface Self-Confidence Signals in LLMs
par: Podolak, Jakub, et autres
Publié: (2025)
par: Podolak, Jakub, et autres
Publié: (2025)
Mind Your Moras: Orthography-Aware Error Analysis of Neural Japanese Morphological Generation
par: Zhang, Wen
Publié: (2026)
par: Zhang, Wen
Publié: (2026)
Views Are My Own, but Also Yours: Benchmarking Theory of Mind Using Common Ground
par: Soubki, Adil, et autres
Publié: (2024)
par: Soubki, Adil, et autres
Publié: (2024)
Mind Your Neighbours: Leveraging Analogous Instances for Rhetorical Role Labeling for Legal Documents
par: Santosh, T. Y. S. S, et autres
Publié: (2024)
par: Santosh, T. Y. S. S, et autres
Publié: (2024)
Accelerating Speculative Decoding with Block Diffusion Draft Trees
par: Ringel, Liran, et autres
Publié: (2026)
par: Ringel, Liran, et autres
Publié: (2026)
Can AI Explanations Make You Change Your Mind?
par: Spillner, Laura, et autres
Publié: (2025)
par: Spillner, Laura, et autres
Publié: (2025)
A Nurse is Blue and Elephant is Rugby: Cross Domain Alignment in Large Language Models Reveal Human-like Patterns
par: Yehudai, Asaf, et autres
Publié: (2024)
par: Yehudai, Asaf, et autres
Publié: (2024)
SAUCE: Synchronous and Asynchronous User-Customizable Environment for Multi-Agent LLM Interaction
par: Neuberger, Shlomo, et autres
Publié: (2024)
par: Neuberger, Shlomo, et autres
Publié: (2024)
CoMMET: To What Extent Can LLMs Perform Theory of Mind Tasks?
par: Chen, Ruirui, et autres
Publié: (2026)
par: Chen, Ruirui, et autres
Publié: (2026)
Re:Verse -- Can Your VLM Read a Manga?
par: Baranwal, Aaditya, et autres
Publié: (2025)
par: Baranwal, Aaditya, et autres
Publié: (2025)
Dependency-Guided Parallel Decoding in Discrete Diffusion Language Models
par: Ringel, Liran, et autres
Publié: (2026)
par: Ringel, Liran, et autres
Publié: (2026)
Embedding And Clustering Your Data Can Improve Contrastive Pretraining
par: Merrick, Luke
Publié: (2024)
par: Merrick, Luke
Publié: (2024)
Can Pruning Improve Reasoning? Revisiting Long-CoT Compression with Capability in Mind for Better Reasoning
par: Zhao, Shangziqi, et autres
Publié: (2025)
par: Zhao, Shangziqi, et autres
Publié: (2025)
When Users Change Their Mind: Evaluating Interruptible Agents in Long-Horizon Web Navigation
par: Zou, Henry Peng, et autres
Publié: (2026)
par: Zou, Henry Peng, et autres
Publié: (2026)
RedacBench: Can AI Erase Your Secrets?
par: Jeon, Hyunjun, et autres
Publié: (2026)
par: Jeon, Hyunjun, et autres
Publié: (2026)
Changing Answer Order Can Decrease MMLU Accuracy
par: Gupta, Vipul, et autres
Publié: (2024)
par: Gupta, Vipul, et autres
Publié: (2024)
How Much of Your Data Can Suck? Thresholds for Domain Performance and Emergent Misalignment in LLMs
par: Ouyang, Jian, et autres
Publié: (2025)
par: Ouyang, Jian, et autres
Publié: (2025)
Good Agentic Friends Do Not Just Give Verbal Advice: They Can Update Your Weights
par: Bao, Wenrui, et autres
Publié: (2026)
par: Bao, Wenrui, et autres
Publié: (2026)
Can Your Model Tell a Negation from an Implicature? Unravelling Challenges With Intent Encoders
par: Zhang, Yuwei, et autres
Publié: (2024)
par: Zhang, Yuwei, et autres
Publié: (2024)
Confidence Improves Self-Consistency in LLMs
par: Taubenfeld, Amir, et autres
Publié: (2025)
par: Taubenfeld, Amir, et autres
Publié: (2025)
Draw Your Mind: Personalized Generation via Condition-Level Modeling in Text-to-Image Diffusion Models
par: Kim, Hyungjin, et autres
Publié: (2025)
par: Kim, Hyungjin, et autres
Publié: (2025)
Speak Your Mind: The Speech Continuation Task as a Probe of Voice-Based Model Bias
par: Satish, Shree Harsha Bokkahalli, et autres
Publié: (2025)
par: Satish, Shree Harsha Bokkahalli, et autres
Publié: (2025)
Large Language Models Can Infer Personality from Free-Form User Interactions
par: Peters, Heinrich, et autres
Publié: (2024)
par: Peters, Heinrich, et autres
Publié: (2024)
Learning a Continue-Thinking Token for Enhanced Test-Time Scaling
par: Ringel, Liran, et autres
Publié: (2025)
par: Ringel, Liran, et autres
Publié: (2025)
Segment-Based Attention Masking for GPTs
par: Katz, Shahar, et autres
Publié: (2024)
par: Katz, Shahar, et autres
Publié: (2024)
Documents similaires
-
Systematic Biases in LLM Simulations of Debates
par: Taubenfeld, Amir, et autres
Publié: (2024) -
A closer look at how large language models trust humans: patterns and biases
par: Lerman, Valeria, et autres
Publié: (2025) -
Do Zombies Understand? A Choose-Your-Own-Adventure Exploration of Machine Cognition
par: Goldstein, Ariel, et autres
Publié: (2024) -
Confident-Knowledge Diversity Drives Human-Human and Human-AI Free Discussion Synergy and Reveals Pure-AI Discussion Shortfalls
par: Sheffer, Tom, et autres
Publié: (2025) -
Tell me who its founders are and I'll tell you what your online community looks like: Online community founders' personality and community attributes
par: Dover, Yaniv, et autres
Publié: (2025)