Saved in:
| Main Authors: | Dutta, Senjuti, Chen, Sherol, Mak, Sunny, Ahmad, Amnah, Collins, Katherine, Butryna, Alena, Ramachandran, Deepak, Dvijotham, Krishnamurthy, Pavlick, Ellie, Rajakumar, Ravi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2403.05576 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Thumbs Up/Down: Untangling Challenges of Fine-Grained Feedback for Text-to-Image Generation
by: Collins, Katherine M., et al.
Published: (2024)
by: Collins, Katherine M., et al.
Published: (2024)
From Prediction to Understanding: Will AI Foundation Models Transform Brain Science?
by: Serre, Thomas, et al.
Published: (2025)
by: Serre, Thomas, et al.
Published: (2025)
Talking Heads: Understanding Inter-layer Communication in Transformer Language Models
by: Merullo, Jack, et al.
Published: (2024)
by: Merullo, Jack, et al.
Published: (2024)
Does Training on Synthetic Data Make Models Less Robust?
by: Zhang, Lingze, et al.
Published: (2025)
by: Zhang, Lingze, et al.
Published: (2025)
How Do Language Models Compose Functions?
by: Khandelwal, Apoorv, et al.
Published: (2025)
by: Khandelwal, Apoorv, et al.
Published: (2025)
Paths Not Taken: Understanding and Mending the Multilingual Factual Recall Pipeline
by: Lu, Meng, et al.
Published: (2025)
by: Lu, Meng, et al.
Published: (2025)
Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG
by: Choudhary, Sarthak, et al.
Published: (2025)
by: Choudhary, Sarthak, et al.
Published: (2025)
Dual Process Learning: Controlling Use of In-Context vs. In-Weights Strategies with Weight Forgetting
by: Anand, Suraj, et al.
Published: (2024)
by: Anand, Suraj, et al.
Published: (2024)
Learning to Receive Help: Intervention-Aware Concept Embedding Models
by: Zarlenga, Mateo Espinosa, et al.
Published: (2023)
by: Zarlenga, Mateo Espinosa, et al.
Published: (2023)
LLMs model how humans induce logically structured rules
by: Loo, Alyssa, et al.
Published: (2025)
by: Loo, Alyssa, et al.
Published: (2025)
mOthello: When Do Cross-Lingual Representation Alignment and Cross-Lingual Transfer Emerge in Multilingual Models?
by: Hua, Tianze, et al.
Published: (2024)
by: Hua, Tianze, et al.
Published: (2024)
What is an "Abstract Reasoner"? Revisiting Experiments and Arguments about Large Language Models
by: Yun, Tian, et al.
Published: (2025)
by: Yun, Tian, et al.
Published: (2025)
Handling and Interpreting Missing Modalities in Patient Clinical Trajectories via Autoregressive Sequence Modeling
by: Wang, Andrew, et al.
Published: (2026)
by: Wang, Andrew, et al.
Published: (2026)
How Do Vision-Language Models Process Conflicting Information Across Modalities?
by: Hua, Tianze, et al.
Published: (2025)
by: Hua, Tianze, et al.
Published: (2025)
A Knapsack by Any Other Name: Presentation impacts LLM performance on NP-hard problems
by: Duchnowski, Alex, et al.
Published: (2025)
by: Duchnowski, Alex, et al.
Published: (2025)
Circuit Component Reuse Across Tasks in Transformer Language Models
by: Merullo, Jack, et al.
Published: (2023)
by: Merullo, Jack, et al.
Published: (2023)
Language Models Implement Simple Word2Vec-style Vector Arithmetic
by: Merullo, Jack, et al.
Published: (2023)
by: Merullo, Jack, et al.
Published: (2023)
Achieving the Tightest Relaxation of Sigmoids for Formal Verification
by: Chevalier, Samuel, et al.
Published: (2024)
by: Chevalier, Samuel, et al.
Published: (2024)
TERRITÓRIOS ÉTNICOS NO PÓS-ABOLIÇÃO: O CASO DO QUILOMBO DA MORMAÇA (RS)
by: Sherol dos Santos
Published: (2009)
by: Sherol dos Santos
Published: (2009)
Instilling Inductive Biases with Subnetworks
by: Zhang, Enyan, et al.
Published: (2023)
by: Zhang, Enyan, et al.
Published: (2023)
The dynamic interplay between in-context and in-weight learning in humans and neural networks
by: Russin, Jacob, et al.
Published: (2024)
by: Russin, Jacob, et al.
Published: (2024)
Uncovering Intermediate Variables in Transformers using Circuit Probing
by: Lepori, Michael A., et al.
Published: (2023)
by: Lepori, Michael A., et al.
Published: (2023)
Source-Modality Monitoring in Vision-Language Models
by: Hua, Etha Tianze, et al.
Published: (2026)
by: Hua, Etha Tianze, et al.
Published: (2026)
Norm-Bounded Low-Rank Adaptation
by: Wang, Ruigang, et al.
Published: (2025)
by: Wang, Ruigang, et al.
Published: (2025)
Monotone, Bi-Lipschitz, and Polyak-Lojasiewicz Networks
by: Wang, Ruigang, et al.
Published: (2024)
by: Wang, Ruigang, et al.
Published: (2024)
Video Finetuning Improves Reasoning Between Frames
by: Yang, Ruiqi, et al.
Published: (2025)
by: Yang, Ruiqi, et al.
Published: (2025)
Transferring Linear Features Across Language Models With Model Stitching
by: Chen, Alan, et al.
Published: (2025)
by: Chen, Alan, et al.
Published: (2025)
LLMs as Models for Analogical Reasoning
by: Musker, Sam, et al.
Published: (2024)
by: Musker, Sam, et al.
Published: (2024)
Equivalence of Almost N ‐Automorphism‐Invariant and Almost N ‐Pseudo‐Injective Modules
by: Amnah A. Alkinani, et al.
Published: (2025)
by: Amnah A. Alkinani, et al.
Published: (2025)
Transformer Mechanisms Mimic Frontostriatal Gating Operations When Trained on Human Working Memory Tasks
by: Traylor, Aaron, et al.
Published: (2024)
by: Traylor, Aaron, et al.
Published: (2024)
Keeping up with dynamic attackers: Certifying robustness to adaptive online data poisoning
by: Bose, Avinandan, et al.
Published: (2025)
by: Bose, Avinandan, et al.
Published: (2025)
Helping or Herding? Reward Model Ensembles Mitigate but do not Eliminate Reward Hacking
by: Eisenstein, Jacob, et al.
Published: (2023)
by: Eisenstein, Jacob, et al.
Published: (2023)
The Same But Different: Structural Similarities and Differences in Multilingual Language Modeling
by: Zhang, Ruochen, et al.
Published: (2024)
by: Zhang, Ruochen, et al.
Published: (2024)
Provably Bounding Neural Network Preimages
by: Kotha, Suhas, et al.
Published: (2023)
by: Kotha, Suhas, et al.
Published: (2023)
Are LLMs Models of Distributional Semantics? A Case Study on Quantifiers
by: Enyan, Zhang, et al.
Published: (2024)
by: Enyan, Zhang, et al.
Published: (2024)
How Adding Metacognitive Requirements in Support of AI Feedback in Practice Exams Transforms Student Learning Behaviors
by: Ahmad, Mak, et al.
Published: (2025)
by: Ahmad, Mak, et al.
Published: (2025)
Exploring Crowdworkers' Perceptions, Current Practices, and Desired Practices Regarding Using Non-Workstation Devices for Crowdwork
by: Dutta, Senjuti, et al.
Published: (2024)
by: Dutta, Senjuti, et al.
Published: (2024)
Unveiling the Inter-Related Preferences of Crowdworkers: Implications for Personalized and Flexible Platform Design
by: Dutta, Senjuti, et al.
Published: (2024)
by: Dutta, Senjuti, et al.
Published: (2024)
Private Gradient Descent for Linear Regression: Tighter Error Bounds and Instance-Specific Uncertainty Estimation
by: Brown, Gavin, et al.
Published: (2024)
by: Brown, Gavin, et al.
Published: (2024)
Adaptive Diffusion Denoised Smoothing : Certified Robustness via Randomized Smoothing with Differentially Private Guided Denoising Diffusion
by: Shpilevskiy, Frederick, et al.
Published: (2025)
by: Shpilevskiy, Frederick, et al.
Published: (2025)
Similar Items
-
Beyond Thumbs Up/Down: Untangling Challenges of Fine-Grained Feedback for Text-to-Image Generation
by: Collins, Katherine M., et al.
Published: (2024) -
From Prediction to Understanding: Will AI Foundation Models Transform Brain Science?
by: Serre, Thomas, et al.
Published: (2025) -
Talking Heads: Understanding Inter-layer Communication in Transformer Language Models
by: Merullo, Jack, et al.
Published: (2024) -
Does Training on Synthetic Data Make Models Less Robust?
by: Zhang, Lingze, et al.
Published: (2025) -
How Do Language Models Compose Functions?
by: Khandelwal, Apoorv, et al.
Published: (2025)