Wait, but Tylenol is Acetaminophen... Investigating and Improving Language Models' Ability to Resist Requests for Misinformation
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Shan, Gao, Mingye, Sasse, Kuleen, Hartvigsen, Thomas, Anthony, Brian, Fan, Lizhou, Aerts, Hugo, Gallifant, Jack, Bitterman, Danielle |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sparse Autoencoder Features for Classifications and Transferability
by: Gallifant, Jack, et al.
Published: (2025)
by: Gallifant, Jack, et al.
Published: (2025)
Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks
by: Gallifant, Jack, et al.
Published: (2024)
by: Gallifant, Jack, et al.
Published: (2024)
debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias
by: Sasse, Kuleen, et al.
Published: (2024)
by: Sasse, Kuleen, et al.
Published: (2024)
KScope: A Framework for Characterizing the Knowledge Status of Language Models
by: Xiao, Yuxin, et al.
Published: (2025)
by: Xiao, Yuxin, et al.
Published: (2025)
MedBrowseComp: Benchmarking Medical Deep Research and Computer Use
by: Chen, Shan, et al.
Published: (2025)
by: Chen, Shan, et al.
Published: (2025)
Controllable Hybrid Captioner for Improved Long-form Video Understanding
by: Sasse, Kuleen, et al.
Published: (2025)
by: Sasse, Kuleen, et al.
Published: (2025)
Cross-Care: Assessing the Healthcare Implications of Pre-training Data on Language Model Bias
by: Chen, Shan, et al.
Published: (2024)
by: Chen, Shan, et al.
Published: (2024)
Improving Clinical NLP Performance through Language Model-Generated Synthetic Clinical Data
by: Chen, Shan, et al.
Published: (2024)
by: Chen, Shan, et al.
Published: (2024)
Can Language Models Identify Side Effects of Breast Cancer Radiation Treatments?
by: Seah, Natalie, et al.
Published: (2026)
by: Seah, Natalie, et al.
Published: (2026)
Disease Entity Recognition and Normalization is Improved with Large Language Model Derived Synthetic Normalized Mentions
by: Sasse, Kuleen, et al.
Published: (2024)
by: Sasse, Kuleen, et al.
Published: (2024)
Autism, leucovorin, Tylenol, and pregnancy
Published: (2025)
Published: (2025)
To Burst or Not to Burst: Generating and Quantifying Improbable Text
by: Sasse, Kuleen, et al.
Published: (2024)
by: Sasse, Kuleen, et al.
Published: (2024)
Making FETCH! Happen: Finding Emergent Dog Whistles Through Common Habitats
by: Sasse, Kuleen, et al.
Published: (2024)
by: Sasse, Kuleen, et al.
Published: (2024)
Analyzing Diversity in Healthcare LLM Research: A Scientometric Perspective
by: Restrepo, David, et al.
Published: (2024)
by: Restrepo, David, et al.
Published: (2024)
Evaluation of ChatGPT Family of Models for Biomedical Reasoning and Classification
by: Chen, Shan, et al.
Published: (2023)
by: Chen, Shan, et al.
Published: (2023)
The use of large language models to enhance cancer clinical trial educational materials
by: Gao, Mingye, et al.
Published: (2024)
by: Gao, Mingye, et al.
Published: (2024)
Seeds of Stereotypes: A Large-Scale Textual Analysis of Race and Gender Associations with Diseases in Online Sources
by: Hansen, Lasse Hyldig, et al.
Published: (2024)
by: Hansen, Lasse Hyldig, et al.
Published: (2024)
WorldMedQA-V: a multilingual, multimodal medical examination dataset for multimodal language models evaluation
by: Matos, João, et al.
Published: (2024)
by: Matos, João, et al.
Published: (2024)
NPHardEval: Dynamic Benchmark on Reasoning Ability of Large Language Models via Complexity Classes
by: Fan, Lizhou, et al.
Published: (2023)
by: Fan, Lizhou, et al.
Published: (2023)
When Models Reason in Your Language: Controlling Thinking Language Comes at the Cost of Accuracy
by: Qi, Jirui, et al.
Published: (2025)
by: Qi, Jirui, et al.
Published: (2025)
EHRmonize: A Framework for Medical Concept Abstraction from Electronic Health Records using Large Language Models
by: Matos, João, et al.
Published: (2024)
by: Matos, João, et al.
Published: (2024)
The German Chambers of Commerce and Industry Self-governance, Service, the General Representation of Interests and the Dual System of Professional Education
by: Eberhard Sasse
by: Eberhard Sasse
Invisible Women: The Children's Librarian in America
by: Sasse, Margo
Published: (1973)
by: Sasse, Margo
Published: (1973)
Math Neurosurgery: Isolating Language Models' Math Reasoning Abilities Using Only Forward Passes
by: Christ, Bryan R., et al.
Published: (2024)
by: Christ, Bryan R., et al.
Published: (2024)
Cued to Queue: Information in Waiting-Line Auctions
by: Hirsch, Jack, et al.
Published: (2025)
by: Hirsch, Jack, et al.
Published: (2025)
Large Language Models to Identify Social Determinants of Health in Electronic Health Records
by: Guevara, Marco, et al.
Published: (2023)
by: Guevara, Marco, et al.
Published: (2023)
The Evolution of Lying in a Spatially-Explicit Prisoner's Dilemma Model
by: Hartvigsen, Gregg
Published: (2026)
by: Hartvigsen, Gregg
Published: (2026)
Finding triangle‐free 2‐factors in general graphs
by: David Hartvigsen
Published: (2024)
by: David Hartvigsen
Published: (2024)
Conversational Inoculation to Enhance Resistance to Misinformation
by: Szabó, Dániel, et al.
Published: (2026)
by: Szabó, Dániel, et al.
Published: (2026)
Triangulating Evidence on Prenatal Acetaminophen Use: Insights From a Large Japanese Cohort
by: Renee M. Gardner, et al.
Published: (2025)
by: Renee M. Gardner, et al.
Published: (2025)
Cooperative Information Network Interlibrary Loan Non-Filled Request Study.
by: Plotkin, Jack
Published: (1975)
by: Plotkin, Jack
Published: (1975)
Proof of Time: A Benchmark for Evaluating Scientific Idea Judgments
by: Ye, Bingyang, et al.
Published: (2026)
by: Ye, Bingyang, et al.
Published: (2026)
When Raw Data Prevails: Are Large Language Model Embeddings Effective in Numerical Data Representation for Medical Machine Learning Applications?
by: Gao, Yanjun, et al.
Published: (2024)
by: Gao, Yanjun, et al.
Published: (2024)
A ELETROBRAS E AS EMPRESAS FORNECEDORAS DE EQUIPAMENTOS PARA O SETOR ELÉTRICO BRASILEIRO (1960-1980)
by: Carla Muller Sasse
Published: (2016)
by: Carla Muller Sasse
Published: (2016)
La evaluación literaria: una retrospectiva
by: Jochen Schulte-Sasse
Published: (2010)
by: Jochen Schulte-Sasse
Published: (2010)
Modeling the Impact of Misinformation Dynamics on Antimicrobial Resistance
by: Fakih, Laurance, et al.
Published: (2025)
by: Fakih, Laurance, et al.
Published: (2025)
Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency
by: Wang, Chenlong, et al.
Published: (2025)
by: Wang, Chenlong, et al.
Published: (2025)
Solar‐Activated Persulphate Oxidation of Acetaminophen in a Compound Parabolic Reactor: A Seasonal Investigation
by: Sylvia B. Kpange, et al.
Published: (2025)
by: Sylvia B. Kpange, et al.
Published: (2025)
MPCG: Multi-Round Persona-Conditioned Generation for Modeling the Evolution of Misinformation with LLMs
by: Chong, Jun Rong Brian, et al.
Published: (2025)
by: Chong, Jun Rong Brian, et al.
Published: (2025)
Stack Attention: Improving the Ability of Transformers to Model Hierarchical Patterns
by: DuSell, Brian, et al.
Published: (2023)
by: DuSell, Brian, et al.
Published: (2023)
Similar Items
-
Sparse Autoencoder Features for Classifications and Transferability
by: Gallifant, Jack, et al.
Published: (2025) -
Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks
by: Gallifant, Jack, et al.
Published: (2024) -
debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias
by: Sasse, Kuleen, et al.
Published: (2024) -
KScope: A Framework for Characterizing the Knowledge Status of Language Models
by: Xiao, Yuxin, et al.
Published: (2025) -
MedBrowseComp: Benchmarking Medical Deep Research and Computer Use
by: Chen, Shan, et al.
Published: (2025)