Neologism Learning for Controllability and Self-Verbalization
Fuente:
arXiv
Saved in:
| Main Authors: | Hewitt, John, Tafjord, Oyvind, Geirhos, Robert, Kim, Been |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
We Can't Understand AI Using our Existing Vocabulary
by: Hewitt, John, et al.
Published: (2025)
by: Hewitt, John, et al.
Published: (2025)
Because we have LLMs, we Can and Should Pursue Agentic Interpretability
by: Kim, Been, et al.
Published: (2025)
by: Kim, Been, et al.
Published: (2025)
Digital Socrates: Evaluating LLMs through Explanation Critiques
by: Gu, Yuling, et al.
Published: (2023)
by: Gu, Yuling, et al.
Published: (2023)
BaRDa: A Belief and Reasoning Dataset that Separates Factual Accuracy and Reasoning Ability
by: Clark, Peter, et al.
Published: (2023)
by: Clark, Peter, et al.
Published: (2023)
Answer, Assemble, Ace: Understanding How LMs Answer Multiple Choice Questions
by: Wiegreffe, Sarah, et al.
Published: (2024)
by: Wiegreffe, Sarah, et al.
Published: (2024)
From 124 Million Tokens to 1,021 Neologisms: A Large-Scale Pipeline for Automatic Neologism Detection
by: Rossini, Diego, et al.
Published: (2026)
by: Rossini, Diego, et al.
Published: (2026)
OLMES: A Standard for Language Model Evaluations
by: Gu, Yuling, et al.
Published: (2024)
by: Gu, Yuling, et al.
Published: (2024)
SimpleToM: Exposing the Gap between Explicit ToM Inference and Implicit ToM Application in LLMs
by: Gu, Yuling, et al.
Published: (2024)
by: Gu, Yuling, et al.
Published: (2024)
Neologism Learning as a Parameter-Efficient Alternative to Fine-Tuning for Model Steering
by: Park, Sungjoon, et al.
Published: (2025)
by: Park, Sungjoon, et al.
Published: (2025)
NeoAMT: Neologism-Aware Agentic Machine Translation with Reinforcement Learning
by: Miao, Zhongtao, et al.
Published: (2026)
by: Miao, Zhongtao, et al.
Published: (2026)
NEO-BENCH: Evaluating Robustness of Large Language Models with Neologisms
by: Zheng, Jonathan, et al.
Published: (2024)
by: Zheng, Jonathan, et al.
Published: (2024)
Don't trust your eyes: on the (un)reliability of feature visualizations
by: Geirhos, Robert, et al.
Published: (2023)
by: Geirhos, Robert, et al.
Published: (2023)
NeoN: A Tool for Automated Detection, Linguistic and LLM-Driven Analysis of Neologisms in Polish
by: Tomaszewska, Aleksandra, et al.
Published: (2025)
by: Tomaszewska, Aleksandra, et al.
Published: (2025)
Reheat Nachos for Dinner? Evaluating AI Support for Cross-Cultural Communication of Neologisms
by: Ki, Dayeon, et al.
Published: (2026)
by: Ki, Dayeon, et al.
Published: (2026)
From Models to Microtheories: Distilling a Model's Topical Knowledge for Grounded Question Answering
by: Weir, Nathaniel, et al.
Published: (2024)
by: Weir, Nathaniel, et al.
Published: (2024)
Do LLMs Know What Luxembourgish Borrows? Probing Lexical Neology in Low-Resource Multilingual Models
by: Hosseini-Kivanani, Nina
Published: (2026)
by: Hosseini-Kivanani, Nina
Published: (2026)
Improving Parametric Knowledge Access in Reasoning Language Models
by: Ma, Melody, et al.
Published: (2026)
by: Ma, Melody, et al.
Published: (2026)
Subliminal Steering: Stronger Encoding of Hidden Signals
by: Morgulis, George, et al.
Published: (2026)
by: Morgulis, George, et al.
Published: (2026)
DISCOVERYWORLD: A Virtual Environment for Developing and Evaluating Automated Scientific Discovery Agents
by: Jansen, Peter, et al.
Published: (2024)
by: Jansen, Peter, et al.
Published: (2024)
Self-Routing RAG: Binding Selective Retrieval with Knowledge Verbalization
by: Wu, Di, et al.
Published: (2025)
by: Wu, Di, et al.
Published: (2025)
Calibrating Verbalized Confidence with Self-Generated Distractors
by: Wang, Victor, et al.
Published: (2025)
by: Wang, Victor, et al.
Published: (2025)
ADVICE: Answer-Dependent Verbalized Confidence Estimation
by: Seo, Ki Jung, et al.
Published: (2025)
by: Seo, Ki Jung, et al.
Published: (2025)
CodeScientist: End-to-End Semi-Automated Scientific Discovery with Code-based Experimentation
by: Jansen, Peter, et al.
Published: (2025)
by: Jansen, Peter, et al.
Published: (2025)
Verbalized Confidence Triggers Self-Verification: Emergent Behavior Without Explicit Reasoning Supervision
by: Jang, Chaeyun, et al.
Published: (2025)
by: Jang, Chaeyun, et al.
Published: (2025)
Verbalized Machine Learning: Revisiting Machine Learning with Language Models
by: Xiao, Tim Z., et al.
Published: (2024)
by: Xiao, Tim Z., et al.
Published: (2024)
PreScience: A Benchmark for Forecasting Scientific Contributions
by: Ajith, Anirudh, et al.
Published: (2026)
by: Ajith, Anirudh, et al.
Published: (2026)
Lexicography of Coronavirus-related Neologisms
Published: (2023)
Published: (2023)
QuestBench: Can LLMs ask the right question to acquire information in reasoning tasks?
by: Li, Belinda Z., et al.
Published: (2025)
by: Li, Belinda Z., et al.
Published: (2025)
Not All Explanations Simulate Equally: Comparing Verbalized Feature Attributions and Self-Generated Rationales
by: Hong, Pingjun, et al.
Published: (2026)
by: Hong, Pingjun, et al.
Published: (2026)
Leveraging Language Models and Machine Learning in Verbal Autopsy Analysis
by: Chu, Yue
Published: (2025)
by: Chu, Yue
Published: (2025)
Think, Verbalize, then Speak: Bridging Complex Thoughts and Comprehensible Speech
by: Woo, Sang Hoon, et al.
Published: (2025)
by: Woo, Sang Hoon, et al.
Published: (2025)
On Verbalized Confidence Scores for LLMs
by: Yang, Daniel, et al.
Published: (2024)
by: Yang, Daniel, et al.
Published: (2024)
OVD: On-policy Verbal Distillation
by: Xiong, Jing, et al.
Published: (2026)
by: Xiong, Jing, et al.
Published: (2026)
Enhancing Systematic Decompositional Natural Language Inference Using Informal Logic
by: Weir, Nathaniel, et al.
Published: (2024)
by: Weir, Nathaniel, et al.
Published: (2024)
Boosting Self-Efficacy and Performance of Large Language Models via Verbal Efficacy Stimulations
by: Chen, Rui, et al.
Published: (2025)
by: Chen, Rui, et al.
Published: (2025)
Learning Visual Composition through Improved Semantic Guidance
by: Stone, Austin, et al.
Published: (2024)
by: Stone, Austin, et al.
Published: (2024)
Inference and Verbalization Functions During In-Context Learning
by: Tao, Junyi, et al.
Published: (2024)
by: Tao, Junyi, et al.
Published: (2024)
On the Robustness of Verbal Confidence of LLMs in Adversarial Attacks
by: Obadinma, Stephen, et al.
Published: (2025)
by: Obadinma, Stephen, et al.
Published: (2025)
Instruction-following Evaluation through Verbalizer Manipulation
by: Li, Shiyang, et al.
Published: (2023)
by: Li, Shiyang, et al.
Published: (2023)
Interpreting and Controlling Model Behavior via Constitutions for Atomic Concept Edits
by: Kalibhat, Neha, et al.
Published: (2026)
by: Kalibhat, Neha, et al.
Published: (2026)
Similar Items
-
We Can't Understand AI Using our Existing Vocabulary
by: Hewitt, John, et al.
Published: (2025) -
Because we have LLMs, we Can and Should Pursue Agentic Interpretability
by: Kim, Been, et al.
Published: (2025) -
Digital Socrates: Evaluating LLMs through Explanation Critiques
by: Gu, Yuling, et al.
Published: (2023) -
BaRDa: A Belief and Reasoning Dataset that Separates Factual Accuracy and Reasoning Ability
by: Clark, Peter, et al.
Published: (2023) -
Answer, Assemble, Ace: Understanding How LMs Answer Multiple Choice Questions
by: Wiegreffe, Sarah, et al.
Published: (2024)