We Can't Understand AI Using our Existing Vocabulary
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hewitt, John, Geirhos, Robert, Kim, Been |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Neologism Learning for Controllability and Self-Verbalization
von: Hewitt, John, et al.
Veröffentlicht: (2025)
von: Hewitt, John, et al.
Veröffentlicht: (2025)
LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench
von: Valmeekam, Karthik, et al.
Veröffentlicht: (2024)
von: Valmeekam, Karthik, et al.
Veröffentlicht: (2024)
QuestBench: Can LLMs ask the right question to acquire information in reasoning tasks?
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
How Can We Effectively Expand the Vocabulary of LLMs with 0.01GB of Target Language Text?
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024)
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024)
Can't say cant? Measuring and Reasoning of Dark Jargons in Large Language Models
von: Ji, Xu, et al.
Veröffentlicht: (2024)
von: Ji, Xu, et al.
Veröffentlicht: (2024)
Beyond I'm Sorry, I Can't: Dissecting Large Language Model Refusal
von: Prakash, Nirmalendu, et al.
Veröffentlicht: (2025)
von: Prakash, Nirmalendu, et al.
Veröffentlicht: (2025)
LLMs Can't Handle Peer Pressure: Crumbling under Multi-Agent Social Interactions
von: Song, Maojia, et al.
Veröffentlicht: (2025)
von: Song, Maojia, et al.
Veröffentlicht: (2025)
If You Can't Use Them, Recycle Them: Optimizing Merging at Scale Mitigates Performance Tradeoffs
von: Khalifa, Muhammad, et al.
Veröffentlicht: (2024)
von: Khalifa, Muhammad, et al.
Veröffentlicht: (2024)
Your Teacher Can't Help You Here: Combating Supervision Fidelity Decay in On-Policy Distillation
von: Liu, Yanjiang, et al.
Veröffentlicht: (2026)
von: Liu, Yanjiang, et al.
Veröffentlicht: (2026)
Climbing the Ladder of Reasoning: What LLMs Can-and Still Can't-Solve after SFT?
von: Sun, Yiyou, et al.
Veröffentlicht: (2025)
von: Sun, Yiyou, et al.
Veröffentlicht: (2025)
Don't trust your eyes: on the (un)reliability of feature visualizations
von: Geirhos, Robert, et al.
Veröffentlicht: (2023)
von: Geirhos, Robert, et al.
Veröffentlicht: (2023)
Bench-2-CoP: Can We Trust Benchmarking for EU AI Compliance?
von: Prandi, Matteo, et al.
Veröffentlicht: (2025)
von: Prandi, Matteo, et al.
Veröffentlicht: (2025)
Can We Trust LLM Detectors?
von: Sandhan, Jivnesh, et al.
Veröffentlicht: (2026)
von: Sandhan, Jivnesh, et al.
Veröffentlicht: (2026)
Can We Locate and Prevent Stereotypes in LLMs?
von: D'Souza, Alex
Veröffentlicht: (2026)
von: D'Souza, Alex
Veröffentlicht: (2026)
A Zero-Shot Open-Vocabulary Pipeline for Dialogue Understanding
von: Safa, Abdulfattah, et al.
Veröffentlicht: (2024)
von: Safa, Abdulfattah, et al.
Veröffentlicht: (2024)
Because we have LLMs, we Can and Should Pursue Agentic Interpretability
von: Kim, Been, et al.
Veröffentlicht: (2025)
von: Kim, Been, et al.
Veröffentlicht: (2025)
You Can't Steal Nothing: Mitigating Prompt Leakages in LLMs via System Vectors
von: Cao, Bochuan, et al.
Veröffentlicht: (2025)
von: Cao, Bochuan, et al.
Veröffentlicht: (2025)
LLMs Can Plan Only If We Tell Them
von: Sel, Bilgehan, et al.
Veröffentlicht: (2025)
von: Sel, Bilgehan, et al.
Veröffentlicht: (2025)
Random Initialization Can't Catch Up: The Advantage of Language Model Transfer for Time Series Forecasting
von: Riachi, Roland, et al.
Veröffentlicht: (2025)
von: Riachi, Roland, et al.
Veröffentlicht: (2025)
LLM-REVal: Can We Trust LLM Reviewers Yet?
von: Li, Rui, et al.
Veröffentlicht: (2025)
von: Li, Rui, et al.
Veröffentlicht: (2025)
Can We Edit LLMs for Long-Tail Biomedical Knowledge?
von: Yi, Xinhao, et al.
Veröffentlicht: (2025)
von: Yi, Xinhao, et al.
Veröffentlicht: (2025)
Can We Verify Step by Step for Incorrect Answer Detection?
von: Xu, Xin, et al.
Veröffentlicht: (2024)
von: Xu, Xin, et al.
Veröffentlicht: (2024)
No Reader Left Behind: Multi-Agent Summaries Everyone Can Understand
von: Jung, Jimin, et al.
Veröffentlicht: (2026)
von: Jung, Jimin, et al.
Veröffentlicht: (2026)
Can't See the Forest for the Trees: Benchmarking Multimodal Safety Awareness for Multimodal LLMs
von: Wang, Wenxuan, et al.
Veröffentlicht: (2025)
von: Wang, Wenxuan, et al.
Veröffentlicht: (2025)
Happiness is Sharing a Vocabulary: A Study of Transliteration Methods
von: Jung, Haeji, et al.
Veröffentlicht: (2025)
von: Jung, Haeji, et al.
Veröffentlicht: (2025)
Even GPT-5.2 Can't Count to Five: The Case for Zero-Error Horizons in Trustworthy LLMs
von: Sato, Ryoma
Veröffentlicht: (2026)
von: Sato, Ryoma
Veröffentlicht: (2026)
Can't Remember Details in Long Documents? You Need Some R&R
von: Agrawal, Devanshu, et al.
Veröffentlicht: (2024)
von: Agrawal, Devanshu, et al.
Veröffentlicht: (2024)
Scaling Laws with Vocabulary: Larger Models Deserve Larger Vocabularies
von: Tao, Chaofan, et al.
Veröffentlicht: (2024)
von: Tao, Chaofan, et al.
Veröffentlicht: (2024)
Efficient and Effective Vocabulary Expansion Towards Multilingual Large Language Models
von: Kim, Seungduk, et al.
Veröffentlicht: (2024)
von: Kim, Seungduk, et al.
Veröffentlicht: (2024)
Puzzled by Puzzles: When Vision-Language Models Can't Take a Hint
von: Lee, Heekyung, et al.
Veröffentlicht: (2025)
von: Lee, Heekyung, et al.
Veröffentlicht: (2025)
SciMaster: Towards General-Purpose Scientific AI Agents, Part I. X-Master as Foundation: Can We Lead on Humanity's Last Exam?
von: Chai, Jingyi, et al.
Veröffentlicht: (2025)
von: Chai, Jingyi, et al.
Veröffentlicht: (2025)
Can We Talk Models Into Seeing the World Differently?
von: Gavrikov, Paul, et al.
Veröffentlicht: (2024)
von: Gavrikov, Paul, et al.
Veröffentlicht: (2024)
Overcoming Vocabulary Mismatch: Vocabulary-agnostic Teacher Guided Language Modeling
von: Shin, Haebin, et al.
Veröffentlicht: (2025)
von: Shin, Haebin, et al.
Veröffentlicht: (2025)
Now It Sounds Like You: Learning Personalized Vocabulary On Device
von: Wang, Sid, et al.
Veröffentlicht: (2023)
von: Wang, Sid, et al.
Veröffentlicht: (2023)
You Can't Get There From Here: Redefining Information Science to address our sociotechnical futures
von: Humr, Scott, et al.
Veröffentlicht: (2025)
von: Humr, Scott, et al.
Veröffentlicht: (2025)
Can AI Be as Creative as Humans?
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
von: Wang, Haonan, et al.
Veröffentlicht: (2024)
Can We Still Hear the Accent? Investigating the Resilience of Native Language Signals in the LLM Era
von: Utami, Nabelanita, et al.
Veröffentlicht: (2026)
von: Utami, Nabelanita, et al.
Veröffentlicht: (2026)
ECCO: Can We Improve Model-Generated Code Efficiency Without Sacrificing Functional Correctness?
von: Waghjale, Siddhant, et al.
Veröffentlicht: (2024)
von: Waghjale, Siddhant, et al.
Veröffentlicht: (2024)
RedacBench: Can AI Erase Your Secrets?
von: Jeon, Hyunjun, et al.
Veröffentlicht: (2026)
von: Jeon, Hyunjun, et al.
Veröffentlicht: (2026)
Dictionaries to the Rescue: Cross-Lingual Vocabulary Transfer for Low-Resource Languages Using Bilingual Dictionaries
von: Sakajo, Haruki, et al.
Veröffentlicht: (2025)
von: Sakajo, Haruki, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Neologism Learning for Controllability and Self-Verbalization
von: Hewitt, John, et al.
Veröffentlicht: (2025) -
LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench
von: Valmeekam, Karthik, et al.
Veröffentlicht: (2024) -
QuestBench: Can LLMs ask the right question to acquire information in reasoning tasks?
von: Li, Belinda Z., et al.
Veröffentlicht: (2025) -
How Can We Effectively Expand the Vocabulary of LLMs with 0.01GB of Target Language Text?
von: Yamaguchi, Atsuki, et al.
Veröffentlicht: (2024) -
Can't say cant? Measuring and Reasoning of Dark Jargons in Large Language Models
von: Ji, Xu, et al.
Veröffentlicht: (2024)