Salvato in:
| Autori principali: | Chang, Ting-Yun, Thomason, Jesse, Jia, Robin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2311.09060 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
When Parts Are Greater Than Sums: Individual LLM Components Can Outperform Full Models
di: Chang, Ting-Yun, et al.
Pubblicazione: (2024)
di: Chang, Ting-Yun, et al.
Pubblicazione: (2024)
Language Models can Infer Action Semantics for Symbolic Planners from Environment Feedback
di: Zhu, Wang, et al.
Pubblicazione: (2024)
di: Zhu, Wang, et al.
Pubblicazione: (2024)
Why Do Some Inputs Break Low-Bit LLM Quantization?
di: Chang, Ting-Yun, et al.
Pubblicazione: (2025)
di: Chang, Ting-Yun, et al.
Pubblicazione: (2025)
PDDL-Mind: Large Language Models are Capable on Belief Reasoning with Reliable State Tracking
di: Zhu, Wang Bill, et al.
Pubblicazione: (2026)
di: Zhu, Wang Bill, et al.
Pubblicazione: (2026)
PSALM-V: Automating Symbolic Planning in Interactive Visual Environments with Large Language Models
di: Zhu, Wang Bill, et al.
Pubblicazione: (2025)
di: Zhu, Wang Bill, et al.
Pubblicazione: (2025)
Guess or Recall? Training CNNs to Classify and Localize Memorization in LLMs
di: Dentan, Jérémie, et al.
Pubblicazione: (2025)
di: Dentan, Jérémie, et al.
Pubblicazione: (2025)
Two Tales of Persona in LLMs: A Survey of Role-Playing and Personalization
di: Tseng, Yu-Min, et al.
Pubblicazione: (2024)
di: Tseng, Yu-Min, et al.
Pubblicazione: (2024)
Adjust for Trust: Mitigating Trust-Induced Inappropriate Reliance on AI Assistance
di: Srinivasan, Tejas, et al.
Pubblicazione: (2025)
di: Srinivasan, Tejas, et al.
Pubblicazione: (2025)
Efficient End-to-End Visual Document Understanding with Rationale Distillation
di: Zhu, Wang, et al.
Pubblicazione: (2023)
di: Zhu, Wang, et al.
Pubblicazione: (2023)
A Tale of Two Structures: Do LLMs Capture the Fractal Complexity of Language?
di: Alabdulmohsin, Ibrahim, et al.
Pubblicazione: (2025)
di: Alabdulmohsin, Ibrahim, et al.
Pubblicazione: (2025)
Large Language Models Do Multi-Label Classification Differently
di: Ma, Marcus, et al.
Pubblicazione: (2025)
di: Ma, Marcus, et al.
Pubblicazione: (2025)
Safetywashing: Do AI Safety Benchmarks Actually Measure Safety Progress?
di: Ren, Richard, et al.
Pubblicazione: (2024)
di: Ren, Richard, et al.
Pubblicazione: (2024)
TwoStep: Multi-agent Task Planning using Classical Planners and Large Language Models
di: Bai, David, et al.
Pubblicazione: (2024)
di: Bai, David, et al.
Pubblicazione: (2024)
Phonological Representation Learning for Isolated Signs Improves Out-of-Vocabulary Generalization
di: Kezar, Lee, et al.
Pubblicazione: (2025)
di: Kezar, Lee, et al.
Pubblicazione: (2025)
When Do LLMs Admit Their Mistakes? Understanding The Role Of Model Belief In Retraction
di: Yang, Yuqing, et al.
Pubblicazione: (2025)
di: Yang, Yuqing, et al.
Pubblicazione: (2025)
Localizing Paragraph Memorization in Language Models
di: Stoehr, Niklas, et al.
Pubblicazione: (2024)
di: Stoehr, Niklas, et al.
Pubblicazione: (2024)
WinoViz: Probing Visual Properties of Objects Under Different States
di: Jin, Woojeong, et al.
Pubblicazione: (2024)
di: Jin, Woojeong, et al.
Pubblicazione: (2024)
Words that make SENSE: Sensorimotor Norms in Learned Lexical Token Representations
di: Gupta, Abhinav, et al.
Pubblicazione: (2026)
di: Gupta, Abhinav, et al.
Pubblicazione: (2026)
Unveiling Over-Memorization in Finetuning LLMs for Reasoning Tasks
di: Ruan, Zhiwen, et al.
Pubblicazione: (2025)
di: Ruan, Zhiwen, et al.
Pubblicazione: (2025)
Short-Context Dominance: How Much Local Context Natural Language Actually Needs?
di: Vakilian, Vala, et al.
Pubblicazione: (2025)
di: Vakilian, Vala, et al.
Pubblicazione: (2025)
LocalBench: Benchmarking LLMs on County-Level Local Knowledge and Reasoning
di: Gao, Zihan, et al.
Pubblicazione: (2025)
di: Gao, Zihan, et al.
Pubblicazione: (2025)
Few-Shot VQA with Frozen LLMs: A Tale of Two Approaches
di: Sterner, Igor, et al.
Pubblicazione: (2024)
di: Sterner, Igor, et al.
Pubblicazione: (2024)
Do LLMs Really Memorize Personally Identifiable Information? Revisiting PII Leakage with a Cue-Controlled Memorization Framework
di: Luo, Xiaoyu, et al.
Pubblicazione: (2026)
di: Luo, Xiaoyu, et al.
Pubblicazione: (2026)
Iterative Formalization and Planning in Partially Observable Environments
di: Gong, Liancheng, et al.
Pubblicazione: (2025)
di: Gong, Liancheng, et al.
Pubblicazione: (2025)
What Do Claim Verification Datasets Actually Test? A Reasoning Trace Analysis
di: Rao, Delip, et al.
Pubblicazione: (2026)
di: Rao, Delip, et al.
Pubblicazione: (2026)
Instructional Goal-Aligned Question Generation for Student Evaluation in Virtual Lab Settings: How Closely Do LLMs Actually Align?
di: Knipper, R. Alexander, et al.
Pubblicazione: (2025)
di: Knipper, R. Alexander, et al.
Pubblicazione: (2025)
Be like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs
di: Hans, Abhimanyu, et al.
Pubblicazione: (2024)
di: Hans, Abhimanyu, et al.
Pubblicazione: (2024)
When Can LLMs Actually Correct Their Own Mistakes? A Critical Survey of Self-Correction of LLMs
di: Kamoi, Ryo, et al.
Pubblicazione: (2024)
di: Kamoi, Ryo, et al.
Pubblicazione: (2024)
From Calibration to Collaboration: LLM Uncertainty Quantification Should Be More Human-Centered
di: Devic, Siddartha, et al.
Pubblicazione: (2025)
di: Devic, Siddartha, et al.
Pubblicazione: (2025)
Benchmarking Chinese Commonsense Reasoning of LLMs: From Chinese-Specifics to Reasoning-Memorization Correlations
di: Sun, Jiaxing, et al.
Pubblicazione: (2024)
di: Sun, Jiaxing, et al.
Pubblicazione: (2024)
Rote Learning Considered Useful: Generalizing over Memorized Data in LLMs
di: Wu, Qinyuan, et al.
Pubblicazione: (2025)
di: Wu, Qinyuan, et al.
Pubblicazione: (2025)
Generating Contextually-Relevant Navigation Instructions for Blind and Low Vision People
di: Merchant, Zain, et al.
Pubblicazione: (2024)
di: Merchant, Zain, et al.
Pubblicazione: (2024)
Decomposing the Delta: What Do Models Actually Learn from Preference Pairs?
di: Lee, Chia-Hsuan, et al.
Pubblicazione: (2026)
di: Lee, Chia-Hsuan, et al.
Pubblicazione: (2026)
Unique Hard Attention: A Tale of Two Sides
di: Jerad, Selim, et al.
Pubblicazione: (2025)
di: Jerad, Selim, et al.
Pubblicazione: (2025)
Alpaca against Vicuna: Using LLMs to Uncover Memorization of LLMs
di: Kassem, Aly M., et al.
Pubblicazione: (2024)
di: Kassem, Aly M., et al.
Pubblicazione: (2024)
Memorization or Reasoning? Exploring the Idiom Understanding of LLMs
di: Kim, Jisu, et al.
Pubblicazione: (2025)
di: Kim, Jisu, et al.
Pubblicazione: (2025)
Mitigating Memorization in LLMs using Activation Steering
di: Suri, Manan, et al.
Pubblicazione: (2025)
di: Suri, Manan, et al.
Pubblicazione: (2025)
Memorization and Knowledge Injection in Gated LLMs
di: Pan, Xu, et al.
Pubblicazione: (2025)
di: Pan, Xu, et al.
Pubblicazione: (2025)
Self Knowledge Re-expression: A Fully Local Method for Adapting LLMs to Tasks Using Intrinsic Knowledge
di: Wang, Mengyu, et al.
Pubblicazione: (2026)
di: Wang, Mengyu, et al.
Pubblicazione: (2026)
When Names Disappear: Revealing What LLMs Actually Understand About Code
di: Le, Cuong Chi, et al.
Pubblicazione: (2025)
di: Le, Cuong Chi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
When Parts Are Greater Than Sums: Individual LLM Components Can Outperform Full Models
di: Chang, Ting-Yun, et al.
Pubblicazione: (2024) -
Language Models can Infer Action Semantics for Symbolic Planners from Environment Feedback
di: Zhu, Wang, et al.
Pubblicazione: (2024) -
Why Do Some Inputs Break Low-Bit LLM Quantization?
di: Chang, Ting-Yun, et al.
Pubblicazione: (2025) -
PDDL-Mind: Large Language Models are Capable on Belief Reasoning with Reliable State Tracking
di: Zhu, Wang Bill, et al.
Pubblicazione: (2026) -
PSALM-V: Automating Symbolic Planning in Interactive Visual Environments with Large Language Models
di: Zhu, Wang Bill, et al.
Pubblicazione: (2025)