Salvato in:
| Autori principali: | Liu, Wenhao, An, Siyu, Lu, Junru, Wu, Muling, Li, Tianlong, Wang, Xiaohua, lv, Changze, Zheng, Xiaoqing, Yin, Di, Sun, Xing, Huang, Xuanjing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2409.16913 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Advancing Parameter Efficiency in Fine-tuning via Representation Editing
di: Wu, Muling, et al.
Pubblicazione: (2024)
di: Wu, Muling, et al.
Pubblicazione: (2024)
Revisiting Jailbreaking for Large Language Models: A Representation Engineering Perspective
di: Li, Tianlong, et al.
Pubblicazione: (2024)
di: Li, Tianlong, et al.
Pubblicazione: (2024)
Aligning Large Language Models with Human Preferences through Representation Engineering
di: Liu, Wenhao, et al.
Pubblicazione: (2023)
di: Liu, Wenhao, et al.
Pubblicazione: (2023)
UPLex: Fine-Grained Personality Control in Large Language Models via Unsupervised Lexical Modulation
di: Li, Tianlong, et al.
Pubblicazione: (2023)
di: Li, Tianlong, et al.
Pubblicazione: (2023)
Promoting Data and Model Privacy in Federated Learning through Quantized LoRA
di: Zhu, JianHao, et al.
Pubblicazione: (2024)
di: Zhu, JianHao, et al.
Pubblicazione: (2024)
NeIn: Telling What You Don't Want
di: Bui, Nhat-Tan, et al.
Pubblicazione: (2024)
di: Bui, Nhat-Tan, et al.
Pubblicazione: (2024)
SpikeCLIP: A Contrastive Language-Image Pretrained Spiking Neural Network
di: Lv, Changze, et al.
Pubblicazione: (2023)
di: Lv, Changze, et al.
Pubblicazione: (2023)
Experts Don't Cheat: Learning What You Don't Know By Predicting Pairs
di: Johnson, Daniel D., et al.
Pubblicazione: (2024)
di: Johnson, Daniel D., et al.
Pubblicazione: (2024)
Visually Dehallucinative Instruction Generation: Know What You Don't Know
di: Cha, Sungguk, et al.
Pubblicazione: (2024)
di: Cha, Sungguk, et al.
Pubblicazione: (2024)
Show Me What You Don't Know: Efficient Sampling from Invariant Sets for Model Validation
di: Rousselot, Armand, et al.
Pubblicazione: (2026)
di: Rousselot, Armand, et al.
Pubblicazione: (2026)
Know What You Don't Know: Uncertainty Calibration of Process Reward Models
di: Park, Young-Jin, et al.
Pubblicazione: (2025)
di: Park, Young-Jin, et al.
Pubblicazione: (2025)
Know What You Don't Know: Selective Prediction for Early Exit DNNs
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2025)
di: Bajpai, Divya Jyoti, et al.
Pubblicazione: (2025)
What LLMs Think When You Don't Tell Them What to Think About?
di: Kwon, Yongchan, et al.
Pubblicazione: (2026)
di: Kwon, Yongchan, et al.
Pubblicazione: (2026)
Imagining What We Don't Know
di: Samuels, Lisa
Pubblicazione: (2026)
di: Samuels, Lisa
Pubblicazione: (2026)
Imagining What We Don't Know
di: Samuels, Lisa
Pubblicazione: (2026)
di: Samuels, Lisa
Pubblicazione: (2026)
Can AI Assistants Know What They Don't Know?
di: Cheng, Qinyuan, et al.
Pubblicazione: (2024)
di: Cheng, Qinyuan, et al.
Pubblicazione: (2024)
Knowing What You Know Is Not Enough: Large Language Model Confidences Don't Align With Their Actions
di: Pal, Arka, et al.
Pubblicazione: (2025)
di: Pal, Arka, et al.
Pubblicazione: (2025)
EpiCaR: Knowing What You Don't Know Matters for Better Reasoning in LLMs
di: Yeom, Jewon, et al.
Pubblicazione: (2026)
di: Yeom, Jewon, et al.
Pubblicazione: (2026)
Enhancing Model Privacy in Federated Learning with Random Masking and Quantization
di: Xu, Zhibo, et al.
Pubblicazione: (2025)
di: Xu, Zhibo, et al.
Pubblicazione: (2025)
What Papers Don't Tell You: Recovering Tacit Knowledge for Automated Paper Reproduction
di: Li, Lehui, et al.
Pubblicazione: (2026)
di: Li, Lehui, et al.
Pubblicazione: (2026)
Progressive Mastery: Customized Curriculum Learning with Guided Prompting for Mathematical Reasoning
di: Wu, Muling, et al.
Pubblicazione: (2025)
di: Wu, Muling, et al.
Pubblicazione: (2025)
Hatevolution: What Static Benchmarks Don't Tell Us
di: Di Bonaventura, Chiara, et al.
Pubblicazione: (2025)
di: Di Bonaventura, Chiara, et al.
Pubblicazione: (2025)
Euphemisms: Tell Me What You Do and I'll Tell You What You Are!
di: Perez, Cenel Augusto
Pubblicazione: (2025)
di: Perez, Cenel Augusto
Pubblicazione: (2025)
When You Don't Know the Answer, Say So
Pubblicazione: (2024)
Pubblicazione: (2024)
Improving Continual Pre-training Through Seamless Data Packing
di: Yin, Ruicheng, et al.
Pubblicazione: (2025)
di: Yin, Ruicheng, et al.
Pubblicazione: (2025)
Insights into LLM Long-Context Failures: When Transformers Know but Don't Tell
di: Lu, Taiming, et al.
Pubblicazione: (2024)
di: Lu, Taiming, et al.
Pubblicazione: (2024)
Large Language Models Must Be Taught to Know What They Don't Know
di: Kapoor, Sanyam, et al.
Pubblicazione: (2024)
di: Kapoor, Sanyam, et al.
Pubblicazione: (2024)
"ChatGPT, Don't Tell Me What to Do": Designing AI for Context Analysis in Humanitarian Frontline Negotiations
di: Ma, ZIlin, et al.
Pubblicazione: (2024)
di: Ma, ZIlin, et al.
Pubblicazione: (2024)
First, Learn What You Don't Know: Active Information Gathering for Driving at the Limits of Handling
di: Davydov, Alexander, et al.
Pubblicazione: (2024)
di: Davydov, Alexander, et al.
Pubblicazione: (2024)
What We Know and What We Don't Know About the Function of γδ T Cells
di: Immo Prinz, et al.
Pubblicazione: (2025)
di: Immo Prinz, et al.
Pubblicazione: (2025)
Don't Say No: Jailbreaking LLM by Suppressing Refusal
di: Zhou, Yukai, et al.
Pubblicazione: (2024)
di: Zhou, Yukai, et al.
Pubblicazione: (2024)
Pulsed‐Field Ablation: What We Know and What We Don't
di: Sanghamitra Mohanty, et al.
Pubblicazione: (2025)
di: Sanghamitra Mohanty, et al.
Pubblicazione: (2025)
Bayesian Mixture-of-Experts: Towards Making LLMs Know What They Don't Know
di: Li, Albus Yizhuo
Pubblicazione: (2025)
di: Li, Albus Yizhuo
Pubblicazione: (2025)
But Don't Play With Me `Cause You're Playing With a Farmer – Electoral Consequences of Proposed Animal Welfare Reforms in Rural Poland
di: Karol Degórski, et al.
Pubblicazione: (2026)
di: Karol Degórski, et al.
Pubblicazione: (2026)
What Librarians Still Don't Know about Free Software
di: Chudnov, Daniel
Pubblicazione: (2009)
di: Chudnov, Daniel
Pubblicazione: (2009)
VIB-Probe: Detecting and Mitigating Hallucinations in Vision-Language Models via Variational Information Bottleneck
di: Zhang, Feiran, et al.
Pubblicazione: (2026)
di: Zhang, Feiran, et al.
Pubblicazione: (2026)
What You Don't Know Won't Hurt You: Self-Consistent Hierarchical Inference with Unknown Follow-up Selection Strategies
di: Essick, Reed, et al.
Pubblicazione: (2026)
di: Essick, Reed, et al.
Pubblicazione: (2026)
Show, Don't Tell: Evaluating Large Language Models Beyond Textual Understanding with ChildPlay
di: de Carvalho, Gonçalo Hora, et al.
Pubblicazione: (2024)
di: de Carvalho, Gonçalo Hora, et al.
Pubblicazione: (2024)
Language Models Don't Know What You Want: Evaluating Personalization in Deep Research Needs Real Users
di: Balepur, Nishant, et al.
Pubblicazione: (2026)
di: Balepur, Nishant, et al.
Pubblicazione: (2026)
How to Computerize Your Serials and Periodicals When You Don't Know How
di: Matthews, Mary, et al.
Pubblicazione: (1970)
di: Matthews, Mary, et al.
Pubblicazione: (1970)
Documenti analoghi
-
Advancing Parameter Efficiency in Fine-tuning via Representation Editing
di: Wu, Muling, et al.
Pubblicazione: (2024) -
Revisiting Jailbreaking for Large Language Models: A Representation Engineering Perspective
di: Li, Tianlong, et al.
Pubblicazione: (2024) -
Aligning Large Language Models with Human Preferences through Representation Engineering
di: Liu, Wenhao, et al.
Pubblicazione: (2023) -
UPLex: Fine-Grained Personality Control in Large Language Models via Unsupervised Lexical Modulation
di: Li, Tianlong, et al.
Pubblicazione: (2023) -
Promoting Data and Model Privacy in Federated Learning through Quantized LoRA
di: Zhu, JianHao, et al.
Pubblicazione: (2024)