Improving Probability-based Prompt Selection Through Unified Evaluation and Analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Sohee, Kim, Jonghyeon, Jang, Joel, Ye, Seonghyeon, Lee, Hyunji, Seo, Minjoon |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
INSTRUCTIR: A Benchmark for Instruction Following of Information Retrieval Models
von: Oh, Hanseok, et al.
Veröffentlicht: (2024)
von: Oh, Hanseok, et al.
Veröffentlicht: (2024)
How Do Large Language Models Acquire Factual Knowledge During Pretraining?
von: Chang, Hoyeon, et al.
Veröffentlicht: (2024)
von: Chang, Hoyeon, et al.
Veröffentlicht: (2024)
How Well Do Large Language Models Truly Ground?
von: Lee, Hyunji, et al.
Veröffentlicht: (2023)
von: Lee, Hyunji, et al.
Veröffentlicht: (2023)
Self-Explore: Enhancing Mathematical Reasoning in Language Models with Fine-grained Rewards
von: Hwang, Hyeonbin, et al.
Veröffentlicht: (2024)
von: Hwang, Hyeonbin, et al.
Veröffentlicht: (2024)
Semiparametric Token-Sequence Co-Supervision
von: Lee, Hyunji, et al.
Veröffentlicht: (2024)
von: Lee, Hyunji, et al.
Veröffentlicht: (2024)
Differential Information Distribution: A Bayesian Perspective on Direct Preference Optimization
von: Won, Yunjae, et al.
Veröffentlicht: (2025)
von: Won, Yunjae, et al.
Veröffentlicht: (2025)
FLASK: Fine-grained Language Model Evaluation based on Alignment Skill Sets
von: Ye, Seonghyeon, et al.
Veröffentlicht: (2023)
von: Ye, Seonghyeon, et al.
Veröffentlicht: (2023)
Exploring the Practicality of Generative Retrieval on Dynamic Corpora
von: Kim, Chaeeun, et al.
Veröffentlicht: (2023)
von: Kim, Chaeeun, et al.
Veröffentlicht: (2023)
KTRL+F: Knowledge-Augmented In-Document Search
von: Oh, Hanseok, et al.
Veröffentlicht: (2023)
von: Oh, Hanseok, et al.
Veröffentlicht: (2023)
Knowledge Entropy Decay during Language Model Pretraining Hinders New Knowledge Acquisition
von: Kim, Jiyeon, et al.
Veröffentlicht: (2024)
von: Kim, Jiyeon, et al.
Veröffentlicht: (2024)
Understanding and Enhancing Mamba-Transformer Hybrids for Memory Recall and Language Modeling
von: Lee, Hyunji, et al.
Veröffentlicht: (2025)
von: Lee, Hyunji, et al.
Veröffentlicht: (2025)
LangBridge: Multilingual Reasoning Without Multilingual Supervision
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2024)
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2024)
How Does Vision-Language Adaptation Impact the Safety of Vision Language Models?
von: Lee, Seongyun, et al.
Veröffentlicht: (2024)
von: Lee, Seongyun, et al.
Veröffentlicht: (2024)
KoDialogBench: Evaluating Conversational Understanding of Language Models with Korean Dialogue Benchmark
von: Jang, Seongbo, et al.
Veröffentlicht: (2024)
von: Jang, Seongbo, et al.
Veröffentlicht: (2024)
On Efficient Language and Vision Assistants for Visually-Situated Natural Language Understanding: What Matters in Reading and Reasoning
von: Kim, Geewook, et al.
Veröffentlicht: (2024)
von: Kim, Geewook, et al.
Veröffentlicht: (2024)
Prometheus-Vision: Vision-Language Model as a Judge for Fine-Grained Evaluation
von: Lee, Seongyun, et al.
Veröffentlicht: (2024)
von: Lee, Seongyun, et al.
Veröffentlicht: (2024)
Efficient Prompt Optimisation for Legal Text Classification with Proxy Prompt Evaluator
von: Lee, Hyunji, et al.
Veröffentlicht: (2025)
von: Lee, Hyunji, et al.
Veröffentlicht: (2025)
Rethinking the Role of Proxy Rewards in Language Model Alignment
von: Kim, Sungdong, et al.
Veröffentlicht: (2024)
von: Kim, Sungdong, et al.
Veröffentlicht: (2024)
How language models extrapolate outside the training data: A case study in Textualized Gridworld
von: Kim, Doyoung, et al.
Veröffentlicht: (2024)
von: Kim, Doyoung, et al.
Veröffentlicht: (2024)
TSLM: Tree-Structured Language Modeling for Divergent Thinking
von: Kim, Doyoung, et al.
Veröffentlicht: (2026)
von: Kim, Doyoung, et al.
Veröffentlicht: (2026)
Reasoning Models Better Express Their Confidence
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2025)
von: Yoon, Dongkeun, et al.
Veröffentlicht: (2025)
Generative Prompt Internalization
von: Shin, Haebin, et al.
Veröffentlicht: (2024)
von: Shin, Haebin, et al.
Veröffentlicht: (2024)
Aligning to Thousands of Preferences via System Message Generalization
von: Lee, Seongyun, et al.
Veröffentlicht: (2024)
von: Lee, Seongyun, et al.
Veröffentlicht: (2024)
Prometheus: Inducing Fine-grained Evaluation Capability in Language Models
von: Kim, Seungone, et al.
Veröffentlicht: (2023)
von: Kim, Seungone, et al.
Veröffentlicht: (2023)
Can Large Language Models Keep Up? Benchmarking Online Adaptation to Continual Knowledge Streams
von: Kim, Jiyeon, et al.
Veröffentlicht: (2026)
von: Kim, Jiyeon, et al.
Veröffentlicht: (2026)
Latent Reasoning via Sentence Embedding Prediction
von: Hwang, Hyeonbin, et al.
Veröffentlicht: (2025)
von: Hwang, Hyeonbin, et al.
Veröffentlicht: (2025)
On the Effectiveness of Integration Methods for Multimodal Dialogue Response Retrieval
von: Jang, Seongbo, et al.
Veröffentlicht: (2025)
von: Jang, Seongbo, et al.
Veröffentlicht: (2025)
Lost in the Noise: How Reasoning Models Fail with Contextual Distractors
von: Lee, Seongyun, et al.
Veröffentlicht: (2026)
von: Lee, Seongyun, et al.
Veröffentlicht: (2026)
Early Decisions Matter: Proximity Bias and Initial Trajectory Shaping in Non-Autoregressive Diffusion Language Models
von: Kim, Jiyeon, et al.
Veröffentlicht: (2026)
von: Kim, Jiyeon, et al.
Veröffentlicht: (2026)
Maximizing Mutual Information Between Prompt and Response Improves LLM Performance With No Additional Data
von: Nam, Hyunji, et al.
Veröffentlicht: (2026)
von: Nam, Hyunji, et al.
Veröffentlicht: (2026)
Instruction Tuning with and without Context: Behavioral Shifts and Downstream Impact
von: Lee, Hyunji, et al.
Veröffentlicht: (2025)
von: Lee, Hyunji, et al.
Veröffentlicht: (2025)
Verbosity-Aware Rationale Reduction: Effective Reduction of Redundant Rationale via Principled Criteria
von: Jang, Joonwon, et al.
Veröffentlicht: (2024)
von: Jang, Joonwon, et al.
Veröffentlicht: (2024)
"As Eastern Powers, I will veto." : An Investigation of Nation-level Bias of Large Language Models in International Relations
von: Choi, Jonghyeon, et al.
Veröffentlicht: (2025)
von: Choi, Jonghyeon, et al.
Veröffentlicht: (2025)
Exploring Language Model's Code Generation Ability with Auxiliary Functions
von: Lee, Seonghyeon, et al.
Veröffentlicht: (2024)
von: Lee, Seonghyeon, et al.
Veröffentlicht: (2024)
Instruction Matters: A Simple yet Effective Task Selection for Optimized Instruction Tuning of Specific Tasks
von: Lee, Changho, et al.
Veröffentlicht: (2024)
von: Lee, Changho, et al.
Veröffentlicht: (2024)
Volcano: Mitigating Multimodal Hallucination through Self-Feedback Guided Revision
von: Lee, Seongyun, et al.
Veröffentlicht: (2023)
von: Lee, Seongyun, et al.
Veröffentlicht: (2023)
UniKnow: A Unified Framework for Reliable Language Model Behavior across Parametric and External Knowledge
von: Kim, Youna, et al.
Veröffentlicht: (2025)
von: Kim, Youna, et al.
Veröffentlicht: (2025)
Data-Driven Mispronunciation Pattern Discovery for Robust Speech Recognition
von: Choi, Anna Seo Gyeong, et al.
Veröffentlicht: (2025)
von: Choi, Anna Seo Gyeong, et al.
Veröffentlicht: (2025)
From What to Respond to When to Respond: Timely Response Generation for Open-domain Dialogue Agents
von: Jang, Seongbo, et al.
Veröffentlicht: (2025)
von: Jang, Seongbo, et al.
Veröffentlicht: (2025)
RoParQ: Paraphrase-Aware Alignment of Large Language Models Towards Robustness to Paraphrased Questions
von: Choi, Minjoon
Veröffentlicht: (2025)
von: Choi, Minjoon
Veröffentlicht: (2025)
Ähnliche Einträge
-
INSTRUCTIR: A Benchmark for Instruction Following of Information Retrieval Models
von: Oh, Hanseok, et al.
Veröffentlicht: (2024) -
How Do Large Language Models Acquire Factual Knowledge During Pretraining?
von: Chang, Hoyeon, et al.
Veröffentlicht: (2024) -
How Well Do Large Language Models Truly Ground?
von: Lee, Hyunji, et al.
Veröffentlicht: (2023) -
Self-Explore: Enhancing Mathematical Reasoning in Language Models with Fine-grained Rewards
von: Hwang, Hyeonbin, et al.
Veröffentlicht: (2024) -
Semiparametric Token-Sequence Co-Supervision
von: Lee, Hyunji, et al.
Veröffentlicht: (2024)