Can Your Model Tell a Negation from an Implicature? Unravelling Challenges With Intent Encoders
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yuwei, Singh, Siffi, Sengupta, Sailik, Shalyminov, Igor, Su, Hang, Song, Hwanjun, Mansour, Saab |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FineSurE: Fine-grained Summarization Evaluation using LLMs
by: Song, Hwanjun, et al.
Published: (2024)
by: Song, Hwanjun, et al.
Published: (2024)
Semi-Supervised Dialogue Abstractive Summarization via High-Quality Pseudolabel Selection
by: He, Jianfeng, et al.
Published: (2024)
by: He, Jianfeng, et al.
Published: (2024)
MAGID: An Automated Pipeline for Generating Synthetic Multi-modal Datasets
by: Aboutalebi, Hossein, et al.
Published: (2024)
by: Aboutalebi, Hossein, et al.
Published: (2024)
Controllable Conversational Theme Detection Track at DSTC 12
by: Shalyminov, Igor, et al.
Published: (2025)
by: Shalyminov, Igor, et al.
Published: (2025)
CERET: Cost-Effective Extrinsic Refinement for Text Generation
by: Cai, Jason, et al.
Published: (2024)
by: Cai, Jason, et al.
Published: (2024)
FLAP: Flow-Adhering Planning with Constrained Decoding in LLMs
by: Roy, Shamik, et al.
Published: (2024)
by: Roy, Shamik, et al.
Published: (2024)
TofuEval: Evaluating Hallucinations of LLMs on Topic-Focused Dialogue Summarization
by: Tang, Liyan, et al.
Published: (2024)
by: Tang, Liyan, et al.
Published: (2024)
GLEAN: Active Generalized Category Discovery with Diverse LLM Feedback
by: Zou, Henry Peng, et al.
Published: (2025)
by: Zou, Henry Peng, et al.
Published: (2025)
Faithful, Unfaithful or Ambiguous? Multi-Agent Debate with Initial Stance for Summary Evaluation
by: Koupaee, Mahnaz, et al.
Published: (2025)
by: Koupaee, Mahnaz, et al.
Published: (2025)
DeAL: Decoding-time Alignment for Large Language Models
by: Huang, James Y., et al.
Published: (2024)
by: Huang, James Y., et al.
Published: (2024)
Controllable Contextualized Image Captioning: Directing the Visual Narrative through User-Defined Highlights
by: Mao, Shunqi, et al.
Published: (2024)
by: Mao, Shunqi, et al.
Published: (2024)
Alignment Tuning for Large Language Models: A Data-Centric Lens on Alignment Data Pipelines
by: Song, Hwanjun
Published: (2026)
by: Song, Hwanjun
Published: (2026)
Implicature in Interaction: Understanding Implicature Improves Alignment in Human-LLM Interaction
by: Hota, Asutosh, et al.
Published: (2025)
by: Hota, Asutosh, et al.
Published: (2025)
Conversational Implicatures: Modelling Relevance Theory Probabilistically
by: Unger, Christoph, et al.
Published: (2025)
by: Unger, Christoph, et al.
Published: (2025)
Conjoined Predication and Scalar Implicature
by: Kandala, Ratna
Published: (2025)
by: Kandala, Ratna
Published: (2025)
UniSumEval: Towards Unified, Fine-Grained, Multi-Dimensional Summarization Evaluation for LLMs
by: Lee, Yuho, et al.
Published: (2024)
by: Lee, Yuho, et al.
Published: (2024)
InsTALL: Context-aware Instructional Task Assistance with Multi-modal Large Language Models
by: Nguyen, Pha, et al.
Published: (2025)
by: Nguyen, Pha, et al.
Published: (2025)
LLM-based User Profile Management for Recommender System
by: Bang, Seunghwan, et al.
Published: (2025)
by: Bang, Seunghwan, et al.
Published: (2025)
User Modeling Challenges in Interactive AI Assistant Systems
by: Su, Megan, et al.
Published: (2024)
by: Su, Megan, et al.
Published: (2024)
Chapter 3 Implicature and explicature
by: Carston, Robyn, et al.
Published: (2019)
by: Carston, Robyn, et al.
Published: (2019)
Tell Your Model Where to Attend: Post-hoc Attention Steering for LLMs
by: Zhang, Qingru, et al.
Published: (2023)
by: Zhang, Qingru, et al.
Published: (2023)
Do Large Language Models Understand Conversational Implicature -- A case study with a chinese sitcom
by: Yue, Shisen, et al.
Published: (2024)
by: Yue, Shisen, et al.
Published: (2024)
DRInQ: Evaluating Conversational Implicature with Controlled Context Variation
by: Arai, Hirona Jacqueline, et al.
Published: (2026)
by: Arai, Hirona Jacqueline, et al.
Published: (2026)
Structured List-Grounded Question Answering
by: Sung, Mujeen, et al.
Published: (2024)
by: Sung, Mujeen, et al.
Published: (2024)
RADIUS: Ranking, Distribution, and Significance - A Comprehensive Alignment Suite for Survey Simulation
by: Łajewska, Weronika, et al.
Published: (2026)
by: Łajewska, Weronika, et al.
Published: (2026)
Cross-Lingual LLM-Judge Transfer via Evaluation Decomposition
by: Sheth, Ivaxi, et al.
Published: (2026)
by: Sheth, Ivaxi, et al.
Published: (2026)
Your Model is Overconfident, and Other Lies We Tell Ourselves
by: Mickus, Timothee, et al.
Published: (2025)
by: Mickus, Timothee, et al.
Published: (2025)
Learning to Summarize from LLM-generated Feedback
by: Song, Hwanjun, et al.
Published: (2024)
by: Song, Hwanjun, et al.
Published: (2024)
DFlow: Diverse Dialogue Flow Simulation with Large Language Models
by: Du, Wanyu, et al.
Published: (2024)
by: Du, Wanyu, et al.
Published: (2024)
Aligning Extraction and Generation for Robust Retrieval-Augmented Generation
by: Song, Hwanjun, et al.
Published: (2025)
by: Song, Hwanjun, et al.
Published: (2025)
IntentGrasp: A Comprehensive Benchmark for Intent Understanding
by: Yin, Yuwei, et al.
Published: (2026)
by: Yin, Yuwei, et al.
Published: (2026)
SeedPrints: Fingerprints Can Even Tell Which Seed Your Large Language Model Was Trained From
by: Tong, Yao, et al.
Published: (2025)
by: Tong, Yao, et al.
Published: (2025)
Using Optimal Transport as Alignment Objective for fine-tuning Multilingual Contextualized Embeddings
by: Alqahtani, Sawsan, et al.
Published: (2021)
by: Alqahtani, Sawsan, et al.
Published: (2021)
Rethinking LLM-Based Recommendations: A Personalized Query-Driven Parallel Integration
by: Han, Donghee, et al.
Published: (2025)
by: Han, Donghee, et al.
Published: (2025)
In-Context Learning with Noisy Labels
by: Kang, Junyong, et al.
Published: (2024)
by: Kang, Junyong, et al.
Published: (2024)
AbsenceBench: Language Models Can't Tell What's Missing
by: Fu, Harvey Yiyun, et al.
Published: (2025)
by: Fu, Harvey Yiyun, et al.
Published: (2025)
SWI: Speaking with Intent in Large Language Models
by: Yin, Yuwei, et al.
Published: (2025)
by: Yin, Yuwei, et al.
Published: (2025)
MDSEval: A Meta-Evaluation Benchmark for Multimodal Dialogue Summarization
by: Liu, Yinhong, et al.
Published: (2025)
by: Liu, Yinhong, et al.
Published: (2025)
What Can String Probability Tell Us About Grammaticality?
by: Hu, Jennifer, et al.
Published: (2025)
by: Hu, Jennifer, et al.
Published: (2025)
Diffusion Is Your Friend in Show, Suggest and Tell
by: Hu, Jia Cheng, et al.
Published: (2025)
by: Hu, Jia Cheng, et al.
Published: (2025)
Similar Items
-
FineSurE: Fine-grained Summarization Evaluation using LLMs
by: Song, Hwanjun, et al.
Published: (2024) -
Semi-Supervised Dialogue Abstractive Summarization via High-Quality Pseudolabel Selection
by: He, Jianfeng, et al.
Published: (2024) -
MAGID: An Automated Pipeline for Generating Synthetic Multi-modal Datasets
by: Aboutalebi, Hossein, et al.
Published: (2024) -
Controllable Conversational Theme Detection Track at DSTC 12
by: Shalyminov, Igor, et al.
Published: (2025) -
CERET: Cost-Effective Extrinsic Refinement for Text Generation
by: Cai, Jason, et al.
Published: (2024)