Unlocking Structure Measuring: Introducing PDD, an Automatic Metric for Positional Discourse Coherence
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Yinhong, Su, Yixuan, Shareghi, Ehsan, Collier, Nigel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Aligning with Logic: Measuring, Evaluating and Improving Logical Preference Consistency in Large Language Models
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
Prompt Compression for Large Language Models: A Survey
von: Li, Zongqian, et al.
Veröffentlicht: (2024)
von: Li, Zongqian, et al.
Veröffentlicht: (2024)
ReasonGraph: Visualisation of Reasoning Paths
von: Li, Zongqian, et al.
Veröffentlicht: (2025)
von: Li, Zongqian, et al.
Veröffentlicht: (2025)
TOAD: Task-Oriented Automatic Dialogs with Diverse Response Styles
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
PiVe: Prompting with Iterative Verification Improving Graph-based Generative Capability of LLMs
von: Han, Jiuzhou, et al.
Veröffentlicht: (2023)
von: Han, Jiuzhou, et al.
Veröffentlicht: (2023)
Aligning with Human Judgement: The Role of Pairwise Preference in Large Language Model Evaluators
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
von: Liu, Yinhong, et al.
Veröffentlicht: (2024)
All Roads Lead to Rome: Graph-Based Confidence Estimation for Large Language Model Reasoning
von: Zhang, Caiqi, et al.
Veröffentlicht: (2025)
von: Zhang, Caiqi, et al.
Veröffentlicht: (2025)
TRIDENT: Benchmarking LLM Safety in Finance, Medicine, and Law
von: Hui, Zheng, et al.
Veröffentlicht: (2025)
von: Hui, Zheng, et al.
Veröffentlicht: (2025)
Privacy-R1: Privacy-Aware Multi-LLM Agent Collaboration via Reinforcement Learning
von: Hui, Zheng, et al.
Veröffentlicht: (2025)
von: Hui, Zheng, et al.
Veröffentlicht: (2025)
500xCompressor: Generalized Prompt Compression for Large Language Models
von: Li, Zongqian, et al.
Veröffentlicht: (2024)
von: Li, Zongqian, et al.
Veröffentlicht: (2024)
A Survey on Prompt Tuning
von: Li, Zongqian, et al.
Veröffentlicht: (2025)
von: Li, Zongqian, et al.
Veröffentlicht: (2025)
PT-MoE: An Efficient Finetuning Framework for Integrating Mixture-of-Experts into Prompt Tuning
von: Li, Zongqian, et al.
Veröffentlicht: (2025)
von: Li, Zongqian, et al.
Veröffentlicht: (2025)
Sparkles: Unlocking Chats Across Multiple Images for Multimodal Instruction-Following Models
von: Huang, Yupan, et al.
Veröffentlicht: (2023)
von: Huang, Yupan, et al.
Veröffentlicht: (2023)
Assessing the Sensitivity and Alignment of FOL Closeness Metrics
von: Thatikonda, Ramya Keerthy, et al.
Veröffentlicht: (2025)
von: Thatikonda, Ramya Keerthy, et al.
Veröffentlicht: (2025)
Equipping Language Models with Tool Use Capability for Tabular Data Analysis in Finance
von: Theuma, Adrian, et al.
Veröffentlicht: (2024)
von: Theuma, Adrian, et al.
Veröffentlicht: (2024)
When Personalization Meets Reality: A Multi-Faceted Analysis of Personalized Preference Learning
von: Dong, Yijiang River, et al.
Veröffentlicht: (2025)
von: Dong, Yijiang River, et al.
Veröffentlicht: (2025)
COFFEE: A Contrastive Oracle-Free Framework for Event Extraction
von: Zhang, Meiru, et al.
Veröffentlicht: (2023)
von: Zhang, Meiru, et al.
Veröffentlicht: (2023)
Cube Bench: A Benchmark for Spatial Visual Reasoning in MLLMs
von: Anand, Dhruv, et al.
Veröffentlicht: (2025)
von: Anand, Dhruv, et al.
Veröffentlicht: (2025)
Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments
von: Zhou, Han, et al.
Veröffentlicht: (2024)
von: Zhou, Han, et al.
Veröffentlicht: (2024)
One STEP at a time: Language Agents are Stepwise Planners
von: Nguyen, Minh, et al.
Veröffentlicht: (2024)
von: Nguyen, Minh, et al.
Veröffentlicht: (2024)
Towards Uncertainty-Aware Language Agent
von: Han, Jiuzhou, et al.
Veröffentlicht: (2024)
von: Han, Jiuzhou, et al.
Veröffentlicht: (2024)
Reward Engineering for Generating Semi-structured Explanation
von: Han, Jiuzhou, et al.
Veröffentlicht: (2023)
von: Han, Jiuzhou, et al.
Veröffentlicht: (2023)
Flexi-LoRA with Input-Adaptive Ranks: Efficient Finetuning for Speech and Reasoning Tasks
von: Li, Zongqian, et al.
Veröffentlicht: (2026)
von: Li, Zongqian, et al.
Veröffentlicht: (2026)
VerifiAgent: a Unified Verification Agent in Language Model Reasoning
von: Han, Jiuzhou, et al.
Veröffentlicht: (2025)
von: Han, Jiuzhou, et al.
Veröffentlicht: (2025)
Uncertainty-Based Methods for Automated Process Reward Data Construction and Output Aggregation in Mathematical Reasoning
von: Han, Jiuzhou, et al.
Veröffentlicht: (2025)
von: Han, Jiuzhou, et al.
Veröffentlicht: (2025)
Quantifying the Persona Effect in LLM Simulations
von: Hu, Tiancheng, et al.
Veröffentlicht: (2024)
von: Hu, Tiancheng, et al.
Veröffentlicht: (2024)
iNews: A Multimodal Dataset for Modeling Personalized Affective Responses to News
von: Hu, Tiancheng, et al.
Veröffentlicht: (2025)
von: Hu, Tiancheng, et al.
Veröffentlicht: (2025)
Jigsaw Puzzles: Splitting Harmful Questions to Jailbreak Large Language Models
von: Yang, Hao, et al.
Veröffentlicht: (2024)
von: Yang, Hao, et al.
Veröffentlicht: (2024)
The Compressor-Retriever Architecture for Language Model OS
von: Yang, Yuan, et al.
Veröffentlicht: (2024)
von: Yang, Yuan, et al.
Veröffentlicht: (2024)
Evaluating LLM-based Approaches to Legal Citation Prediction: Domain-specific Pre-training, Fine-tuning, or RAG? A Benchmark and an Australian Law Case Study
von: Han, Jiuzhou, et al.
Veröffentlicht: (2024)
von: Han, Jiuzhou, et al.
Veröffentlicht: (2024)
Logical Reasoning with Outcome Reward Models for Test-Time Scaling
von: Thatikonda, Ramya Keerthy, et al.
Veröffentlicht: (2025)
von: Thatikonda, Ramya Keerthy, et al.
Veröffentlicht: (2025)
Strategies for Improving NL-to-FOL Translation with LLMs: Data Generation, Incremental Fine-Tuning, and Verification
von: Thatikonda, Ramya Keerthy, et al.
Veröffentlicht: (2024)
von: Thatikonda, Ramya Keerthy, et al.
Veröffentlicht: (2024)
A Closer Look at Logical Reasoning with LLMs: The Choice of Tool Matters
von: Lam, Long Hei Matthew, et al.
Veröffentlicht: (2024)
von: Lam, Long Hei Matthew, et al.
Veröffentlicht: (2024)
Towards Inference-time Scaling for Continuous Space Reasoning
von: Wang, Minghan, et al.
Veröffentlicht: (2025)
von: Wang, Minghan, et al.
Veröffentlicht: (2025)
Attention Instruction: Amplifying Attention in the Middle via Prompting
von: Zhang, Meiru, et al.
Veröffentlicht: (2024)
von: Zhang, Meiru, et al.
Veröffentlicht: (2024)
Time to Revist Exact Match
von: Abbood, Auss, et al.
Veröffentlicht: (2025)
von: Abbood, Auss, et al.
Veröffentlicht: (2025)
Can LLMs Reason in the Wild with Programs?
von: Yang, Yuan, et al.
Veröffentlicht: (2024)
von: Yang, Yuan, et al.
Veröffentlicht: (2024)
LUQ: Long-text Uncertainty Quantification for LLMs
von: Zhang, Caiqi, et al.
Veröffentlicht: (2024)
von: Zhang, Caiqi, et al.
Veröffentlicht: (2024)
Improving Symbolic Translation of Language Models for Logical Reasoning
von: Thatikonda, Ramya Keerthy, et al.
Veröffentlicht: (2026)
von: Thatikonda, Ramya Keerthy, et al.
Veröffentlicht: (2026)
Towards Probing Speech-Specific Risks in Large Multimodal Models: A Taxonomy, Benchmark, and Insights
von: Yang, Hao, et al.
Veröffentlicht: (2024)
von: Yang, Hao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Aligning with Logic: Measuring, Evaluating and Improving Logical Preference Consistency in Large Language Models
von: Liu, Yinhong, et al.
Veröffentlicht: (2024) -
Prompt Compression for Large Language Models: A Survey
von: Li, Zongqian, et al.
Veröffentlicht: (2024) -
ReasonGraph: Visualisation of Reasoning Paths
von: Li, Zongqian, et al.
Veröffentlicht: (2025) -
TOAD: Task-Oriented Automatic Dialogs with Diverse Response Styles
von: Liu, Yinhong, et al.
Veröffentlicht: (2024) -
PiVe: Prompting with Iterative Verification Improving Graph-based Generative Capability of LLMs
von: Han, Jiuzhou, et al.
Veröffentlicht: (2023)