Overview of PerpectiveArg2024: The First Shared Task on Perspective Argument Retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Falk, Neele, Waldis, Andreas, Gurevych, Iryna |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
How to Handle Different Types of Out-of-Distribution Scenarios in Computational Argumentation? A Comprehensive and Fine-Grained Field Study
von: Waldis, Andreas, et al.
Veröffentlicht: (2023)
von: Waldis, Andreas, et al.
Veröffentlicht: (2023)
Diversity Over Size: On the Effect of Sample and Topic Sizes for Topic-Dependent Argument Mining Datasets
von: Schiller, Benjamin, et al.
Veröffentlicht: (2022)
von: Schiller, Benjamin, et al.
Veröffentlicht: (2022)
Dive into the Chasm: Probing the Gap between In- and Cross-Topic Generalization
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
The Lou Dataset -- Exploring the Impact of Gender-Fair Language in German Text Classification
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
Holmes: A Benchmark to Assess the Linguistic Competence of Language Models
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
Aligned Probing: Relating Toxic Behavior and Model Internals
von: Waldis, Andreas, et al.
Veröffentlicht: (2025)
von: Waldis, Andreas, et al.
Veröffentlicht: (2025)
Investigating Subjective Factors of Argument Strength: Storytelling, Emotions, and Hedging
von: Quensel, Carlotta, et al.
Veröffentlicht: (2025)
von: Quensel, Carlotta, et al.
Veröffentlicht: (2025)
ArgBench: Benchmarking LLMs on Computational Argumentation Tasks
von: Ajjour, Yamen, et al.
Veröffentlicht: (2026)
von: Ajjour, Yamen, et al.
Veröffentlicht: (2026)
DAPR: A Benchmark on Document-Aware Passage Retrieval
von: Wang, Kexin, et al.
Veröffentlicht: (2023)
von: Wang, Kexin, et al.
Veröffentlicht: (2023)
Automatic Reviewers Fail to Detect Faulty Reasoning in Research Papers: A New Counterfactual Evaluation Framework
von: Dycke, Nils, et al.
Veröffentlicht: (2025)
von: Dycke, Nils, et al.
Veröffentlicht: (2025)
Like a Good Nearest Neighbor: Practical Content Moderation and Text Classification
von: Bates, Luke, et al.
Veröffentlicht: (2023)
von: Bates, Luke, et al.
Veröffentlicht: (2023)
Citation Failure: Definition, Analysis and Efficient Mitigation
von: Buchmann, Jan, et al.
Veröffentlicht: (2025)
von: Buchmann, Jan, et al.
Veröffentlicht: (2025)
GRITHopper: Decomposition-Free Multi-Hop Dense Retrieval
von: Erker, Justus-Jonas, et al.
Veröffentlicht: (2025)
von: Erker, Justus-Jonas, et al.
Veröffentlicht: (2025)
Systematic Task Exploration with LLMs: A Study in Citation Text Generation
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2024)
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2024)
Annotator-Centric Active Learning for Subjective NLP Tasks
von: van der Meer, Michiel, et al.
Veröffentlicht: (2024)
von: van der Meer, Michiel, et al.
Veröffentlicht: (2024)
A Course Shared Task on Evaluating LLM Output for Clinical Questions
von: Hou, Yufang, et al.
Veröffentlicht: (2024)
von: Hou, Yufang, et al.
Veröffentlicht: (2024)
Learning from Implicit User Feedback, Emotions and Demographic Information in Task-Oriented and Document-Grounded Dialogues
von: Petrak, Dominic, et al.
Veröffentlicht: (2024)
von: Petrak, Dominic, et al.
Veröffentlicht: (2024)
Overview of the NLPCC 2024 Shared Task on Chinese Metaphor Generation
von: Qu, Xingwei, et al.
Veröffentlicht: (2024)
von: Qu, Xingwei, et al.
Veröffentlicht: (2024)
SciCoQA: Quality Assurance for Scientific Paper--Code Alignment
von: Baumgärtner, Tim, et al.
Veröffentlicht: (2026)
von: Baumgärtner, Tim, et al.
Veröffentlicht: (2026)
A Pipeline to Assess Merging Methods via Behavior and Internals
von: Sigrist, Yutaro, et al.
Veröffentlicht: (2025)
von: Sigrist, Yutaro, et al.
Veröffentlicht: (2025)
ArgCMV: An Argument Summarization Benchmark for the LLM-era
von: Gurjar, Omkar, et al.
Veröffentlicht: (2025)
von: Gurjar, Omkar, et al.
Veröffentlicht: (2025)
Commitment Checklist: Auditing Author Commitments in Peer Review
von: Chen, Chung-Chi, et al.
Veröffentlicht: (2026)
von: Chen, Chung-Chi, et al.
Veröffentlicht: (2026)
Overview of AI-Debater 2023: The Challenges of Argument Generation Tasks
von: Lin, Jiayu, et al.
Veröffentlicht: (2024)
von: Lin, Jiayu, et al.
Veröffentlicht: (2024)
Robust Utility-Preserving Text Anonymization Based on Large Language Models
von: Yang, Tianyu, et al.
Veröffentlicht: (2024)
von: Yang, Tianyu, et al.
Veröffentlicht: (2024)
Re3: A Holistic Framework and Dataset for Modeling Collaborative Document Revision
von: Ruan, Qian, et al.
Veröffentlicht: (2024)
von: Ruan, Qian, et al.
Veröffentlicht: (2024)
DARA: Decomposition-Alignment-Reasoning Autonomous Language Agent for Question Answering over Knowledge Graphs
von: Fang, Haishuo, et al.
Veröffentlicht: (2024)
von: Fang, Haishuo, et al.
Veröffentlicht: (2024)
LLM Roleplay: Simulating Human-Chatbot Interaction
von: Tamoyan, Hovhannes, et al.
Veröffentlicht: (2024)
von: Tamoyan, Hovhannes, et al.
Veröffentlicht: (2024)
Are Large Language Models Good Classifiers? A Study on Edit Intent Classification in Scientific Document Revisions
von: Ruan, Qian, et al.
Veröffentlicht: (2024)
von: Ruan, Qian, et al.
Veröffentlicht: (2024)
Attribute or Abstain: Large Language Models as Long Document Assistants
von: Buchmann, Jan, et al.
Veröffentlicht: (2024)
von: Buchmann, Jan, et al.
Veröffentlicht: (2024)
Identifying Aspects in Peer Reviews
von: Lu, Sheng, et al.
Veröffentlicht: (2025)
von: Lu, Sheng, et al.
Veröffentlicht: (2025)
Token Weighting for Long-Range Language Modeling
von: Helm, Falko, et al.
Veröffentlicht: (2025)
von: Helm, Falko, et al.
Veröffentlicht: (2025)
COVE: COntext and VEracity prediction for out-of-context images
von: Tonglet, Jonathan, et al.
Veröffentlicht: (2025)
von: Tonglet, Jonathan, et al.
Veröffentlicht: (2025)
Turning Logic Against Itself : Probing Model Defenses Through Contrastive Questions
von: Sachdeva, Rachneet, et al.
Veröffentlicht: (2025)
von: Sachdeva, Rachneet, et al.
Veröffentlicht: (2025)
Expert Preference-based Evaluation of Automated Related Work Generation
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2025)
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2025)
M4FC: a Multimodal, Multilingual, Multicultural, Multitask Real-World Fact-Checking Dataset
von: Geng, Jiahui, et al.
Veröffentlicht: (2025)
von: Geng, Jiahui, et al.
Veröffentlicht: (2025)
CATfOOD: Counterfactual Augmented Training for Improving Out-of-Domain Performance and Calibration
von: Sachdeva, Rachneet, et al.
Veröffentlicht: (2023)
von: Sachdeva, Rachneet, et al.
Veröffentlicht: (2023)
How are Prompts Different in Terms of Sensitivity?
von: Lu, Sheng, et al.
Veröffentlicht: (2023)
von: Lu, Sheng, et al.
Veröffentlicht: (2023)
Reward Modeling for Scientific Writing Evaluation
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2026)
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2026)
IRCoder: Intermediate Representations Make Language Models Robust Multilingual Code Generators
von: Paul, Indraneil, et al.
Veröffentlicht: (2024)
von: Paul, Indraneil, et al.
Veröffentlicht: (2024)
RIRAG: Regulatory Information Retrieval and Answer Generation
von: Gokhan, Tuba, et al.
Veröffentlicht: (2024)
von: Gokhan, Tuba, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
How to Handle Different Types of Out-of-Distribution Scenarios in Computational Argumentation? A Comprehensive and Fine-Grained Field Study
von: Waldis, Andreas, et al.
Veröffentlicht: (2023) -
Diversity Over Size: On the Effect of Sample and Topic Sizes for Topic-Dependent Argument Mining Datasets
von: Schiller, Benjamin, et al.
Veröffentlicht: (2022) -
Dive into the Chasm: Probing the Gap between In- and Cross-Topic Generalization
von: Waldis, Andreas, et al.
Veröffentlicht: (2024) -
The Lou Dataset -- Exploring the Impact of Gender-Fair Language in German Text Classification
von: Waldis, Andreas, et al.
Veröffentlicht: (2024) -
Holmes: A Benchmark to Assess the Linguistic Competence of Language Models
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)