Limited Generalizability in Argument Mining: State-Of-The-Art Models Learn Datasets, Not Arguments
Fuente:
arXiv
Saved in:
| Main Authors: | Feger, Marc, Boland, Katarina, Dietze, Stefan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TACO -- Twitter Arguments from COnversations
by: Feger, Marc, et al.
Published: (2024)
by: Feger, Marc, et al.
Published: (2024)
OpenDebateEvidence: A Massive-Scale Argument Mining and Summarization Dataset
by: Roush, Allen, et al.
Published: (2024)
by: Roush, Allen, et al.
Published: (2024)
Small Models Are (Still) Effective Cross-Domain Argument Extractors
by: Gantt, William, et al.
Published: (2024)
by: Gantt, William, et al.
Published: (2024)
Understanding Enthymemes in Argument Maps: Bridging Argument Mining and Logic-based Argumentation
by: Ben-Naim, Jonathan, et al.
Published: (2024)
by: Ben-Naim, Jonathan, et al.
Published: (2024)
A Logical Fallacy-Informed Framework for Argument Generation
by: Mouchel, Luca, et al.
Published: (2024)
by: Mouchel, Luca, et al.
Published: (2024)
End-to-End Argument Mining through Autoregressive Argumentative Structure Prediction
by: Das, Nilmadhab, et al.
Published: (2025)
by: Das, Nilmadhab, et al.
Published: (2025)
A Reply to Makelov et al. (2023)'s "Interpretability Illusion" Arguments
by: Wu, Zhengxuan, et al.
Published: (2024)
by: Wu, Zhengxuan, et al.
Published: (2024)
In-Context Learning and Fine-Tuning GPT for Argument Mining
by: Cabessa, Jérémie, et al.
Published: (2024)
by: Cabessa, Jérémie, et al.
Published: (2024)
No Argument Left Behind: Overlapping Chunks for Faster Processing of Arbitrarily Long Legal Texts
by: Fama, Israel, et al.
Published: (2024)
by: Fama, Israel, et al.
Published: (2024)
A Hybrid Intelligence Method for Argument Mining
by: van der Meer, Michiel, et al.
Published: (2024)
by: van der Meer, Michiel, et al.
Published: (2024)
Breaking Down Bias: On The Limits of Generalizable Pruning Strategies
by: Ma, Sibo, et al.
Published: (2025)
by: Ma, Sibo, et al.
Published: (2025)
Measuring Faithfulness and Abstention: An Automated Pipeline for Evaluating LLM-Generated 3-ply Case-Based Legal Arguments
by: Zhang, Li, et al.
Published: (2025)
by: Zhang, Li, et al.
Published: (2025)
Towards Generalizable Generic Harmful Speech Datasets for Implicit Hate Speech Detection
by: Almohaimeed, Saad, et al.
Published: (2025)
by: Almohaimeed, Saad, et al.
Published: (2025)
Uncovering Latent Arguments in Social Media Messaging by Employing LLMs-in-the-Loop Strategy
by: Islam, Tunazzina, et al.
Published: (2024)
by: Islam, Tunazzina, et al.
Published: (2024)
LASP: Surveying the State-of-the-Art in Large Language Model-Assisted AI Planning
by: Li, Haoming, et al.
Published: (2024)
by: Li, Haoming, et al.
Published: (2024)
FinGen: A Dataset for Argument Generation in Finance
by: Chen, Chung-Chi, et al.
Published: (2024)
by: Chen, Chung-Chi, et al.
Published: (2024)
Stacking Small Language Models for Generalizability
by: Liang, Laurence
Published: (2024)
by: Liang, Laurence
Published: (2024)
Persona-Based Conversational AI: State of the Art and Challenges
by: Liu, Junfeng, et al.
Published: (2022)
by: Liu, Junfeng, et al.
Published: (2022)
Abstractive Text Summarization: State of the Art, Challenges, and Improvements
by: Shakil, Hassan, et al.
Published: (2024)
by: Shakil, Hassan, et al.
Published: (2024)
Exploration of Marker-Based Approaches in Argument Mining through Augmented Natural Language
by: Das, Nilmadhab, et al.
Published: (2024)
by: Das, Nilmadhab, et al.
Published: (2024)
A State-of-the-Art SQL Reasoning Model using RLVR
by: Ali, Alnur, et al.
Published: (2025)
by: Ali, Alnur, et al.
Published: (2025)
Large Language Models in Cybersecurity: State-of-the-Art
by: Motlagh, Farzad Nourmohammadzadeh, et al.
Published: (2024)
by: Motlagh, Farzad Nourmohammadzadeh, et al.
Published: (2024)
Can Large Language Models perform Relation-based Argument Mining?
by: Gorur, Deniz, et al.
Published: (2024)
by: Gorur, Deniz, et al.
Published: (2024)
Can Large Language Models Play Text Games Well? Current State-of-the-Art and Open Questions
by: Tsai, Chen Feng, et al.
Published: (2023)
by: Tsai, Chen Feng, et al.
Published: (2023)
Alice in Wonderland: Simple Tasks Showing Complete Reasoning Breakdown in State-Of-the-Art Large Language Models
by: Nezhurina, Marianna, et al.
Published: (2024)
by: Nezhurina, Marianna, et al.
Published: (2024)
Argument Mining in Data Scarce Settings: Cross-lingual Transfer and Few-shot Techniques
by: Yeginbergen, Anar, et al.
Published: (2024)
by: Yeginbergen, Anar, et al.
Published: (2024)
Mitigating Manipulation and Enhancing Persuasion: A Reflective Multi-Agent Approach for Legal Argument Generation
by: Zhang, Li, et al.
Published: (2025)
by: Zhang, Li, et al.
Published: (2025)
SQBC: Active Learning using LLM-Generated Synthetic Data for Stance Detection in Online Political Discussions
by: Wagner, Stefan Sylvius, et al.
Published: (2024)
by: Wagner, Stefan Sylvius, et al.
Published: (2024)
Object-Centric Neuro-Argumentative Learning
by: Jacob, Abdul Rahman, et al.
Published: (2025)
by: Jacob, Abdul Rahman, et al.
Published: (2025)
RewardAnything: Generalizable Principle-Following Reward Models
by: Yu, Zhuohao, et al.
Published: (2025)
by: Yu, Zhuohao, et al.
Published: (2025)
Large Language Models as Generalizable Policies for Embodied Tasks
by: Szot, Andrew, et al.
Published: (2023)
by: Szot, Andrew, et al.
Published: (2023)
AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset
by: Moshkov, Ivan, et al.
Published: (2025)
by: Moshkov, Ivan, et al.
Published: (2025)
The Argument is the Explanation: Structured Argumentation for Trust in Agents
by: Cakar, Ege, et al.
Published: (2025)
by: Cakar, Ege, et al.
Published: (2025)
Mining Intrinsic Rewards from LLM Hidden States for Efficient Best-of-N Sampling
by: Guo, Jizhou, et al.
Published: (2025)
by: Guo, Jizhou, et al.
Published: (2025)
Limits of Transformer Language Models on Learning to Compose Algorithms
by: Thomm, Jonathan, et al.
Published: (2024)
by: Thomm, Jonathan, et al.
Published: (2024)
Qalb: Largest State-of-the-Art Urdu Large Language Model for 230M Speakers with Systematic Continued Pre-training
by: Hassan, Muhammad Taimoor, et al.
Published: (2026)
by: Hassan, Muhammad Taimoor, et al.
Published: (2026)
Small LLMs Do Not Learn a Generalizable Theory of Mind via Reinforcement Learning
by: Sarangi, Sneheel, et al.
Published: (2025)
by: Sarangi, Sneheel, et al.
Published: (2025)
Boosting Protein Language Models with Negative Sample Mining
by: Xu, Yaoyao, et al.
Published: (2024)
by: Xu, Yaoyao, et al.
Published: (2024)
Mind the Gap: A Review of Arabic Post-Training Datasets and Their Limitations
by: Alkhowaiter, Mohammed, et al.
Published: (2025)
by: Alkhowaiter, Mohammed, et al.
Published: (2025)
Generalizable and Stable Finetuning of Pretrained Language Models on Low-Resource Texts
by: Somayajula, Sai Ashish, et al.
Published: (2024)
by: Somayajula, Sai Ashish, et al.
Published: (2024)
Similar Items
-
TACO -- Twitter Arguments from COnversations
by: Feger, Marc, et al.
Published: (2024) -
OpenDebateEvidence: A Massive-Scale Argument Mining and Summarization Dataset
by: Roush, Allen, et al.
Published: (2024) -
Small Models Are (Still) Effective Cross-Domain Argument Extractors
by: Gantt, William, et al.
Published: (2024) -
Understanding Enthymemes in Argument Maps: Bridging Argument Mining and Logic-based Argumentation
by: Ben-Naim, Jonathan, et al.
Published: (2024) -
A Logical Fallacy-Informed Framework for Argument Generation
by: Mouchel, Luca, et al.
Published: (2024)