Small Models Are (Still) Effective Cross-Domain Argument Extractors
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Gantt, William, White, Aaron Steven |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Event-Keyed Summarization
par: Gantt, William, et autres
Publié: (2024)
par: Gantt, William, et autres
Publié: (2024)
Limited Generalizability in Argument Mining: State-Of-The-Art Models Learn Datasets, Not Arguments
par: Feger, Marc, et autres
Publié: (2025)
par: Feger, Marc, et autres
Publié: (2025)
Small Language Model Makes an Effective Long Text Extractor
par: Chen, Yelin, et autres
Publié: (2025)
par: Chen, Yelin, et autres
Publié: (2025)
Domain-Adapted Small Language Models for Reliable Clinical Triage
par: Aljohani, Manar, et autres
Publié: (2026)
par: Aljohani, Manar, et autres
Publié: (2026)
Still "Talking About Large Language Models": Some Clarifications
par: Shanahan, Murray
Publié: (2024)
par: Shanahan, Murray
Publié: (2024)
Simple and Effective Masked Diffusion Language Models
par: Sahoo, Subham Sekhar, et autres
Publié: (2024)
par: Sahoo, Subham Sekhar, et autres
Publié: (2024)
Is Child-Directed Speech Effective Training Data for Language Models?
par: Feng, Steven Y., et autres
Publié: (2024)
par: Feng, Steven Y., et autres
Publié: (2024)
Frontier LLMs Still Struggle with Simple Reasoning Tasks
par: Malek, Alan, et autres
Publié: (2025)
par: Malek, Alan, et autres
Publié: (2025)
Winning Big with Small Models: Knowledge Distillation vs. Self-Training for Reducing Hallucination in Product QA Agents
par: Lewis, Ashley, et autres
Publié: (2025)
par: Lewis, Ashley, et autres
Publié: (2025)
Can Large Language Models Still Explain Themselves? Investigating the Impact of Quantization on Self-Explanations
par: Wang, Qianli, et autres
Publié: (2026)
par: Wang, Qianli, et autres
Publié: (2026)
Language Models Optimized to Fool Detectors Still Have a Distinct Style (And How to Change It)
par: Soto, Rafael Rivera, et autres
Publié: (2025)
par: Soto, Rafael Rivera, et autres
Publié: (2025)
When Domains Interact: Asymmetric and Order-Sensitive Cross-Domain Effects in Reinforcement Learning for Reasoning
par: Yang, Wang, et autres
Publié: (2026)
par: Yang, Wang, et autres
Publié: (2026)
A Logical Fallacy-Informed Framework for Argument Generation
par: Mouchel, Luca, et autres
Publié: (2024)
par: Mouchel, Luca, et autres
Publié: (2024)
Bringing Up a Bilingual BabyLM: Investigating Multilingual Language Acquisition Using Small-Scale Models
par: Zeng, Linda, et autres
Publié: (2026)
par: Zeng, Linda, et autres
Publié: (2026)
Learn from Weaknesses: Automated Domain Specialization for Small Computer-Use Agents
par: Kim, Suji, et autres
Publié: (2026)
par: Kim, Suji, et autres
Publié: (2026)
MLSD: A Novel Few-Shot Learning Approach to Enhance Cross-Target and Cross-Domain Stance Detection
par: Gera, Parush, et autres
Publié: (2025)
par: Gera, Parush, et autres
Publié: (2025)
A Reply to Makelov et al. (2023)'s "Interpretability Illusion" Arguments
par: Wu, Zhengxuan, et autres
Publié: (2024)
par: Wu, Zhengxuan, et autres
Publié: (2024)
X-Cross: Dynamic Integration of Language Models for Cross-Domain Sequential Recommendation
par: Hadad, Guy, et autres
Publié: (2025)
par: Hadad, Guy, et autres
Publié: (2025)
A Nurse is Blue and Elephant is Rugby: Cross Domain Alignment in Large Language Models Reveal Human-like Patterns
par: Yehudai, Asaf, et autres
Publié: (2024)
par: Yehudai, Asaf, et autres
Publié: (2024)
Probabilities of Chat LLMs Are Miscalibrated but Still Predict Correctness on Multiple-Choice Q&A
par: Plaut, Benjamin, et autres
Publié: (2024)
par: Plaut, Benjamin, et autres
Publié: (2024)
Your Next State-of-the-Art Could Come from Another Domain: A Cross-Domain Analysis of Hierarchical Text Classification
par: Li, Nan, et autres
Publié: (2024)
par: Li, Nan, et autres
Publié: (2024)
Climbing the Ladder of Reasoning: What LLMs Can-and Still Can't-Solve after SFT?
par: Sun, Yiyou, et autres
Publié: (2025)
par: Sun, Yiyou, et autres
Publié: (2025)
Revisiting Reinforcement Learning for LLM Reasoning from A Cross-Domain Perspective
par: Cheng, Zhoujun, et autres
Publié: (2025)
par: Cheng, Zhoujun, et autres
Publié: (2025)
OpenDebateEvidence: A Massive-Scale Argument Mining and Summarization Dataset
par: Roush, Allen, et autres
Publié: (2024)
par: Roush, Allen, et autres
Publié: (2024)
MultiMUC: Multilingual Template Filling on MUC-4
par: Gantt, William, et autres
Publié: (2024)
par: Gantt, William, et autres
Publié: (2024)
MixCE: Training Autoregressive Language Models by Mixing Forward and Reverse Cross-Entropies
par: Zhang, Shiyue, et autres
Publié: (2023)
par: Zhang, Shiyue, et autres
Publié: (2023)
CP-Prompt: Composition-Based Cross-modal Prompting for Domain-Incremental Continual Learning
par: Feng, Yu, et autres
Publié: (2024)
par: Feng, Yu, et autres
Publié: (2024)
Stacking Small Language Models for Generalizability
par: Liang, Laurence
Publié: (2024)
par: Liang, Laurence
Publié: (2024)
All Language Models Large and Small
par: Chen, Zhixun, et autres
Publié: (2024)
par: Chen, Zhixun, et autres
Publié: (2024)
Interpretable Cross-Examination Technique (ICE-T): Using highly informative features to boost LLM performance
par: Muric, Goran, et autres
Publié: (2024)
par: Muric, Goran, et autres
Publié: (2024)
Small Language Models: Survey, Measurements, and Insights
par: Lu, Zhenyan, et autres
Publié: (2024)
par: Lu, Zhenyan, et autres
Publié: (2024)
Squat: Quant Small Language Models on the Edge
par: Shen, Xuan, et autres
Publié: (2024)
par: Shen, Xuan, et autres
Publié: (2024)
Towards Reasoning Ability of Small Language Models
par: Srivastava, Gaurav, et autres
Publié: (2025)
par: Srivastava, Gaurav, et autres
Publié: (2025)
SmallToLarge (S2L): Scalable Data Selection for Fine-tuning Large Language Models by Summarizing Training Trajectories of Small Models
par: Yang, Yu, et autres
Publié: (2024)
par: Yang, Yu, et autres
Publié: (2024)
Language Representation Favored Zero-Shot Cross-Domain Cognitive Diagnosis
par: Liu, Shuo, et autres
Publié: (2025)
par: Liu, Shuo, et autres
Publié: (2025)
Are Small Language Models Ready to Compete with Large Language Models for Practical Applications?
par: Sinha, Neelabh, et autres
Publié: (2024)
par: Sinha, Neelabh, et autres
Publié: (2024)
Linear Representation Transferability Hypothesis: Leveraging Small Models to Steer Large Models
par: Bello, Femi, et autres
Publié: (2025)
par: Bello, Femi, et autres
Publié: (2025)
Simplifying Outcomes of Language Model Component Analyses with ELIA
par: Eidt, Aaron Louis, et autres
Publié: (2026)
par: Eidt, Aaron Louis, et autres
Publié: (2026)
Fox-1: Open Small Language Model for Cloud and Edge
par: Hu, Zijian, et autres
Publié: (2024)
par: Hu, Zijian, et autres
Publié: (2024)
Hymba: A Hybrid-head Architecture for Small Language Models
par: Dong, Xin, et autres
Publié: (2024)
par: Dong, Xin, et autres
Publié: (2024)
Documents similaires
-
Event-Keyed Summarization
par: Gantt, William, et autres
Publié: (2024) -
Limited Generalizability in Argument Mining: State-Of-The-Art Models Learn Datasets, Not Arguments
par: Feger, Marc, et autres
Publié: (2025) -
Small Language Model Makes an Effective Long Text Extractor
par: Chen, Yelin, et autres
Publié: (2025) -
Domain-Adapted Small Language Models for Reliable Clinical Triage
par: Aljohani, Manar, et autres
Publié: (2026) -
Still "Talking About Large Language Models": Some Clarifications
par: Shanahan, Murray
Publié: (2024)