Systematic Task Exploration with LLMs: A Study in Citation Text Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Şahinuç, Furkan, Kuznetsov, Ilia, Hou, Yufang, Gurevych, Iryna |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Expert Preference-based Evaluation of Automated Related Work Generation
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2025)
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2025)
Reward Modeling for Scientific Writing Evaluation
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2026)
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2026)
Efficient Performance Tracking: Leveraging Large Language Models for Automated Construction of Scientific Leaderboards
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2024)
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2024)
Are Large Language Models Good Classifiers? A Study on Edit Intent Classification in Scientific Document Revisions
von: Ruan, Qian, et al.
Veröffentlicht: (2024)
von: Ruan, Qian, et al.
Veröffentlicht: (2024)
How to Handle Different Types of Out-of-Distribution Scenarios in Computational Argumentation? A Comprehensive and Fine-Grained Field Study
von: Waldis, Andreas, et al.
Veröffentlicht: (2023)
von: Waldis, Andreas, et al.
Veröffentlicht: (2023)
Re3: A Holistic Framework and Dataset for Modeling Collaborative Document Revision
von: Ruan, Qian, et al.
Veröffentlicht: (2024)
von: Ruan, Qian, et al.
Veröffentlicht: (2024)
Dive into the Chasm: Probing the Gap between In- and Cross-Topic Generalization
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
Identifying Aspects in Peer Reviews
von: Lu, Sheng, et al.
Veröffentlicht: (2025)
von: Lu, Sheng, et al.
Veröffentlicht: (2025)
Citation Failure: Definition, Analysis and Efficient Mitigation
von: Buchmann, Jan, et al.
Veröffentlicht: (2025)
von: Buchmann, Jan, et al.
Veröffentlicht: (2025)
Is Peer Review Really in Decline? Analyzing Review Quality across Venues and Time
von: Kuznetsov, Ilia, et al.
Veröffentlicht: (2026)
von: Kuznetsov, Ilia, et al.
Veröffentlicht: (2026)
Exposía: Teaching and Assessment of Academic Writing Skills for Research Project Proposals and Peer Feedback
von: Zyska, Dennis, et al.
Veröffentlicht: (2026)
von: Zyska, Dennis, et al.
Veröffentlicht: (2026)
STRICTA: Structured Reasoning in Critical Text Assessment for Peer Review and Beyond
von: Dycke, Nils, et al.
Veröffentlicht: (2024)
von: Dycke, Nils, et al.
Veröffentlicht: (2024)
Missci: Reconstructing Fallacies in Misrepresented Science
von: Glockner, Max, et al.
Veröffentlicht: (2024)
von: Glockner, Max, et al.
Veröffentlicht: (2024)
Grounding Fallacies Misrepresenting Scientific Publications in Evidence
von: Glockner, Max, et al.
Veröffentlicht: (2024)
von: Glockner, Max, et al.
Veröffentlicht: (2024)
ABCD-LINK: Annotation Bootstrapping for Cross-Document Fine-Grained Links
von: Basch, Serwar, et al.
Veröffentlicht: (2025)
von: Basch, Serwar, et al.
Veröffentlicht: (2025)
The Nature of NLP: Analyzing Contributions in NLP Papers
von: Pramanick, Aniket, et al.
Veröffentlicht: (2024)
von: Pramanick, Aniket, et al.
Veröffentlicht: (2024)
Transforming Scholarly Landscapes: Influence of Large Language Models on Academic Fields beyond Computer Science
von: Pramanick, Aniket, et al.
Veröffentlicht: (2024)
von: Pramanick, Aniket, et al.
Veröffentlicht: (2024)
ClaimFlow: Tracing the Evolution of Scientific Claims in NLP
von: Pramanick, Aniket, et al.
Veröffentlicht: (2026)
von: Pramanick, Aniket, et al.
Veröffentlicht: (2026)
Holmes: A Benchmark to Assess the Linguistic Competence of Language Models
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
Document Structure in Long Document Transformers
von: Buchmann, Jan, et al.
Veröffentlicht: (2024)
von: Buchmann, Jan, et al.
Veröffentlicht: (2024)
Like a Good Nearest Neighbor: Practical Content Moderation and Text Classification
von: Bates, Luke, et al.
Veröffentlicht: (2023)
von: Bates, Luke, et al.
Veröffentlicht: (2023)
M2QA: Multi-domain Multilingual Question Answering
von: Engländer, Leon, et al.
Veröffentlicht: (2024)
von: Engländer, Leon, et al.
Veröffentlicht: (2024)
A Course Shared Task on Evaluating LLM Output for Clinical Questions
von: Hou, Yufang, et al.
Veröffentlicht: (2024)
von: Hou, Yufang, et al.
Veröffentlicht: (2024)
Code Prompting Elicits Conditional Reasoning Abilities in Text+Code LLMs
von: Puerto, Haritz, et al.
Veröffentlicht: (2024)
von: Puerto, Haritz, et al.
Veröffentlicht: (2024)
Overview of PerpectiveArg2024: The First Shared Task on Perspective Argument Retrieval
von: Falk, Neele, et al.
Veröffentlicht: (2024)
von: Falk, Neele, et al.
Veröffentlicht: (2024)
Robust Utility-Preserving Text Anonymization Based on Large Language Models
von: Yang, Tianyu, et al.
Veröffentlicht: (2024)
von: Yang, Tianyu, et al.
Veröffentlicht: (2024)
Automatic Reviewers Fail to Detect Faulty Reasoning in Research Papers: A New Counterfactual Evaluation Framework
von: Dycke, Nils, et al.
Veröffentlicht: (2025)
von: Dycke, Nils, et al.
Veröffentlicht: (2025)
The Inherent Limits of Pretrained LLMs: The Unexpected Convergence of Instruction Tuning and In-Context Learning Capabilities
von: Bigoulaeva, Irina, et al.
Veröffentlicht: (2025)
von: Bigoulaeva, Irina, et al.
Veröffentlicht: (2025)
Socratic Reasoning Improves Positive Text Rewriting
von: Goel, Anmol, et al.
Veröffentlicht: (2024)
von: Goel, Anmol, et al.
Veröffentlicht: (2024)
ChartAttack: Testing the Vulnerability of LLMs to Malicious Prompting in Chart Generation
von: Ortiz-Barajas, Jesus-German, et al.
Veröffentlicht: (2026)
von: Ortiz-Barajas, Jesus-German, et al.
Veröffentlicht: (2026)
MiDe22: An Annotated Multi-Event Tweet Dataset for Misinformation Detection
von: Toraman, Cagri, et al.
Veröffentlicht: (2022)
von: Toraman, Cagri, et al.
Veröffentlicht: (2022)
IRCoder: Intermediate Representations Make Language Models Robust Multilingual Code Generators
von: Paul, Indraneil, et al.
Veröffentlicht: (2024)
von: Paul, Indraneil, et al.
Veröffentlicht: (2024)
Learning from Implicit User Feedback, Emotions and Demographic Information in Task-Oriented and Document-Grounded Dialogues
von: Petrak, Dominic, et al.
Veröffentlicht: (2024)
von: Petrak, Dominic, et al.
Veröffentlicht: (2024)
The Lou Dataset -- Exploring the Impact of Gender-Fair Language in German Text Classification
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
von: Waldis, Andreas, et al.
Veröffentlicht: (2024)
SciCoQA: Quality Assurance for Scientific Paper--Code Alignment
von: Baumgärtner, Tim, et al.
Veröffentlicht: (2026)
von: Baumgärtner, Tim, et al.
Veröffentlicht: (2026)
Are Multilingual LLMs Culturally-Diverse Reasoners? An Investigation into Multicultural Proverbs and Sayings
von: Liu, Chen Cecilia, et al.
Veröffentlicht: (2023)
von: Liu, Chen Cecilia, et al.
Veröffentlicht: (2023)
Pull Requests From The Classroom: Co-Developing Curriculum And Code
von: Zyska, Dennis, et al.
Veröffentlicht: (2025)
von: Zyska, Dennis, et al.
Veröffentlicht: (2025)
Hierarchical Latent Structures in Data Generation Process Unify Mechanistic Phenomena across Scale
von: Rohweder, Jonas, et al.
Veröffentlicht: (2026)
von: Rohweder, Jonas, et al.
Veröffentlicht: (2026)
MAGneT: Coordinated Multi-Agent Generation of Synthetic Multi-Turn Mental Health Counseling Sessions
von: Mandal, Aishik, et al.
Veröffentlicht: (2025)
von: Mandal, Aishik, et al.
Veröffentlicht: (2025)
LLMs as Cultural Archives: Cultural Commonsense Knowledge Graph Extraction
von: Tonga, Junior Cedric, et al.
Veröffentlicht: (2026)
von: Tonga, Junior Cedric, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Expert Preference-based Evaluation of Automated Related Work Generation
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2025) -
Reward Modeling for Scientific Writing Evaluation
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2026) -
Efficient Performance Tracking: Leveraging Large Language Models for Automated Construction of Scientific Leaderboards
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2024) -
Are Large Language Models Good Classifiers? A Study on Edit Intent Classification in Scientific Document Revisions
von: Ruan, Qian, et al.
Veröffentlicht: (2024) -
How to Handle Different Types of Out-of-Distribution Scenarios in Computational Argumentation? A Comprehensive and Fine-Grained Field Study
von: Waldis, Andreas, et al.
Veröffentlicht: (2023)