Cuckoo: An IE Free Rider Hatched by Massive Nutrition in LLM's Nest
Fuente:
arXiv
Saved in:
| Main Authors: | Peng, Letian, Wang, Zilong, Yao, Feng, Shang, Jingbo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MetaIE: Distilling a Meta Model from LLM for All Kinds of Information Extraction Tasks
by: Peng, Letian, et al.
Published: (2024)
by: Peng, Letian, et al.
Published: (2024)
Incubating Text Classifiers Following User Instruction with Nothing but LLM
by: Peng, Letian, et al.
Published: (2024)
by: Peng, Letian, et al.
Published: (2024)
Codifying Character Logic in Role-Playing
by: Peng, Letian, et al.
Published: (2025)
by: Peng, Letian, et al.
Published: (2025)
Quantifying and Optimizing Global Faithfulness in Persona-driven Role-playing
by: Peng, Letian, et al.
Published: (2024)
by: Peng, Letian, et al.
Published: (2024)
The Price of Format: Diversity Collapse in LLMs
by: Yun, Longfei, et al.
Published: (2025)
by: Yun, Longfei, et al.
Published: (2025)
Correlation and Navigation in the Vocabulary Key Representation Space of Language Models
by: Peng, Letian, et al.
Published: (2024)
by: Peng, Letian, et al.
Published: (2024)
UltraGen: Extremely Fine-grained Controllable Generation via Attribute Reconstruction and Global Preference Optimization
by: Yun, Longfei, et al.
Published: (2025)
by: Yun, Longfei, et al.
Published: (2025)
Controllable Data Augmentation for Few-Shot Text Mining with Chain-of-Thought Attribute Manipulation
by: Peng, Letian, et al.
Published: (2023)
by: Peng, Letian, et al.
Published: (2023)
Answer is All You Need: Instruction-following Text Embedding via Answering the Question
by: Peng, Letian, et al.
Published: (2024)
by: Peng, Letian, et al.
Published: (2024)
Codified Finite-state Machines for Role-playing
by: Peng, Letian, et al.
Published: (2026)
by: Peng, Letian, et al.
Published: (2026)
Linear Correlation in LM's Compositional Generalization and Hallucination
by: Peng, Letian, et al.
Published: (2025)
by: Peng, Letian, et al.
Published: (2025)
Text Grafting: Near-Distribution Weak Supervision for Minority Classes in Text Classification
by: Peng, Letian, et al.
Published: (2024)
by: Peng, Letian, et al.
Published: (2024)
Deriving Character Logic from Storyline as Codified Decision Trees
by: Peng, Letian, et al.
Published: (2026)
by: Peng, Letian, et al.
Published: (2026)
Codified Foreshadowing-Payoff Text Generation
by: Yun, Longfei, et al.
Published: (2026)
by: Yun, Longfei, et al.
Published: (2026)
Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step
by: Zhong, Li, et al.
Published: (2024)
by: Zhong, Li, et al.
Published: (2024)
DOCMASTER: A Unified Platform for Annotation, Training, & Inference in Document Question-Answering
by: Nguyen, Alex, et al.
Published: (2024)
by: Nguyen, Alex, et al.
Published: (2024)
Towards Few-shot Entity Recognition in Document Images: A Graph Neural Network Approach Robust to Image Manipulation
by: Krishnan, Prashant, et al.
Published: (2023)
by: Krishnan, Prashant, et al.
Published: (2023)
Evaluating the Smooth Control of Attribute Intensity in Text Generation with LLMs
by: Zhou, Shang, et al.
Published: (2024)
by: Zhou, Shang, et al.
Published: (2024)
Bidirectional LMs are Better Knowledge Memorizers? A Benchmark for Real-world Knowledge Injection
by: Zhang, Yuwei, et al.
Published: (2025)
by: Zhang, Yuwei, et al.
Published: (2025)
BOOKMARKS: Efficient Active Storyline Memory for Role-playing
by: Peng, Letian, et al.
Published: (2026)
by: Peng, Letian, et al.
Published: (2026)
Attention Reveals More Than Tokens: Training-Free Long-Context Reasoning with Attention-guided Retrieval
by: Zhang, Yuwei, et al.
Published: (2025)
by: Zhang, Yuwei, et al.
Published: (2025)
Watermarks for Language Models via Probabilistic Automata
by: Wang, Yangkun, et al.
Published: (2025)
by: Wang, Yangkun, et al.
Published: (2025)
Corrupted by Reasoning: Reasoning Language Models Become Free-Riders in Public Goods Games
by: Piedrahita, David Guzman, et al.
Published: (2025)
by: Piedrahita, David Guzman, et al.
Published: (2025)
OfficeBench: Benchmarking Language Agents across Multiple Applications for Office Automation
by: Wang, Zilong, et al.
Published: (2024)
by: Wang, Zilong, et al.
Published: (2024)
Training Language Models to Generate Quality Code with Program Analysis Feedback
by: Yao, Feng, et al.
Published: (2025)
by: Yao, Feng, et al.
Published: (2025)
Simulating Organized Group Behavior: New Framework, Benchmark, and Analysis
by: Zou, Xinkai, et al.
Published: (2026)
by: Zou, Xinkai, et al.
Published: (2026)
Finish First, Perfect Later: Test-Time Token-Level Cross-Validation for Diffusion Large Language Models
by: Tian, Runchu, et al.
Published: (2025)
by: Tian, Runchu, et al.
Published: (2025)
VeriLocc: End-to-End Cross-Architecture Register Allocation via LLM
by: Jin, Lesheng, et al.
Published: (2025)
by: Jin, Lesheng, et al.
Published: (2025)
Order Matters: Rethinking Prompt Construction in In-Context Learning
by: Li, Warren, et al.
Published: (2025)
by: Li, Warren, et al.
Published: (2025)
Entangled Relations: Leveraging NLI and Meta-analysis to Enhance Biomedical Relation Extraction
by: Hogan, William, et al.
Published: (2024)
by: Hogan, William, et al.
Published: (2024)
Smaller Language Models are capable of selecting Instruction-Tuning Training Data for Larger Language Models
by: Mekala, Dheeraj, et al.
Published: (2024)
by: Mekala, Dheeraj, et al.
Published: (2024)
IE as Cache: Information Extraction Enhanced Agentic Reasoning
by: Lv, Hang, et al.
Published: (2026)
by: Lv, Hang, et al.
Published: (2026)
Data Contamination Can Cross Language Barriers
by: Yao, Feng, et al.
Published: (2024)
by: Yao, Feng, et al.
Published: (2024)
Memorize or Generalize? Evaluating LLM Code Generation with Code Rewriting
by: Zhang, Lizhe, et al.
Published: (2025)
by: Zhang, Lizhe, et al.
Published: (2025)
Hybrid OCR-LLM Framework for Enterprise-Scale Document Information Extraction Under Copy-heavy Task
by: Wang, Zilong, et al.
Published: (2025)
by: Wang, Zilong, et al.
Published: (2025)
Next-Token Prediction Task Assumes Optimal Data Ordering for LLM Training in Proof Generation
by: An, Chenyang, et al.
Published: (2024)
by: An, Chenyang, et al.
Published: (2024)
READ: Improving Relation Extraction from an ADversarial Perspective
by: Li, Dawei, et al.
Published: (2024)
by: Li, Dawei, et al.
Published: (2024)
A Tale of LLMs and Induced Small Proxies: Scalable Agents for Knowledge Mining
by: Zhang, Sipeng, et al.
Published: (2025)
by: Zhang, Sipeng, et al.
Published: (2025)
Marco-LLM: Bridging Languages via Massive Multilingual Training for Cross-Lingual Enhancement
by: Ming, Lingfeng, et al.
Published: (2024)
by: Ming, Lingfeng, et al.
Published: (2024)
Composited-Nested-Learning with Data Augmentation for Nested Named Entity Recognition
by: Liao, Xingming, et al.
Published: (2024)
by: Liao, Xingming, et al.
Published: (2024)
Similar Items
-
MetaIE: Distilling a Meta Model from LLM for All Kinds of Information Extraction Tasks
by: Peng, Letian, et al.
Published: (2024) -
Incubating Text Classifiers Following User Instruction with Nothing but LLM
by: Peng, Letian, et al.
Published: (2024) -
Codifying Character Logic in Role-Playing
by: Peng, Letian, et al.
Published: (2025) -
Quantifying and Optimizing Global Faithfulness in Persona-driven Role-playing
by: Peng, Letian, et al.
Published: (2024) -
The Price of Format: Diversity Collapse in LLMs
by: Yun, Longfei, et al.
Published: (2025)