Learning to Generate Structured Output with Schema Reinforcement Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Lu, Yaxi, Li, Haolun, Cong, Xin, Zhang, Zhong, Wu, Yesai, Lin, Yankai, Liu, Zhiyuan, Liu, Fangming, Sun, Maosong |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Proactive Agent: Shifting LLM Agents from Reactive Responses to Active Assistance
por: Lu, Yaxi, et al.
Publicado: (2024)
por: Lu, Yaxi, et al.
Publicado: (2024)
RepoAgent: An LLM-Powered Open-Source Framework for Repository-level Code Documentation Generation
por: Luo, Qinyu, et al.
Publicado: (2024)
por: Luo, Qinyu, et al.
Publicado: (2024)
Adaptive Schema-aware Event Extraction with Retrieval-Augmented Generation
por: Liang, Sheng, et al.
Publicado: (2025)
por: Liang, Sheng, et al.
Publicado: (2025)
Schema as Parameterized Tools for Universal Information Extraction
por: Liang, Sheng, et al.
Publicado: (2025)
por: Liang, Sheng, et al.
Publicado: (2025)
ConPET: Continual Parameter-Efficient Tuning for Large Language Models
por: Song, Chenyang, et al.
Publicado: (2023)
por: Song, Chenyang, et al.
Publicado: (2023)
Effective and Efficient Schema-aware Information Extraction Using On-Device Large Language Models
por: Wen, Zhihao, et al.
Publicado: (2025)
por: Wen, Zhihao, et al.
Publicado: (2025)
AgentCPM-GUI: Building Mobile-Use Agents with Reinforcement Fine-Tuning
por: Zhang, Zhong, et al.
Publicado: (2025)
por: Zhang, Zhong, et al.
Publicado: (2025)
Sparsing Law: Towards Large Language Models with Greater Activation Sparsity
por: Luo, Yuqi, et al.
Publicado: (2024)
por: Luo, Yuqi, et al.
Publicado: (2024)
Low-Resource Court Judgment Summarization for Common Law Systems
por: Liu, Shuaiqi, et al.
Publicado: (2024)
por: Liu, Shuaiqi, et al.
Publicado: (2024)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
por: Oketunji, Abiodun Finbarrs
Publicado: (2023)
por: Oketunji, Abiodun Finbarrs
Publicado: (2023)
LCFO: Long Context and Long Form Output Dataset and Benchmarking
por: Costa-jussà, Marta R., et al.
Publicado: (2024)
por: Costa-jussà, Marta R., et al.
Publicado: (2024)
MALoRA: Mixture of Asymmetric Low-Rank Adaptation for Enhanced Multi-Task Learning
por: Wang, Xujia, et al.
Publicado: (2024)
por: Wang, Xujia, et al.
Publicado: (2024)
LLM-GLOBE: A Benchmark Evaluating the Cultural Values Embedded in LLM Output
por: Karinshak, Elise, et al.
Publicado: (2024)
por: Karinshak, Elise, et al.
Publicado: (2024)
AI Can Learn Scientific Taste
por: Tong, Jingqi, et al.
Publicado: (2026)
por: Tong, Jingqi, et al.
Publicado: (2026)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
por: Ashuach, Tomer, et al.
Publicado: (2025)
por: Ashuach, Tomer, et al.
Publicado: (2025)
AsyncTLS: Efficient Generative LLM Inference with Asynchronous Two-level Sparse Attention
por: Hu, Yuxuan, et al.
Publicado: (2026)
por: Hu, Yuxuan, et al.
Publicado: (2026)
Reinforcement Learning for Latent-Space Thinking in LLMs
por: Özeren, Enes, et al.
Publicado: (2025)
por: Özeren, Enes, et al.
Publicado: (2025)
ICG: Improving Cover Image Generation via MLLM-based Prompting and Personalized Preference Alignment
por: Bian, Zhipeng, et al.
Publicado: (2026)
por: Bian, Zhipeng, et al.
Publicado: (2026)
Evaluating an evidence-guided reinforcement learning framework in aligning light-parameter large language models with decision-making cognition in psychiatric clinical reasoning
por: Lin, Xinxin, et al.
Publicado: (2026)
por: Lin, Xinxin, et al.
Publicado: (2026)
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models
por: Liu, Zhongxin, et al.
Publicado: (2025)
por: Liu, Zhongxin, et al.
Publicado: (2025)
Surprise Calibration for Better In-Context Learning
por: Tan, Zhihang, et al.
Publicado: (2025)
por: Tan, Zhihang, et al.
Publicado: (2025)
Learning Translations via Matrix Completion
por: Wijaya, Derry, et al.
Publicado: (2024)
por: Wijaya, Derry, et al.
Publicado: (2024)
Policy-driven Knowledge Selection and Response Generation for Document-grounded Dialogue
por: Ma, Longxuan, et al.
Publicado: (2024)
por: Ma, Longxuan, et al.
Publicado: (2024)
Active Few-Shot Learning for Text Classification
por: Ahmadnia, Saeed, et al.
Publicado: (2025)
por: Ahmadnia, Saeed, et al.
Publicado: (2025)
ARAGOG: Advanced RAG Output Grading
por: Eibich, Matouš, et al.
Publicado: (2024)
por: Eibich, Matouš, et al.
Publicado: (2024)
Vision-Language Reasoning for Geolocalization: A Reinforcement Learning Approach
por: Wu, Biao, et al.
Publicado: (2026)
por: Wu, Biao, et al.
Publicado: (2026)
I run as fast as a rabbit, can you? A Multilingual Simile Dialogue Dataset
por: Ma, Longxuan, et al.
Publicado: (2023)
por: Ma, Longxuan, et al.
Publicado: (2023)
Learning and communication pressures in neural networks: Lessons from emergent communication
por: Galke, Lukas, et al.
Publicado: (2024)
por: Galke, Lukas, et al.
Publicado: (2024)
Combining Denoising Autoencoders with Contrastive Learning to fine-tune Transformer Models
por: Lopez-Avila, Alejo, et al.
Publicado: (2024)
por: Lopez-Avila, Alejo, et al.
Publicado: (2024)
WSM: Decay-Free Learning Rate Schedule via Checkpoint Merging for LLM Pre-training
por: Tian, Changxin, et al.
Publicado: (2025)
por: Tian, Changxin, et al.
Publicado: (2025)
ScoreRAG: A Retrieval-Augmented Generation Framework with Consistency-Relevance Scoring and Structured Summarization for News Generation
por: Lin, Pei-Yun, et al.
Publicado: (2025)
por: Lin, Pei-Yun, et al.
Publicado: (2025)
Detection of ChatGPT Fake Science with the xFakeSci Learning Algorithm
por: Hamed, Ahmed Abdeen, et al.
Publicado: (2023)
por: Hamed, Ahmed Abdeen, et al.
Publicado: (2023)
Beyond Prefixes: Graph-as-Memory Cross-Attention for Knowledge Graph Completion with Large Language Models
por: Liu, Ruitong, et al.
Publicado: (2025)
por: Liu, Ruitong, et al.
Publicado: (2025)
ProSparse: Introducing and Enhancing Intrinsic Activation Sparsity within Large Language Models
por: Song, Chenyang, et al.
Publicado: (2024)
por: Song, Chenyang, et al.
Publicado: (2024)
Efficient Few-shot Learning for Multi-label Classification of Scientific Documents with Many Classes
por: Schopf, Tim, et al.
Publicado: (2024)
por: Schopf, Tim, et al.
Publicado: (2024)
MIX : a Multi-task Learning Approach to Solve Open-Domain Question Answering
por: Chaybouti, Sofian, et al.
Publicado: (2020)
por: Chaybouti, Sofian, et al.
Publicado: (2020)
Lacuna Language Learning: Leveraging RNNs for Ranked Text Completion in Digitized Coptic Manuscripts
por: Levine, Lauren, et al.
Publicado: (2024)
por: Levine, Lauren, et al.
Publicado: (2024)
Enhancing Emotion Prediction in News Headlines: Insights from ChatGPT and Seq2Seq Models for Free-Text Generation
por: Gao, Ge, et al.
Publicado: (2024)
por: Gao, Ge, et al.
Publicado: (2024)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
por: Saji, Alan, et al.
Publicado: (2025)
por: Saji, Alan, et al.
Publicado: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
por: Peters, Sydney, et al.
Publicado: (2025)
por: Peters, Sydney, et al.
Publicado: (2025)
Ejemplares similares
-
Proactive Agent: Shifting LLM Agents from Reactive Responses to Active Assistance
por: Lu, Yaxi, et al.
Publicado: (2024) -
RepoAgent: An LLM-Powered Open-Source Framework for Repository-level Code Documentation Generation
por: Luo, Qinyu, et al.
Publicado: (2024) -
Adaptive Schema-aware Event Extraction with Retrieval-Augmented Generation
por: Liang, Sheng, et al.
Publicado: (2025) -
Schema as Parameterized Tools for Universal Information Extraction
por: Liang, Sheng, et al.
Publicado: (2025) -
ConPET: Continual Parameter-Efficient Tuning for Large Language Models
por: Song, Chenyang, et al.
Publicado: (2023)