UniCreative: Unifying Long-form Logic and Short-form Sparkle via Reference-Free Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Xiaolong, Zhu, Zerun, Niu, Simin, Zhang, Xingyu, Yu, Peiying, Xiao, Changxuan, Li, Yuchen, Yang, Jicheng, Zhao, Zhejun, Meng, Chong, Xia, Long, Shi, Daiting |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Igniting Creative Writing in Small Language Models: LLM-as-a-Judge versus Multi-Agent Refined Rewards
by: Wei, Xiaolong, et al.
Published: (2025)
by: Wei, Xiaolong, et al.
Published: (2025)
TURA: Tool-Augmented Unified Retrieval Agent for AI Search
by: Zhao, Zhejun, et al.
Published: (2025)
by: Zhao, Zhejun, et al.
Published: (2025)
Sparkling bubbles in chiral active fluids
by: Petrini, Alessandro, et al.
Published: (2026)
by: Petrini, Alessandro, et al.
Published: (2026)
Beyond ReAct: A Planner-Centric Framework for Complex Tool-Augmented LLM Reasoning
by: Wei, Xiaolong, et al.
Published: (2025)
by: Wei, Xiaolong, et al.
Published: (2025)
LFQA-E: Carefully Benchmarking Long-form QA Evaluation
by: Fan, Yuchen, et al.
Published: (2024)
by: Fan, Yuchen, et al.
Published: (2024)
Progressive Video Condensation with MLLM Agent for Long-form Video Understanding
by: Yin, Yufei, et al.
Published: (2026)
by: Yin, Yufei, et al.
Published: (2026)
Long-form RewardBench: Evaluating Reward Models for Long-form Generation
by: Huang, Hui, et al.
Published: (2026)
by: Huang, Hui, et al.
Published: (2026)
LitCab: Lightweight Language Model Calibration over Short- and Long-form Responses
by: Liu, Xin, et al.
Published: (2023)
by: Liu, Xin, et al.
Published: (2023)
Sparkling saddle loops of vector fields on surfaces
by: Shilin, Ivan
Published: (2019)
by: Shilin, Ivan
Published: (2019)
OmniWeaving: Towards Unified Video Generation with Free-form Composition and Reasoning
by: Pan, Kaihang, et al.
Published: (2026)
by: Pan, Kaihang, et al.
Published: (2026)
Short-form Text Rewriting with Phi Silica
by: Tadimeti, Divya, et al.
Published: (2026)
by: Tadimeti, Divya, et al.
Published: (2026)
Meta-Chunking: Learning Text Segmentation and Semantic Completion via Logical Perception
by: Zhao, Jihao, et al.
Published: (2024)
by: Zhao, Jihao, et al.
Published: (2024)
UniCustom: Unified Visual Conditioning for Multi-Reference Image Generation
by: Xu, Yiyan, et al.
Published: (2026)
by: Xu, Yiyan, et al.
Published: (2026)
UniMLVG: Unified Framework for Multi-view Long Video Generation with Comprehensive Control Capabilities for Autonomous Driving
by: Chen, Rui, et al.
Published: (2024)
by: Chen, Rui, et al.
Published: (2024)
An Impulse-formed Navier-Stokes Solver based on Long-range Particle Flow Maps
by: Li, Zhiqi, et al.
Published: (2026)
by: Li, Zhiqi, et al.
Published: (2026)
EVA-Score: Evaluating Abstractive Long-form Summarization on Informativeness through Extraction and Validation
by: Fan, Yuchen, et al.
Published: (2024)
by: Fan, Yuchen, et al.
Published: (2024)
A Baseline $T\log^2 T$ Upper Bound for KL-Regularized Prime--Zero Optimal Transport
by: Yang, Zhejun
Published: (2025)
by: Yang, Zhejun
Published: (2025)
ULU: A Unified Activation Function
by: Huo, Simin
Published: (2025)
by: Huo, Simin
Published: (2025)
UniSymNet: A Unified Symbolic Network Guided by Transformer
by: Li, Xinxin, et al.
Published: (2025)
by: Li, Xinxin, et al.
Published: (2025)
Non-contractible closed geodesics on compact Finsler space forms without self-intersections
by: Wang, Yuchen
Published: (2024)
by: Wang, Yuchen
Published: (2024)
UniCtrl: Improving the Spatiotemporal Consistency of Text-to-Video Diffusion Models via Training-Free Unified Attention Control
by: Xia, Tian, et al.
Published: (2024)
by: Xia, Tian, et al.
Published: (2024)
Data-Free Layer-Adaptive Merging via Fisher Information for Long-to-Short Reasoning LLMs
by: Xia, Tian
Published: (2026)
by: Xia, Tian
Published: (2026)
Decoding the Flow: CauseMotion for Emotional Causality Analysis in Long-form Conversations
by: Zhang, Yuxuan, et al.
Published: (2025)
by: Zhang, Yuxuan, et al.
Published: (2025)
UniPixel: Unified Object Referring and Segmentation for Pixel-Level Visual Reasoning
by: Liu, Ye, et al.
Published: (2025)
by: Liu, Ye, et al.
Published: (2025)
Primes of the form $ax+by$ in certain intervals with small solutions
by: Ding, Yuchen, et al.
Published: (2025)
by: Ding, Yuchen, et al.
Published: (2025)
Long-form evaluation of model editing
by: Rosati, Domenic, et al.
Published: (2024)
by: Rosati, Domenic, et al.
Published: (2024)
Uni-ISP: Toward Unifying the Learning of ISPs from Multiple Mobile Cameras
by: Li, Lingen, et al.
Published: (2024)
by: Li, Lingen, et al.
Published: (2024)
Writing-RL: Advancing Long-form Writing via Adaptive Curriculum Reinforcement Learning
by: Lei, Xuanyu, et al.
Published: (2025)
by: Lei, Xuanyu, et al.
Published: (2025)
OpenReward: Learning to Reward Long-form Agentic Tasks via Reinforcement Learning
by: Hu, Ziyou, et al.
Published: (2025)
by: Hu, Ziyou, et al.
Published: (2025)
ACE-RL: Adaptive Constraint-Enhanced Reward for Long-form Generation Reinforcement Learning
by: Chen, Jianghao, et al.
Published: (2025)
by: Chen, Jianghao, et al.
Published: (2025)
SimulRAG: Simulator-based RAG for Grounding LLMs in Long-form Scientific QA
by: Xu, Haozhou, et al.
Published: (2025)
by: Xu, Haozhou, et al.
Published: (2025)
Analysis of Polysulfides in Aged Sparkling Wines From Different Vintages
by: Susanne Dekker, et al.
Published: (2024)
by: Susanne Dekker, et al.
Published: (2024)
Short form of the Spanish adaptation of the State-Trait Anxiety Inventory
by: Buela-Casal, Gualberto, et al.
Published: (2017)
by: Buela-Casal, Gualberto, et al.
Published: (2017)
USV: Towards Understanding the User-generated Short-form Videos
by: Cheng, Haoyue, et al.
Published: (2026)
by: Cheng, Haoyue, et al.
Published: (2026)
KVQ: Kwai Video Quality Assessment for Short-form Videos
by: Lu, Yiting, et al.
Published: (2024)
by: Lu, Yiting, et al.
Published: (2024)
SILC: Lookahead Caching for Short-form Video Delivery Systems
by: Masood, Maleeha, et al.
Published: (2026)
by: Masood, Maleeha, et al.
Published: (2026)
FGSVQA: Frequency-Guided Short-form Video Quality Assessment
by: Wang, Xinyi, et al.
Published: (2026)
by: Wang, Xinyi, et al.
Published: (2026)
Commanding Humanoid by Free-form Language: A Large Language Action Model with Unified Motion Vocabulary
by: Liu, Zhirui, et al.
Published: (2025)
by: Liu, Zhirui, et al.
Published: (2025)
UniFormer: Unified and Efficient Transformer for Reasoning Across General and Custom Computing
by: Ran, Zhuoheng, et al.
Published: (2025)
by: Ran, Zhuoheng, et al.
Published: (2025)
Safe Reinforcement Learning with Free-form Natural Language Constraints and Pre-Trained Language Models
by: Lou, Xingzhou, et al.
Published: (2024)
by: Lou, Xingzhou, et al.
Published: (2024)
Similar Items
-
Igniting Creative Writing in Small Language Models: LLM-as-a-Judge versus Multi-Agent Refined Rewards
by: Wei, Xiaolong, et al.
Published: (2025) -
TURA: Tool-Augmented Unified Retrieval Agent for AI Search
by: Zhao, Zhejun, et al.
Published: (2025) -
Sparkling bubbles in chiral active fluids
by: Petrini, Alessandro, et al.
Published: (2026) -
Beyond ReAct: A Planner-Centric Framework for Complex Tool-Augmented LLM Reasoning
by: Wei, Xiaolong, et al.
Published: (2025) -
LFQA-E: Carefully Benchmarking Long-form QA Evaluation
by: Fan, Yuchen, et al.
Published: (2024)