Beyond Single-Task: Robust Multi-Task Length Generalization for LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hu, Yi, Kang, Shijia, Yang, Haotong, Xu, Haotian, Zhang, Muhan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Number Cookbook: Number Understanding of Language Models and How to Improve It
von: Yang, Haotong, et al.
Veröffentlicht: (2024)
von: Yang, Haotong, et al.
Veröffentlicht: (2024)
Parrot Mind: Towards Explaining the Complex Task Reasoning of Pretrained Large Language Models with Template-Content Structure
von: Yang, Haotong, et al.
Veröffentlicht: (2023)
von: Yang, Haotong, et al.
Veröffentlicht: (2023)
Case-Based or Rule-Based: How Do Transformers Do the Math?
von: Hu, Yi, et al.
Veröffentlicht: (2024)
von: Hu, Yi, et al.
Veröffentlicht: (2024)
Proof-RM: A Scalable and Generalizable Reward Model for Math Proof
von: Yang, Haotong, et al.
Veröffentlicht: (2026)
von: Yang, Haotong, et al.
Veröffentlicht: (2026)
LiteToken: Removing Intermediate Merge Residues From BPE Tokenizers
von: Sun, Yike, et al.
Veröffentlicht: (2026)
von: Sun, Yike, et al.
Veröffentlicht: (2026)
Dynamic Prompt Fusion for Multi-Task and Cross-Domain Adaptation in LLMs
von: Hu, Xin, et al.
Veröffentlicht: (2025)
von: Hu, Xin, et al.
Veröffentlicht: (2025)
LIFT: A Novel Framework for Enhancing Long-Context Understanding of LLMs via Long Input Fine-Tuning
von: Mao, Yansheng, et al.
Veröffentlicht: (2025)
von: Mao, Yansheng, et al.
Veröffentlicht: (2025)
Large Language Models Still Face Challenges in Multi-Hop Reasoning with External Knowledge
von: Zhang, Haotong
Veröffentlicht: (2024)
von: Zhang, Haotong
Veröffentlicht: (2024)
GL-Fusion: Rethinking the Combination of Graph Neural Network and Large Language model
von: Yang, Haotong, et al.
Veröffentlicht: (2024)
von: Yang, Haotong, et al.
Veröffentlicht: (2024)
What Affects the Effective Depth of Large Language Models?
von: Hu, Yi, et al.
Veröffentlicht: (2025)
von: Hu, Yi, et al.
Veröffentlicht: (2025)
Towards Understanding Multi-Task Learning (Generalization) of LLMs via Detecting and Exploring Task-Specific Neurons
von: Leng, Yongqi, et al.
Veröffentlicht: (2024)
von: Leng, Yongqi, et al.
Veröffentlicht: (2024)
The Road Less Traveled: Enhancing Exploration in LLMs via Sequential Sampling
von: Kang, Shijia, et al.
Veröffentlicht: (2025)
von: Kang, Shijia, et al.
Veröffentlicht: (2025)
Efficient Knowledge Transfer in Multi-Task Learning through Task-Adaptive Low-Rank Representation
von: Zhang, Xiao, et al.
Veröffentlicht: (2025)
von: Zhang, Xiao, et al.
Veröffentlicht: (2025)
TaskCraft: Automated Generation of Agentic Tasks
von: Shi, Dingfeng, et al.
Veröffentlicht: (2025)
von: Shi, Dingfeng, et al.
Veröffentlicht: (2025)
Streamlining the Collaborative Chain of Models into A Single Forward Pass in Generation-Based Tasks
von: Lyu, Yuanjie, et al.
Veröffentlicht: (2025)
von: Lyu, Yuanjie, et al.
Veröffentlicht: (2025)
RedStar: Does Scaling Long-CoT Data Unlock Better Slow-Reasoning Systems?
von: Xu, Haotian, et al.
Veröffentlicht: (2025)
von: Xu, Haotian, et al.
Veröffentlicht: (2025)
Multi-Task Learning with LLMs for Implicit Sentiment Analysis: Data-level and Task-level Automatic Weight Learning
von: Lai, Wenna, et al.
Veröffentlicht: (2024)
von: Lai, Wenna, et al.
Veröffentlicht: (2024)
SubTokenTest: A Practical Benchmark for Real-World Sub-token Understanding
von: Hou, Shuyang, et al.
Veröffentlicht: (2026)
von: Hou, Shuyang, et al.
Veröffentlicht: (2026)
Explicitly Encoding Structural Symmetry is Key to Length Generalization in Arithmetic Tasks
von: Sabbaghi, Mahdi, et al.
Veröffentlicht: (2024)
von: Sabbaghi, Mahdi, et al.
Veröffentlicht: (2024)
MaiNLP at SemEval-2024 Task 1: Analyzing Source Language Selection in Cross-Lingual Textual Relatedness
von: Zhou, Shijia, et al.
Veröffentlicht: (2024)
von: Zhou, Shijia, et al.
Veröffentlicht: (2024)
AutoTask: Task Aware Multi-Faceted Single Model for Multi-Task Ads Relevance
von: Guo, Shouchang, et al.
Veröffentlicht: (2024)
von: Guo, Shouchang, et al.
Veröffentlicht: (2024)
Existing LLMs Are Not Self-Consistent For Simple Tasks
von: Lin, Zhenru, et al.
Veröffentlicht: (2025)
von: Lin, Zhenru, et al.
Veröffentlicht: (2025)
Multi-Task Learning for Front-End Text Processing in TTS
von: Kang, Wonjune, et al.
Veröffentlicht: (2024)
von: Kang, Wonjune, et al.
Veröffentlicht: (2024)
Beyond English: The Impact of Prompt Translation Strategies across Languages and Tasks in Multilingual LLMs
von: Mondshine, Itai, et al.
Veröffentlicht: (2025)
von: Mondshine, Itai, et al.
Veröffentlicht: (2025)
Aligning Paralinguistic Understanding and Generation in Speech LLMs via Multi-Task Reinforcement Learning
von: Chen, Jingxiang, et al.
Veröffentlicht: (2026)
von: Chen, Jingxiang, et al.
Veröffentlicht: (2026)
Balancing the Reasoning Load: Difficulty-Differentiated Policy Optimization with Length Redistribution for Efficient and Robust Reinforcement Learning
von: Xia, Yinan, et al.
Veröffentlicht: (2026)
von: Xia, Yinan, et al.
Veröffentlicht: (2026)
CA-LoRA: Adapting Existing LoRA for Compressed LLMs to Enable Efficient Multi-Tasking on Personal Devices
von: Zhao, Weilin, et al.
Veröffentlicht: (2023)
von: Zhao, Weilin, et al.
Veröffentlicht: (2023)
CATMark: A Context-Aware Thresholding Framework for Robust Cross-Task Watermarking in Large Language Models
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
von: Zhang, Yu, et al.
Veröffentlicht: (2025)
Beyond Elicitation: Provision-based Prompt Optimization for Knowledge-Intensive Tasks
von: Xu, Yunzhe, et al.
Veröffentlicht: (2025)
von: Xu, Yunzhe, et al.
Veröffentlicht: (2025)
Predictable Emergent Abilities of LLMs: Proxy Tasks Are All You Need
von: Zhang, Bo-Wen, et al.
Veröffentlicht: (2024)
von: Zhang, Bo-Wen, et al.
Veröffentlicht: (2024)
Unveiling Over-Memorization in Finetuning LLMs for Reasoning Tasks
von: Ruan, Zhiwen, et al.
Veröffentlicht: (2025)
von: Ruan, Zhiwen, et al.
Veröffentlicht: (2025)
Finetuning LLMs for Comparative Assessment Tasks
von: Raina, Vatsal, et al.
Veröffentlicht: (2024)
von: Raina, Vatsal, et al.
Veröffentlicht: (2024)
Systematic Task Exploration with LLMs: A Study in Citation Text Generation
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2024)
von: Şahinuç, Furkan, et al.
Veröffentlicht: (2024)
From Instance Training to Instruction Learning: Task Adapters Generation from Instructions
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024)
von: Liao, Huanxuan, et al.
Veröffentlicht: (2024)
Are Long-LLMs A Necessity For Long-Context Tasks?
von: Qian, Hongjin, et al.
Veröffentlicht: (2024)
von: Qian, Hongjin, et al.
Veröffentlicht: (2024)
Dynamic Task Vector Grouping for Efficient Multi-Task Prompt Tuning
von: Zhang, Pieyi, et al.
Veröffentlicht: (2025)
von: Zhang, Pieyi, et al.
Veröffentlicht: (2025)
PsychiatryBench: A Multi-Task Benchmark for LLMs in Psychiatry
von: Fouda, Aya E., et al.
Veröffentlicht: (2025)
von: Fouda, Aya E., et al.
Veröffentlicht: (2025)
Emotion-Enhanced Multi-Task Learning with LLMs for Aspect Category Sentiment Analysis
von: Chai, Yaping, et al.
Veröffentlicht: (2025)
von: Chai, Yaping, et al.
Veröffentlicht: (2025)
On Lexical Invariance on Multisets and Graphs
von: Zhang, Muhan
Veröffentlicht: (2024)
von: Zhang, Muhan
Veröffentlicht: (2024)
Length Controlled Generation for Black-box LLMs
von: Gu, Yuxuan, et al.
Veröffentlicht: (2024)
von: Gu, Yuxuan, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Number Cookbook: Number Understanding of Language Models and How to Improve It
von: Yang, Haotong, et al.
Veröffentlicht: (2024) -
Parrot Mind: Towards Explaining the Complex Task Reasoning of Pretrained Large Language Models with Template-Content Structure
von: Yang, Haotong, et al.
Veröffentlicht: (2023) -
Case-Based or Rule-Based: How Do Transformers Do the Math?
von: Hu, Yi, et al.
Veröffentlicht: (2024) -
Proof-RM: A Scalable and Generalizable Reward Model for Math Proof
von: Yang, Haotong, et al.
Veröffentlicht: (2026) -
LiteToken: Removing Intermediate Merge Residues From BPE Tokenizers
von: Sun, Yike, et al.
Veröffentlicht: (2026)