SwitchCIT: Switching for Continual Instruction Tuning
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Xinbo, Hartman, Max, Jayaraman, Vidhata Arjun, Varshney, Lav R. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Skip-It? Theoretical Conditions for Layer Skipping in Vision-Language Models
by: Hartman, Max, et al.
Published: (2025)
by: Hartman, Max, et al.
Published: (2025)
Transformer-based Causal Language Models Perform Clustering
by: Wu, Xinbo, et al.
Published: (2024)
by: Wu, Xinbo, et al.
Published: (2024)
Hiding in Plain Sight: Detectability-Aware Antidistillation of Reasoning Models
by: Hartman, Max, et al.
Published: (2026)
by: Hartman, Max, et al.
Published: (2026)
A Meta-Learning Perspective on Transformers for Causal Language Modeling
by: Wu, Xinbo, et al.
Published: (2023)
by: Wu, Xinbo, et al.
Published: (2023)
SparseJEPA: Sparse Representation Learning of Joint Embedding Predictive Architectures
by: Hartman, Max, et al.
Published: (2025)
by: Hartman, Max, et al.
Published: (2025)
Context-Gated Associative Retrieval: From Theory to Transformers
by: Choraria, Moulik, et al.
Published: (2026)
by: Choraria, Moulik, et al.
Published: (2026)
Beyond Bilingual Transfer: Multilingual Code-Switching in Instruction Tuning
by: Asano, Shunta, et al.
Published: (2026)
by: Asano, Shunta, et al.
Published: (2026)
Hallucination Basins: A Dynamic Framework for Understanding and Controlling LLM Hallucinations
by: Cherukuri, Kalyan, et al.
Published: (2026)
by: Cherukuri, Kalyan, et al.
Published: (2026)
Efficient Model-Agnostic Multi-Group Equivariant Networks
by: Baltaji, Razan, et al.
Published: (2023)
by: Baltaji, Razan, et al.
Published: (2023)
ProSwitch: Knowledge-Guided Instruction Tuning to Switch Between Professional and Non-Professional Responses
by: Zong, Chang, et al.
Published: (2024)
by: Zong, Chang, et al.
Published: (2024)
Persona Inconstancy in Multi-Agent LLM Collaboration: Conformity, Confabulation, and Impersonation
by: Baltaji, Razan, et al.
Published: (2024)
by: Baltaji, Razan, et al.
Published: (2024)
Concealment of Intent: A Game-Theoretic Analysis
by: Wu, Xinbo, et al.
Published: (2025)
by: Wu, Xinbo, et al.
Published: (2025)
Fed-SB: A Silver Bullet for Extreme Communication Efficiency and Performance in (Private) Federated LoRA Fine-Tuning
by: Singhal, Raghav, et al.
Published: (2025)
by: Singhal, Raghav, et al.
Published: (2025)
Beyond Continuity: Challenges of Context Switching in Multi-Turn Dialogue with LLMs
by: Sinha, Aditya, et al.
Published: (2026)
by: Sinha, Aditya, et al.
Published: (2026)
When to Continue Thinking: Adaptive Thinking Mode Switching for Efficient Reasoning
by: Zhang, Xiaoyun, et al.
Published: (2025)
by: Zhang, Xiaoyun, et al.
Published: (2025)
Energy-Aware Routing to Large Reasoning Models
by: Ellis-Mohr, Austin R., et al.
Published: (2025)
by: Ellis-Mohr, Austin R., et al.
Published: (2025)
AdaSwitch: Balancing Exploration and Guidance in Knowledge Distillation via Adaptive Switching
by: Peng, Jingyu, et al.
Published: (2025)
by: Peng, Jingyu, et al.
Published: (2025)
SwitchLingua: The First Large-Scale Multilingual and Multi-Ethnic Code-Switching Dataset
by: Xie, Peng, et al.
Published: (2025)
by: Xie, Peng, et al.
Published: (2025)
ExpliCIT-QA: Explainable Code-Based Image Table Question Answering
by: Lagos, Maximiliano Hormazábal, et al.
Published: (2025)
by: Lagos, Maximiliano Hormazábal, et al.
Published: (2025)
A Theoretical Game of Attacks via Compositional Skills
by: Wu, Xinbo, et al.
Published: (2026)
by: Wu, Xinbo, et al.
Published: (2026)
Parsing the Switch: LLM-Based UD Annotation for Complex Code-Switched and Low-Resource Languages
by: Kellert, Olga, et al.
Published: (2025)
by: Kellert, Olga, et al.
Published: (2025)
Instruction Tuning With Loss Over Instructions
by: Shi, Zhengyan, et al.
Published: (2024)
by: Shi, Zhengyan, et al.
Published: (2024)
MixReasoning: Switching Modes to Think
by: Lu, Haiquan, et al.
Published: (2025)
by: Lu, Haiquan, et al.
Published: (2025)
Enhancing Multimodal Continual Instruction Tuning with BranchLoRA
by: Zhang, Duzhen, et al.
Published: (2025)
by: Zhang, Duzhen, et al.
Published: (2025)
Towards Automatic Continual Learning: A Self-Adaptive Framework for Continual Instruction Tuning
by: Lin, Peiyi, et al.
Published: (2025)
by: Lin, Peiyi, et al.
Published: (2025)
Computational Approaches to Arabic-English Code-Switching
by: Sabty, Caroline
Published: (2024)
by: Sabty, Caroline
Published: (2024)
Indic-TunedLens: Interpreting Multilingual Models in Indian Languages
by: Panchal, Mihir, et al.
Published: (2026)
by: Panchal, Mihir, et al.
Published: (2026)
Conditioning LLMs to Generate Code-Switched Text
by: Heredia, Maite, et al.
Published: (2025)
by: Heredia, Maite, et al.
Published: (2025)
Mind the Gap: Conformative Decoding to Improve Output Diversity of Instruction-Tuned Large Language Models
by: Peeperkorn, Max, et al.
Published: (2025)
by: Peeperkorn, Max, et al.
Published: (2025)
Finding the Translation Switch: Discovering and Exploiting the Task-Initiation Features in LLMs
by: Wu, Xinwei, et al.
Published: (2026)
by: Wu, Xinwei, et al.
Published: (2026)
Don't Half-listen: Capturing Key-part Information in Continual Instruction Tuning
by: He, Yongquan, et al.
Published: (2024)
by: He, Yongquan, et al.
Published: (2024)
MLLM-CTBench: A Benchmark for Continual Instruction Tuning with Reasoning Process Diagnosis
by: Guo, Haiyun, et al.
Published: (2025)
by: Guo, Haiyun, et al.
Published: (2025)
ME-Switch: A Memory-Efficient Expert Switching Framework for Large Language Models
by: Liu, Jing, et al.
Published: (2024)
by: Liu, Jing, et al.
Published: (2024)
SwitchLoRA: Switched Low-Rank Adaptation Can Learn Full-Rank Information
by: Zhou, Kaiye, et al.
Published: (2024)
by: Zhou, Kaiye, et al.
Published: (2024)
Detecting Propaganda Techniques in Code-Switched Social Media Text
by: Salman, Muhammad Umar, et al.
Published: (2023)
by: Salman, Muhammad Umar, et al.
Published: (2023)
Rethinking Table Instruction Tuning
by: Deng, Naihao, et al.
Published: (2025)
by: Deng, Naihao, et al.
Published: (2025)
Many LLMs Are More Utilitarian Than One
by: Keshmirian, Anita, et al.
Published: (2025)
by: Keshmirian, Anita, et al.
Published: (2025)
ConCSE: Unified Contrastive Learning and Augmentation for Code-Switched Embeddings
by: Jeon, Jangyeong, et al.
Published: (2024)
by: Jeon, Jangyeong, et al.
Published: (2024)
Code-Switching Red-Teaming: LLM Evaluation for Safety and Multilingual Understanding
by: Yoo, Haneul, et al.
Published: (2024)
by: Yoo, Haneul, et al.
Published: (2024)
MANTIS: Interleaved Multi-Image Instruction Tuning
by: Jiang, Dongfu, et al.
Published: (2024)
by: Jiang, Dongfu, et al.
Published: (2024)
Similar Items
-
Skip-It? Theoretical Conditions for Layer Skipping in Vision-Language Models
by: Hartman, Max, et al.
Published: (2025) -
Transformer-based Causal Language Models Perform Clustering
by: Wu, Xinbo, et al.
Published: (2024) -
Hiding in Plain Sight: Detectability-Aware Antidistillation of Reasoning Models
by: Hartman, Max, et al.
Published: (2026) -
A Meta-Learning Perspective on Transformers for Causal Language Modeling
by: Wu, Xinbo, et al.
Published: (2023) -
SparseJEPA: Sparse Representation Learning of Joint Embedding Predictive Architectures
by: Hartman, Max, et al.
Published: (2025)