FiRST: Finetuning Router-Selective Transformers for Input-Adaptive Latency Reduction
Fuente:
arXiv
Saved in:
| Main Authors: | Jain, Akriti, Sharma, Saransh, Mukherjee, Koyel, Pal, Soumyabrata |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Answer is just the Start: Related Insight Generation for Open-Ended Document-Grounded QA
by: Sharma, Saransh, et al.
Published: (2026)
by: Sharma, Saransh, et al.
Published: (2026)
PromptRefine: Enhancing Few-Shot Performance on Low-Resource Indic Languages with Example Selection from Related Example Banks
by: Ghosal, Soumya Suvra, et al.
Published: (2024)
by: Ghosal, Soumya Suvra, et al.
Published: (2024)
Modeling Contextual Passage Utility for Multihop Question Answering
by: Jain, Akriti, et al.
Published: (2025)
by: Jain, Akriti, et al.
Published: (2025)
Knowing What's Missing: Assessing Information Sufficiency in Question Answering
by: Jain, Akriti, et al.
Published: (2025)
by: Jain, Akriti, et al.
Published: (2025)
From Tokens to Steps: Verification-Aware Speculative Decoding for Efficient Multi-Step Reasoning
by: Purohit, Kiran, et al.
Published: (2026)
by: Purohit, Kiran, et al.
Published: (2026)
eRST: A Signaled Graph Theory of Discourse Relations and Organization
by: Zeldes, Amir, et al.
Published: (2024)
by: Zeldes, Amir, et al.
Published: (2024)
SHA256 at SemEval-2025 Task 4: Selective Amnesia -- Constrained Unlearning for Large Language Models via Knowledge Isolation
by: Agrawal, Saransh, et al.
Published: (2025)
by: Agrawal, Saransh, et al.
Published: (2025)
Doc2Chart: Intent-Driven Zero-Shot Chart Generation from Documents
by: Jain, Akriti, et al.
Published: (2025)
by: Jain, Akriti, et al.
Published: (2025)
Text Takes Over: A Study of Modality Bias in Multimodal Intent Detection
by: Mullick, Ankan, et al.
Published: (2025)
by: Mullick, Ankan, et al.
Published: (2025)
Flexi-LoRA with Input-Adaptive Ranks: Efficient Finetuning for Speech and Reasoning Tasks
by: Li, Zongqian, et al.
Published: (2026)
by: Li, Zongqian, et al.
Published: (2026)
Can we obtain significant success in RST discourse parsing by using Large Language Models?
by: Maekawa, Aru, et al.
Published: (2024)
by: Maekawa, Aru, et al.
Published: (2024)
ContextFocus: Activation Steering for Contextual Faithfulness in Large Language Models
by: Anand, Nikhil, et al.
Published: (2026)
by: Anand, Nikhil, et al.
Published: (2026)
Lightweight Domain Adaptation of a Large Language Model for Legal Assistance in the Indian Context
by: Gupta, Jatin, et al.
Published: (2025)
by: Gupta, Jatin, et al.
Published: (2025)
CP-Router: An Uncertainty-Aware Router Between LLM and LRM
by: Su, Jiayuan, et al.
Published: (2025)
by: Su, Jiayuan, et al.
Published: (2025)
Router Upcycling: Leveraging Mixture-of-Routers in Mixture-of-Experts Upcycling
by: Ran, Junfeng, et al.
Published: (2025)
by: Ran, Junfeng, et al.
Published: (2025)
Cross-Document Cross-Lingual NLI via RST-Enhanced Graph Fusion and Interpretability Prediction
by: Yuan, Mengying, et al.
Published: (2025)
by: Yuan, Mengying, et al.
Published: (2025)
IDALC: A Semi-Supervised Framework for Intent Detection and Active Learning based Correction
by: Mullick, Ankan, et al.
Published: (2025)
by: Mullick, Ankan, et al.
Published: (2025)
Building FKG.in: a Knowledge Graph for Indian Food
by: Gupta, Saransh Kumar, et al.
Published: (2024)
by: Gupta, Saransh Kumar, et al.
Published: (2024)
Decisive: Guiding User Decisions with Optimal Preference Elicitation from Unstructured Documents
by: Jain, Akriti, et al.
Published: (2026)
by: Jain, Akriti, et al.
Published: (2026)
Router-Tuning: A Simple and Effective Approach for Enabling Dynamic-Depth in Transformers
by: He, Shwai, et al.
Published: (2024)
by: He, Shwai, et al.
Published: (2024)
Retrieval-Augmented Reasoning for Chartered Accountancy
by: Gupta, Jatin, et al.
Published: (2026)
by: Gupta, Jatin, et al.
Published: (2026)
Mixture of Routers
by: Zhang, Jia-Chen, et al.
Published: (2025)
by: Zhang, Jia-Chen, et al.
Published: (2025)
Towards Optimizing the Costs of LLM Usage
by: Shekhar, Shivanshu, et al.
Published: (2024)
by: Shekhar, Shivanshu, et al.
Published: (2024)
WebRouter: Query-specific Router via Variational Information Bottleneck for Cost-sensitive Web Agent
by: Li, Tao, et al.
Published: (2025)
by: Li, Tao, et al.
Published: (2025)
AgentRouter: A Knowledge-Graph-Guided LLM Router for Collaborative Multi-Agent Question Answering
by: Zhang, Zheyuan, et al.
Published: (2025)
by: Zhang, Zheyuan, et al.
Published: (2025)
Latency Adjustable Transformer Encoder for Language Understanding
by: Kachuee, Sajjad, et al.
Published: (2022)
by: Kachuee, Sajjad, et al.
Published: (2022)
LLMs and Finetuning: Benchmarking cross-domain performance for hate speech detection
by: Nasir, Ahmad, et al.
Published: (2023)
by: Nasir, Ahmad, et al.
Published: (2023)
Efficient Reinforcement Finetuning via Adaptive Curriculum Learning
by: Shi, Taiwei, et al.
Published: (2025)
by: Shi, Taiwei, et al.
Published: (2025)
Layerwise Recurrent Router for Mixture-of-Experts
by: Qiu, Zihan, et al.
Published: (2024)
by: Qiu, Zihan, et al.
Published: (2024)
Making the Most of your Model: Methods for Finetuning and Applying Pretrained Transformers
by: Yoshida, Davis
Published: (2024)
by: Yoshida, Davis
Published: (2024)
XL-DURel: Finetuning Sentence Transformers for Ordinal Word-in-Context Classification
by: Yadav, Sachin, et al.
Published: (2025)
by: Yadav, Sachin, et al.
Published: (2025)
FrugalRAG: Less is More in RL Finetuning for Multi-Hop Question Answering
by: Java, Abhinav, et al.
Published: (2025)
by: Java, Abhinav, et al.
Published: (2025)
RST-LoRA: A Discourse-Aware Low-Rank Adaptation for Long Document Abstractive Summarization
by: Liu, Dongqi, et al.
Published: (2024)
by: Liu, Dongqi, et al.
Published: (2024)
Adaptive BPE Tokenization for Enhanced Vocabulary Adaptation in Finetuning Pretrained Language Models
by: Balde, Gunjan, et al.
Published: (2024)
by: Balde, Gunjan, et al.
Published: (2024)
Maya: An Instruction Finetuned Multilingual Multimodal Model
by: Alam, Nahid, et al.
Published: (2024)
by: Alam, Nahid, et al.
Published: (2024)
Enhancing FKG.in: automating Indian food composition analysis
by: Gupta, Saransh Kumar, et al.
Published: (2024)
by: Gupta, Saransh Kumar, et al.
Published: (2024)
LLM Judges Inconsistently Disagree Across Safety Criteria and Harm Categories
by: Vishnubhotla, Krishnapriya, et al.
Published: (2026)
by: Vishnubhotla, Krishnapriya, et al.
Published: (2026)
TripTide: A Benchmark for Adaptive Travel Planning under Disruptions
by: Karmakar, Priyanshu, et al.
Published: (2025)
by: Karmakar, Priyanshu, et al.
Published: (2025)
Part-Of-Speech Sensitivity of Routers in Mixture of Experts Models
by: Antoine, Elie, et al.
Published: (2024)
by: Antoine, Elie, et al.
Published: (2024)
Arch-Router: Aligning LLM Routing with Human Preferences
by: Tran, Co, et al.
Published: (2025)
by: Tran, Co, et al.
Published: (2025)
Similar Items
-
An Answer is just the Start: Related Insight Generation for Open-Ended Document-Grounded QA
by: Sharma, Saransh, et al.
Published: (2026) -
PromptRefine: Enhancing Few-Shot Performance on Low-Resource Indic Languages with Example Selection from Related Example Banks
by: Ghosal, Soumya Suvra, et al.
Published: (2024) -
Modeling Contextual Passage Utility for Multihop Question Answering
by: Jain, Akriti, et al.
Published: (2025) -
Knowing What's Missing: Assessing Information Sufficiency in Question Answering
by: Jain, Akriti, et al.
Published: (2025) -
From Tokens to Steps: Verification-Aware Speculative Decoding for Efficient Multi-Step Reasoning
by: Purohit, Kiran, et al.
Published: (2026)