On the Runway Cascade of Transformers for Language Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Hunjae, Clark, Corey |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Transformer Modeling for Both Scalability and Performance in Multivariate Time Series
by: Lee, Hunjae, et al.
Published: (2025)
by: Lee, Hunjae, et al.
Published: (2025)
Dynamic Relational Priming Improves Transformer in Multivariate Time Series
by: Lee, Hunjae, et al.
Published: (2025)
by: Lee, Hunjae, et al.
Published: (2025)
Understanding the Failure Modes of Transformers through the Lens of Graph Neural Networks
by: Lee, Hunjae
Published: (2025)
by: Lee, Hunjae
Published: (2025)
Blind Evaluation Framework for Fully Homomorphic Encryption and Privacy-Preserving Machine Learning
by: Lee, Hunjae "Timothy", et al.
Published: (2023)
by: Lee, Hunjae "Timothy", et al.
Published: (2023)
Synthetic Data for Robust Runway Detection
by: Chigot, Estelle, et al.
Published: (2025)
by: Chigot, Estelle, et al.
Published: (2025)
Runway vs. Taxiway: Challenges in Automated Line Identification and Notation Approaches
by: Ganeriwala, Parth, et al.
Published: (2025)
by: Ganeriwala, Parth, et al.
Published: (2025)
Robustness Assessment of a Runway Object Classifier for Safe Aircraft Taxiing
by: Elboher, Yizhak, et al.
Published: (2024)
by: Elboher, Yizhak, et al.
Published: (2024)
Cascade-Aware Training of Language Models
by: Wang, Congchao, et al.
Published: (2024)
by: Wang, Congchao, et al.
Published: (2024)
Automatic Pruning of Fine-tuning Datasets for Transformer-based Language Models
by: Tayaranian, Mohammadreza, et al.
Published: (2024)
by: Tayaranian, Mohammadreza, et al.
Published: (2024)
Cascading Adversarial Bias from Injection to Distillation in Language Models
by: Chaudhari, Harsh, et al.
Published: (2025)
by: Chaudhari, Harsh, et al.
Published: (2025)
Nemotron-Cascade: Scaling Cascaded Reinforcement Learning for General-Purpose Reasoning Models
by: Wang, Boxin, et al.
Published: (2025)
by: Wang, Boxin, et al.
Published: (2025)
SIG: A Synthetic Identity Generation Pipeline for Generating Evaluation Datasets for Face Recognition
by: Nzalasse, Kassi, et al.
Published: (2024)
by: Nzalasse, Kassi, et al.
Published: (2024)
An EM Gradient Algorithm for Mixture Models with Components Derived from the Manly Transformation
by: Clark, Katharine M., et al.
Published: (2024)
by: Clark, Katharine M., et al.
Published: (2024)
On Memory: A comparison of memory mechanisms in world models
by: Laird, Eli J., et al.
Published: (2025)
by: Laird, Eli J., et al.
Published: (2025)
C3PO: Optimized Large Language Model Cascades with Probabilistic Cost Constraints for Reasoning
by: Valkanas, Antonios, et al.
Published: (2025)
by: Valkanas, Antonios, et al.
Published: (2025)
Language Model Cascades: Token-level uncertainty and beyond
by: Gupta, Neha, et al.
Published: (2024)
by: Gupta, Neha, et al.
Published: (2024)
Bi-directional Model Cascading with Proxy Confidence
by: Warren, David, et al.
Published: (2025)
by: Warren, David, et al.
Published: (2025)
Variational Neurons in Transformers for Language Modeling
by: Ruffenach, Yves
Published: (2026)
by: Ruffenach, Yves
Published: (2026)
SAP: Syntactic Attention Pruning for Transformer-based Language Models
by: Lee, Tzu-Yun, et al.
Published: (2025)
by: Lee, Tzu-Yun, et al.
Published: (2025)
GeoTransolver: Learning Physics on Irregular Domains Using Multi-scale Geometry Aware Physics Attention Transformer
by: Adams, Corey, et al.
Published: (2025)
by: Adams, Corey, et al.
Published: (2025)
Gatekeeper: Improving Model Cascades Through Confidence Tuning
by: Rabanser, Stephan, et al.
Published: (2025)
by: Rabanser, Stephan, et al.
Published: (2025)
CascadeServe: Unlocking Model Cascades for Inference Serving
by: Kossmann, Ferdi, et al.
Published: (2024)
by: Kossmann, Ferdi, et al.
Published: (2024)
Linking In-context Learning in Transformers to Human Episodic Memory
by: Ji-An, Li, et al.
Published: (2024)
by: Ji-An, Li, et al.
Published: (2024)
Cascaded Diffusion Models for Neural Motion Planning
by: Sharma, Mohit, et al.
Published: (2025)
by: Sharma, Mohit, et al.
Published: (2025)
Cascading Robustness Verification: Toward Efficient Model-Agnostic Certification
by: Maleki, Mohammadreza, et al.
Published: (2026)
by: Maleki, Mohammadreza, et al.
Published: (2026)
Cascading Reinforcement Learning
by: Du, Yihan, et al.
Published: (2024)
by: Du, Yihan, et al.
Published: (2024)
Uncertainty-Aware Transformers: Conformal Prediction for Language Models
by: Vellore, Abhiram, et al.
Published: (2026)
by: Vellore, Abhiram, et al.
Published: (2026)
Evolutionary Large Language Model for Automated Feature Transformation
by: Gong, Nanxu, et al.
Published: (2024)
by: Gong, Nanxu, et al.
Published: (2024)
A Kolmogorov-Arnold Neural Model for Cascading Extremes
by: de Carvalho, Miguel, et al.
Published: (2025)
by: de Carvalho, Miguel, et al.
Published: (2025)
Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning
by: Yue, Murong, et al.
Published: (2023)
by: Yue, Murong, et al.
Published: (2023)
Model Cascading for Code: A Cascaded Black-Box Multi-Model Framework for Cost-Efficient Code Completion with Self-Testing
by: Chen, Boyuan, et al.
Published: (2024)
by: Chen, Boyuan, et al.
Published: (2024)
Exploring Quantization for Efficient Pre-Training of Transformer Language Models
by: Chitsaz, Kamran, et al.
Published: (2024)
by: Chitsaz, Kamran, et al.
Published: (2024)
AffineQuant: Affine Transformation Quantization for Large Language Models
by: Ma, Yuexiao, et al.
Published: (2024)
by: Ma, Yuexiao, et al.
Published: (2024)
Anatomical Heterogeneity in Transformer Language Models
by: Wietrzykowski, Tomasz
Published: (2026)
by: Wietrzykowski, Tomasz
Published: (2026)
Cascaded Cross-Modal Transformer for Audio-Textual Classification
by: Ristea, Nicolae-Catalin, et al.
Published: (2024)
by: Ristea, Nicolae-Catalin, et al.
Published: (2024)
Modeling Matches as Language: A Generative Transformer Approach for Counterfactual Player Valuation in Football
by: Hong, Miru, et al.
Published: (2026)
by: Hong, Miru, et al.
Published: (2026)
Calibrate-Then-Delegate: Safety Monitoring with Risk and Budget Guarantees via Model Cascades
by: Pona, Edoardo, et al.
Published: (2026)
by: Pona, Edoardo, et al.
Published: (2026)
Transformation-Augmented GRPO for Enhancing Exploration in Reasoning of Large Language Models
by: Le, Khiem, et al.
Published: (2026)
by: Le, Khiem, et al.
Published: (2026)
A Free Probabilistic Framework for Analyzing the Transformer-based Language Models
by: Das, Swagatam
Published: (2025)
by: Das, Swagatam
Published: (2025)
Cascading Bandits Robust to Adversarial Corruptions
by: Xie, Jize, et al.
Published: (2025)
by: Xie, Jize, et al.
Published: (2025)
Similar Items
-
Transformer Modeling for Both Scalability and Performance in Multivariate Time Series
by: Lee, Hunjae, et al.
Published: (2025) -
Dynamic Relational Priming Improves Transformer in Multivariate Time Series
by: Lee, Hunjae, et al.
Published: (2025) -
Understanding the Failure Modes of Transformers through the Lens of Graph Neural Networks
by: Lee, Hunjae
Published: (2025) -
Blind Evaluation Framework for Fully Homomorphic Encryption and Privacy-Preserving Machine Learning
by: Lee, Hunjae "Timothy", et al.
Published: (2023) -
Synthetic Data for Robust Runway Detection
by: Chigot, Estelle, et al.
Published: (2025)