Improving Sequence-to-Sequence Models for Abstractive Text Summarization Using Meta Heuristic Approaches
Fuente:
arXiv
Saved in:
| Main Authors: | Saxena, Aditya, Ranjan, Ashutosh |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SpikingSSMs: Learning Long Sequences with Sparse and Parallel Spiking State Space Models
by: Shen, Shuaijie, et al.
Published: (2024)
by: Shen, Shuaijie, et al.
Published: (2024)
Delay Embedding Theory of Neural Sequence Models
by: Ostrow, Mitchell, et al.
Published: (2024)
by: Ostrow, Mitchell, et al.
Published: (2024)
Long-Sequence LSTM Modeling for NBA Game Outcome Prediction Using a Novel Multi-Season Dataset
by: Rios, Charles, et al.
Published: (2025)
by: Rios, Charles, et al.
Published: (2025)
State-Space Constraints Can Improve the Generalisation of the Differentiable Neural Computer to Input Sequences With Unseen Length
by: Ofner, Patrick, et al.
Published: (2021)
by: Ofner, Patrick, et al.
Published: (2021)
EvoX: Meta-Evolution for Automated Discovery
by: Liu, Shu, et al.
Published: (2026)
by: Liu, Shu, et al.
Published: (2026)
Learning Long Sequences in Spiking Neural Networks
by: Stan, Matei Ioan, et al.
Published: (2023)
by: Stan, Matei Ioan, et al.
Published: (2023)
B'MOJO: Hybrid State Space Realizations of Foundation Models with Eidetic and Fading Memory
by: Zancato, Luca, et al.
Published: (2024)
by: Zancato, Luca, et al.
Published: (2024)
HubRouter: A Pluggable Sub-Quadratic Routing Primitive for Hybrid Sequence Models
by: Basu, Abhinaba
Published: (2026)
by: Basu, Abhinaba
Published: (2026)
Closed-Form Test Functions for Biophysical Sequence Optimization Algorithms
by: Stanton, Samuel, et al.
Published: (2024)
by: Stanton, Samuel, et al.
Published: (2024)
Improving Language Plasticity via Pretraining with Active Forgetting
by: Chen, Yihong, et al.
Published: (2023)
by: Chen, Yihong, et al.
Published: (2023)
Neuromorphic Spiking Neural Network Based Classification of COVID-19 Spike Sequences
by: Murad, Taslim, et al.
Published: (2024)
by: Murad, Taslim, et al.
Published: (2024)
Man-Made Heuristics Are Dead. Long Live Code Generators!
by: Dwivedula, Rohit, et al.
Published: (2025)
by: Dwivedula, Rohit, et al.
Published: (2025)
The Alpha-Alternator: Dynamic Adaptation To Varying Noise Levels In Sequences Using The Vendi Score For Improved Robustness and Performance
by: Rezaei, Mohammad Reza, et al.
Published: (2025)
by: Rezaei, Mohammad Reza, et al.
Published: (2025)
Indian Wedding System Optimization (IWSO): A Novel Socially Inspired Metaheuristic with Operational Design and Analysis
by: Saxena, Deepika, et al.
Published: (2026)
by: Saxena, Deepika, et al.
Published: (2026)
Are Large Language Models In-Context Personalized Summarizers? Get an iCOPERNICUS Test Done!
by: Patel, Divya, et al.
Published: (2024)
by: Patel, Divya, et al.
Published: (2024)
Heuristic Search as Language-Guided Program Optimization
by: Yu, Mingxin, et al.
Published: (2026)
by: Yu, Mingxin, et al.
Published: (2026)
Large Language Models for Tuning Evolution Strategies
by: Kramer, Oliver
Published: (2024)
by: Kramer, Oliver
Published: (2024)
Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers
by: Kadlčík, Marek, et al.
Published: (2025)
by: Kadlčík, Marek, et al.
Published: (2025)
Sorbet: A Neuromorphic Hardware-Compatible Transformer-Based Spiking Language Model
by: Tang, Kaiwen, et al.
Published: (2024)
by: Tang, Kaiwen, et al.
Published: (2024)
SpikeGPT: Generative Pre-trained Language Model with Spiking Neural Networks
by: Zhu, Rui-Jie, et al.
Published: (2023)
by: Zhu, Rui-Jie, et al.
Published: (2023)
Cancer-inspired Genomics Mapper Model for the Generation of Synthetic DNA Sequences with Desired Genomics Signatures
by: Lazebnik, Teddy, et al.
Published: (2023)
by: Lazebnik, Teddy, et al.
Published: (2023)
Pruner-Zero: Evolving Symbolic Pruning Metric from scratch for Large Language Models
by: Dong, Peijie, et al.
Published: (2024)
by: Dong, Peijie, et al.
Published: (2024)
Genetic Instruct: Scaling up Synthetic Generation of Coding Instructions for Large Language Models
by: Majumdar, Somshubra, et al.
Published: (2024)
by: Majumdar, Somshubra, et al.
Published: (2024)
Large Language Models Suffer From Their Own Output: An Analysis of the Self-Consuming Training Loop
by: Briesch, Martin, et al.
Published: (2023)
by: Briesch, Martin, et al.
Published: (2023)
SpikeLM: Towards General Spike-Driven Language Modeling via Elastic Bi-Spiking Mechanisms
by: Xing, Xingrun, et al.
Published: (2024)
by: Xing, Xingrun, et al.
Published: (2024)
SpikeLLM: Scaling up Spiking Neural Network to Large Language Models via Saliency-based Spiking
by: Xing, Xingrun, et al.
Published: (2024)
by: Xing, Xingrun, et al.
Published: (2024)
Jailbreaking LLMs' Safeguard with Universal Magic Words for Text Embedding Models
by: Liang, Haoyu, et al.
Published: (2025)
by: Liang, Haoyu, et al.
Published: (2025)
AP-BMM: Approximating Capability-Cost Pareto Sets of LLMs via Asynchronous Prior-Guided Bayesian Model Merging
by: Chen, Kesheng, et al.
Published: (2025)
by: Chen, Kesheng, et al.
Published: (2025)
Optimizing PM2.5 Forecasting Accuracy with Hybrid Meta-Heuristic and Machine Learning Models
by: Ghafariasl, Parviz, et al.
Published: (2024)
by: Ghafariasl, Parviz, et al.
Published: (2024)
Statistical Properties of the King Wen Sequence: An Anti-Habituation Structure That Does Not Improve Neural Network Training
by: Chan, Augustin
Published: (2026)
by: Chan, Augustin
Published: (2026)
An Imbalanced Dataset with Multiple Feature Representations for Studying Quality Control of Next-Generation Sequencing
by: Röchner, Philipp, et al.
Published: (2026)
by: Röchner, Philipp, et al.
Published: (2026)
From Heuristic Selection to Automated Algorithm Design: LLMs Benefit from Strong Priors
by: Huang, Qi, et al.
Published: (2026)
by: Huang, Qi, et al.
Published: (2026)
Effective Predictive Modeling for Emergency Department Visits and Evaluating Exogenous Variables Impact: Using Explainable Meta-learning Gradient Boosting
by: Neshat, Mehdi, et al.
Published: (2024)
by: Neshat, Mehdi, et al.
Published: (2024)
QL-LSTM: A Parameter-Efficient LSTM for Stable Long-Sequence Modeling
by: Nti, Isaac Kofi
Published: (2025)
by: Nti, Isaac Kofi
Published: (2025)
Children's Acquisition of Tail-recursion Sequences: A Review of Locative Recursion and Possessive Recursion as Examples
by: Wang, Xiaoyi, et al.
Published: (2024)
by: Wang, Xiaoyi, et al.
Published: (2024)
WASHH: An Anchor-Aware Whale-Guided Selection Hyper-Heuristic for Continuous Optimization and SVC Configuration
by: Zhao, Yifu, et al.
Published: (2026)
by: Zhao, Yifu, et al.
Published: (2026)
GLU Attention Improve Transformer
by: Wang, Zehao
Published: (2025)
by: Wang, Zehao
Published: (2025)
Hysteresis Activation Function for Efficient Inference
by: Kimhi, Moshe, et al.
Published: (2024)
by: Kimhi, Moshe, et al.
Published: (2024)
On the Power of Convolution Augmented Transformer
by: Li, Mingchen, et al.
Published: (2024)
by: Li, Mingchen, et al.
Published: (2024)
An enhanced Teaching-Learning-Based Optimization (TLBO) with Grey Wolf Optimizer (GWO) for text feature selection and clustering
by: Azarshab, Mahsa, et al.
Published: (2024)
by: Azarshab, Mahsa, et al.
Published: (2024)
Similar Items
-
SpikingSSMs: Learning Long Sequences with Sparse and Parallel Spiking State Space Models
by: Shen, Shuaijie, et al.
Published: (2024) -
Delay Embedding Theory of Neural Sequence Models
by: Ostrow, Mitchell, et al.
Published: (2024) -
Long-Sequence LSTM Modeling for NBA Game Outcome Prediction Using a Novel Multi-Season Dataset
by: Rios, Charles, et al.
Published: (2025) -
State-Space Constraints Can Improve the Generalisation of the Differentiable Neural Computer to Input Sequences With Unseen Length
by: Ofner, Patrick, et al.
Published: (2021) -
EvoX: Meta-Evolution for Automated Discovery
by: Liu, Shu, et al.
Published: (2026)