Orchid: Flexible and Data-Dependent Convolution for Sequence Modeling
Fuente:
arXiv
Saved in:
| Main Authors: | Karami, Mahdi, Ghodsi, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Auto-Regressive Masked Diffusion Models
by: Karami, Mahdi, et al.
Published: (2026)
by: Karami, Mahdi, et al.
Published: (2026)
MS-SSM: A Multi-Scale State Space Model for Efficient Sequence Modeling
by: Karami, Mahdi, et al.
Published: (2025)
by: Karami, Mahdi, et al.
Published: (2025)
How Many Heads Make an SSM? A Unified Framework for Attention and State Space Models
by: Ghodsi, Ali
Published: (2025)
by: Ghodsi, Ali
Published: (2025)
Best of Both Worlds: Advantages of Hybrid Graph Sequence Models
by: Behrouz, Ali, et al.
Published: (2024)
by: Behrouz, Ali, et al.
Published: (2024)
HiGen: Hierarchical Graph Generative Networks
by: Karami, Mahdi
Published: (2023)
by: Karami, Mahdi
Published: (2023)
WERank: Towards Rank Degradation Prevention for Self-Supervised Learning Using Weight Regularization
by: Pasand, Ali Saheb, et al.
Published: (2024)
by: Pasand, Ali Saheb, et al.
Published: (2024)
Trellis: Learning to Compress Key-Value Memory in Attention Models
by: Karami, Mahdi, et al.
Published: (2025)
by: Karami, Mahdi, et al.
Published: (2025)
Scalable Graph Self-Supervised Learning
by: Pasand, Ali Saheb, et al.
Published: (2024)
by: Pasand, Ali Saheb, et al.
Published: (2024)
Lattice: Learning to Efficiently Compress the Memory
by: Karami, Mahdi, et al.
Published: (2025)
by: Karami, Mahdi, et al.
Published: (2025)
Disentangling the Complex Multiplexed DIA Spectra in De Novo Peptide Sequencing
by: Ma, Zheng, et al.
Published: (2024)
by: Ma, Zheng, et al.
Published: (2024)
GraphPI: Efficient Protein Inference with Graph Neural Networks
by: Ma, Zheng, et al.
Published: (2026)
by: Ma, Zheng, et al.
Published: (2026)
Learning Chemotherapy Drug Action via Universal Physics-Informed Neural Networks
by: Podina, Lena, et al.
Published: (2024)
by: Podina, Lena, et al.
Published: (2024)
Multi-Mixer Models: Flexible Sequence Modeling with Shared Representations
by: Li, Kevin Y., et al.
Published: (2026)
by: Li, Kevin Y., et al.
Published: (2026)
BlockGen: Flexible Blockwise Sequence Modeling with Hybrid Samplers
by: Deschenaux, Justin, et al.
Published: (2026)
by: Deschenaux, Justin, et al.
Published: (2026)
Graph Anomaly Detection in Time Series: A Survey
by: Ho, Thi Kieu Khanh, et al.
Published: (2023)
by: Ho, Thi Kieu Khanh, et al.
Published: (2023)
Sorted LLaMA: Unlocking the Potential of Intermediate Layers of Large Language Models for Dynamic Inference
by: Kavehzadeh, Parsa, et al.
Published: (2023)
by: Kavehzadeh, Parsa, et al.
Published: (2023)
Flexible Infinite-Width Graph Convolutional Neural Networks
by: Anson, Ben, et al.
Published: (2024)
by: Anson, Ben, et al.
Published: (2024)
Reparameterized Multi-Resolution Convolutions for Long Sequence Modelling
by: Cunningham, Harry Jake, et al.
Published: (2024)
by: Cunningham, Harry Jake, et al.
Published: (2024)
TNT: Improving Chunkwise Training for Test-Time Memorization
by: Li, Zeman, et al.
Published: (2025)
by: Li, Zeman, et al.
Published: (2025)
AI-based Mapping of the Conservation Status of Orchid Assemblages at Global Scale
by: Estopinan, Joaquim, et al.
Published: (2024)
by: Estopinan, Joaquim, et al.
Published: (2024)
FutureFill: Fast Generation from Convolutional Sequence Models
by: Agarwal, Naman, et al.
Published: (2024)
by: Agarwal, Naman, et al.
Published: (2024)
One RNG to Rule Them All: How Randomness Becomes an Attack Vector in Machine Learning
by: Prabhu, Kotekar Annapoorna, et al.
Published: (2026)
by: Prabhu, Kotekar Annapoorna, et al.
Published: (2026)
SortedNet: A Scalable and Generalized Framework for Training Modular Deep Neural Networks
by: Valipour, Mojtaba, et al.
Published: (2023)
by: Valipour, Mojtaba, et al.
Published: (2023)
Mamba Meets Scheduling: Learning to Solve Flexible Job Shop Scheduling with Efficient Sequence Modeling
by: Cao, Zhi, et al.
Published: (2026)
by: Cao, Zhi, et al.
Published: (2026)
Boosting t-SNE Efficiency for Sequencing Data: Insights from Kernel Selection
by: Jan, Avais, et al.
Published: (2025)
by: Jan, Avais, et al.
Published: (2025)
Flash Inference: Near Linear Time Inference for Long Convolution Sequence Models and Beyond
by: Oncescu, Costin-Andrei, et al.
Published: (2024)
by: Oncescu, Costin-Andrei, et al.
Published: (2024)
mGRADE: Minimal Recurrent Gating Meets Delay Convolutions for Lightweight Sequence Modeling
by: Torchet, Tristan, et al.
Published: (2025)
by: Torchet, Tristan, et al.
Published: (2025)
GSA-Forecaster: Forecasting Graph-Based Time-Dependent Data with Graph Sequence Attention
by: Li, Yang, et al.
Published: (2021)
by: Li, Yang, et al.
Published: (2021)
TEE4EHR: Transformer Event Encoder for Better Representation Learning in Electronic Health Records
by: Karami, Hojjat, et al.
Published: (2024)
by: Karami, Hojjat, et al.
Published: (2024)
A Comprehensive Forecasting-Based Framework for Time Series Anomaly Detection: Benchmarking on the Numenta Anomaly Benchmark (NAB)
by: Karami, Mohammad, et al.
Published: (2025)
by: Karami, Mohammad, et al.
Published: (2025)
Short-Long Convolutions Help Hardware-Efficient Linear Attention to Focus on Long Sequences
by: Liu, Zicheng, et al.
Published: (2024)
by: Liu, Zicheng, et al.
Published: (2024)
Orchid: Image Latent Diffusion for Joint Appearance and Geometry Generation
by: Krishnan, Akshay, et al.
Published: (2025)
by: Krishnan, Akshay, et al.
Published: (2025)
EchoAtt: Attend, Copy, then Adjust for More Efficient Large Language Models
by: Rajabzadeh, Hossein, et al.
Published: (2024)
by: Rajabzadeh, Hossein, et al.
Published: (2024)
QDyLoRA: Quantized Dynamic Low-Rank Adaptation for Efficient Large Language Model Tuning
by: Rajabzadeh, Hossein, et al.
Published: (2024)
by: Rajabzadeh, Hossein, et al.
Published: (2024)
Algorithm- and Data-Dependent Generalization Bounds for Diffusion Models
by: Dupuis, Benjamin, et al.
Published: (2025)
by: Dupuis, Benjamin, et al.
Published: (2025)
Optimizing Portfolio with Two-Sided Transactions and Lending: A Reinforcement Learning Framework
by: Habibnia, Ali, et al.
Published: (2024)
by: Habibnia, Ali, et al.
Published: (2024)
CAREER: A Foundation Model for Labor Sequence Data
by: Vafa, Keyon, et al.
Published: (2022)
by: Vafa, Keyon, et al.
Published: (2022)
Sequence-to-Image Transformation for Sequence Classification Using Rips Complex Construction and Chaos Game Representation
by: Ali, Sarwan, et al.
Published: (2025)
by: Ali, Sarwan, et al.
Published: (2025)
Cross-Paradigm Knowledge Distillation: A Comprehensive Study of Bidirectional Transfer Between Random Forests and Deep Neural Networks for Big Data Applications
by: Moghadasi, Mahdi Naser
Published: (2026)
by: Moghadasi, Mahdi Naser
Published: (2026)
Transformer Model for Alzheimer's Disease Progression Prediction Using Longitudinal Visit Sequences
by: Moghaddami, Mahdi, et al.
Published: (2025)
by: Moghaddami, Mahdi, et al.
Published: (2025)
Similar Items
-
Auto-Regressive Masked Diffusion Models
by: Karami, Mahdi, et al.
Published: (2026) -
MS-SSM: A Multi-Scale State Space Model for Efficient Sequence Modeling
by: Karami, Mahdi, et al.
Published: (2025) -
How Many Heads Make an SSM? A Unified Framework for Attention and State Space Models
by: Ghodsi, Ali
Published: (2025) -
Best of Both Worlds: Advantages of Hybrid Graph Sequence Models
by: Behrouz, Ali, et al.
Published: (2024) -
HiGen: Hierarchical Graph Generative Networks
by: Karami, Mahdi
Published: (2023)