Dual-Encoders for Extreme Multi-Label Classification
Fuente:
arXiv
Saved in:
| Main Authors: | Gupta, Nilesh, Khatri, Devvrit, Rawat, Ankit S, Bhojanapalli, Srinadh, Jain, Prateek, Dhillon, Inderjit |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scalable In-context Ranking with Generative Models
by: Gupta, Nilesh, et al.
Published: (2025)
by: Gupta, Nilesh, et al.
Published: (2025)
Autoregressive Ranking: Bridging the Gap Between Dual and Cross Encoders
by: Rozonoyer, Benjamin, et al.
Published: (2026)
by: Rozonoyer, Benjamin, et al.
Published: (2026)
EHI: End-to-end Learning of Hierarchical Index for Efficient Dense Retrieval
by: Kumar, Ramnath, et al.
Published: (2023)
by: Kumar, Ramnath, et al.
Published: (2023)
Compressing Many-Shots in In-Context Learning
by: Khatri, Devvrit, et al.
Published: (2025)
by: Khatri, Devvrit, et al.
Published: (2025)
LUCID: Attention with Preconditioned Representations
by: Duvvuri, Sai Surya, et al.
Published: (2026)
by: Duvvuri, Sai Surya, et al.
Published: (2026)
UniDEC : Unified Dual Encoder and Classifier Training for Extreme Multi-Label Classification
by: Kharbanda, Siddhant, et al.
Published: (2024)
by: Kharbanda, Siddhant, et al.
Published: (2024)
Interleaved Head Attention
by: Duvvuri, Sai Surya, et al.
Published: (2026)
by: Duvvuri, Sai Surya, et al.
Published: (2026)
HiRE: High Recall Approximate Top-$k$ Estimation for Efficient LLM Inference
by: L, Yashas Samaga B, et al.
Published: (2024)
by: L, Yashas Samaga B, et al.
Published: (2024)
Arithmetic Transformers Can Length-Generalize in Both Operand Length and Count
by: Cho, Hanseul, et al.
Published: (2024)
by: Cho, Hanseul, et al.
Published: (2024)
The Art of Scaling Reinforcement Learning Compute for LLMs
by: Khatri, Devvrit, et al.
Published: (2025)
by: Khatri, Devvrit, et al.
Published: (2025)
LLM-guided Hierarchical Search for End-to-end Reasoning Intensive Retrieval
by: Gupta, Nilesh, et al.
Published: (2025)
by: Gupta, Nilesh, et al.
Published: (2025)
MatFormer: Nested Transformer for Elastic Inference
by: Devvrit, et al.
Published: (2023)
by: Devvrit, et al.
Published: (2023)
Position Coupling: Improving Length Generalization of Arithmetic Transformers Using Task Structure
by: Cho, Hanseul, et al.
Published: (2024)
by: Cho, Hanseul, et al.
Published: (2024)
On student-teacher deviations in distillation: does it pay to disobey?
by: Nagarajan, Vaishnavh, et al.
Published: (2023)
by: Nagarajan, Vaishnavh, et al.
Published: (2023)
Mimetic Initialization Helps State Space Models Learn to Recall
by: Trockman, Asher, et al.
Published: (2024)
by: Trockman, Asher, et al.
Published: (2024)
LASER: Attention with Exponential Transformation
by: Duvvuri, Sai Surya, et al.
Published: (2024)
by: Duvvuri, Sai Surya, et al.
Published: (2024)
Learning, Fast and Slow: Towards LLMs That Adapt Continually
by: Tiwari, Rishabh, et al.
Published: (2026)
by: Tiwari, Rishabh, et al.
Published: (2026)
ODRPO: Ordinal Decompositions of Discrete Rewards for Robust Policy Optimization
by: Patel, Nirmal, et al.
Published: (2026)
by: Patel, Nirmal, et al.
Published: (2026)
Geometric Median (GM) Matching for Robust Data Pruning
by: Acharya, Anish, et al.
Published: (2024)
by: Acharya, Anish, et al.
Published: (2024)
Graph Regularized Encoder Training for Extreme Classification
by: Mittal, Anshul, et al.
Published: (2024)
by: Mittal, Anshul, et al.
Published: (2024)
Efficient Language Model Architectures for Differentially Private Federated Learning
by: Ro, Jae Hun, et al.
Published: (2024)
by: Ro, Jae Hun, et al.
Published: (2024)
Geometric Median Matching for Robust k-Subset Selection from Noisy Data
by: Acharya, Anish, et al.
Published: (2025)
by: Acharya, Anish, et al.
Published: (2025)
Retraining with Predicted Hard Labels Provably Increases Model Accuracy
by: Das, Rudrajit, et al.
Published: (2024)
by: Das, Rudrajit, et al.
Published: (2024)
Learning label-label correlations in Extreme Multi-label Classification via Label Features
by: Kharbanda, Siddhant, et al.
Published: (2024)
by: Kharbanda, Siddhant, et al.
Published: (2024)
Towards Quantifying the Preconditioning Effect of Adam
by: Das, Rudrajit, et al.
Published: (2024)
by: Das, Rudrajit, et al.
Published: (2024)
Matryoshka Model Learning for Improved Elastic Student Models
by: Verma, Chetan, et al.
Published: (2025)
by: Verma, Chetan, et al.
Published: (2025)
Functional Interpolation for Relative Positions Improves Long Context Transformers
by: Li, Shanda, et al.
Published: (2023)
by: Li, Shanda, et al.
Published: (2023)
Multi-Head Encoding for Extreme Label Classification
by: Liang, Daojun, et al.
Published: (2024)
by: Liang, Daojun, et al.
Published: (2024)
ICXML: An In-Context Learning Framework for Zero-Shot Extreme Multi-Label Classification
by: Zhu, Yaxin, et al.
Published: (2023)
by: Zhu, Yaxin, et al.
Published: (2023)
LightTopoGAT: Enhancing Graph Attention Networks with Topological Features for Efficient Graph Classification
by: Sharma, Ankit, et al.
Published: (2025)
by: Sharma, Ankit, et al.
Published: (2025)
Labels in Extremes: How Well Calibrated are Extreme Multi-label Classifiers?
by: Ullah, Nasib, et al.
Published: (2024)
by: Ullah, Nasib, et al.
Published: (2024)
Label Cluster Chains for Multi-Label Classification
by: Gatto, Elaine Cecília, et al.
Published: (2024)
by: Gatto, Elaine Cecília, et al.
Published: (2024)
Spark Transformer: Reactivating Sparsity in FFN and Attention
by: You, Chong, et al.
Published: (2025)
by: You, Chong, et al.
Published: (2025)
Estimating Causal Effects in Gaussian Linear SCMs with Finite Data
by: Maiti, Aurghya, et al.
Published: (2026)
by: Maiti, Aurghya, et al.
Published: (2026)
AEMLO: AutoEncoder-Guided Multi-Label Oversampling
by: Zhou, Ao, et al.
Published: (2024)
by: Zhou, Ao, et al.
Published: (2024)
Multi-Objective Trajectory Planning with Dual-Encoder
by: Zhang, Beibei, et al.
Published: (2024)
by: Zhang, Beibei, et al.
Published: (2024)
On the Necessity of World Knowledge for Mitigating Missing Labels in Extreme Classification
by: Prakash, Jatin, et al.
Published: (2024)
by: Prakash, Jatin, et al.
Published: (2024)
Scalable Label Distribution Learning for Multi-Label Classification
by: Zhao, Xingyu, et al.
Published: (2023)
by: Zhao, Xingyu, et al.
Published: (2023)
MultiPruner: Balanced Structure Removal in Foundation Models
by: Muñoz, J. Pablo, et al.
Published: (2025)
by: Muñoz, J. Pablo, et al.
Published: (2025)
Unsupervised Waste Classification By Dual-Encoder Contrastive Learning and Multi-Clustering Voting (DECMCV)
by: Huang, Kui, et al.
Published: (2025)
by: Huang, Kui, et al.
Published: (2025)
Similar Items
-
Scalable In-context Ranking with Generative Models
by: Gupta, Nilesh, et al.
Published: (2025) -
Autoregressive Ranking: Bridging the Gap Between Dual and Cross Encoders
by: Rozonoyer, Benjamin, et al.
Published: (2026) -
EHI: End-to-end Learning of Hierarchical Index for Efficient Dense Retrieval
by: Kumar, Ramnath, et al.
Published: (2023) -
Compressing Many-Shots in In-Context Learning
by: Khatri, Devvrit, et al.
Published: (2025) -
LUCID: Attention with Preconditioned Representations
by: Duvvuri, Sai Surya, et al.
Published: (2026)