Adapting Vision-Language Models for E-commerce Understanding at Scale
Fuente:
arXiv
Guardado en:
| Autores principales: | Nulli, Matteo, Orshulevich, Vladimir, Bazazo, Tala, Herold, Christian, Kozielski, Michael, Mazur, Marcin, Tuzel, Szymon, Snoek, Cees G. M., Hashemi, Seyyed Hadi, Javed, Omar, Versley, Yannick, Khadivi, Shahram |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Domain Adaptation of Foundation LLMs for e-Commerce
por: Herold, Christian, et al.
Publicado: (2025)
por: Herold, Christian, et al.
Publicado: (2025)
Vocabulary Customization for Efficient Domain-Specific LLM Deployment
por: Herold, Christian, et al.
Publicado: (2025)
por: Herold, Christian, et al.
Publicado: (2025)
LiLiuM: eBay's Large Language Models for e-commerce
por: Herold, Christian, et al.
Publicado: (2024)
por: Herold, Christian, et al.
Publicado: (2024)
ClusComp: A Simple Paradigm for Model Compression and Efficient Finetuning
por: Liao, Baohao, et al.
Publicado: (2025)
por: Liao, Baohao, et al.
Publicado: (2025)
Unilogit: Robust Machine Unlearning for LLMs Using Uniform-Target Self-Distillation
por: Vasilev, Stefan, et al.
Publicado: (2025)
por: Vasilev, Stefan, et al.
Publicado: (2025)
Meeting SLOs, Slashing Hours: Automated Enterprise LLM Optimization with OptiKIT
por: Santavas, Nicholas, et al.
Publicado: (2026)
por: Santavas, Nicholas, et al.
Publicado: (2026)
ITEm: Unsupervised Image-Text Embedding Learning for eCommerce
por: Liao, Baohao, et al.
Publicado: (2023)
por: Liao, Baohao, et al.
Publicado: (2023)
ApiQ: Finetuning of 2-Bit Quantized Large Language Model
por: Liao, Baohao, et al.
Publicado: (2024)
por: Liao, Baohao, et al.
Publicado: (2024)
IKUN for WMT24 General MT Task: LLMs Are here for Multilingual Machine Translation
por: Liao, Baohao, et al.
Publicado: (2024)
por: Liao, Baohao, et al.
Publicado: (2024)
Beyond Model Adaptation at Test Time: A Survey
por: Xiao, Zehao, et al.
Publicado: (2024)
por: Xiao, Zehao, et al.
Publicado: (2024)
CONGRAD:Conflicting Gradient Filtering for Multilingual Preference Alignment
por: Li, Jiangnan, et al.
Publicado: (2025)
por: Li, Jiangnan, et al.
Publicado: (2025)
Analgesic Efficacy of Bromelain and Bromelain Plus Turmeric for Pain Control After Orthodontic Separator Placement: A Triple‐Blind Randomized Clinical Trial
por: Shabnam Ajami, et al.
Publicado: (2025)
por: Shabnam Ajami, et al.
Publicado: (2025)
Roadmap to Precision 3D Printing of Cellulose: Rheology‐Guided Formulation, Fidelity Assessment, and Application Horizons (Adv. Mater. Technol. 8/2026)
por: Majed Amini, et al.
Publicado: (2026)
por: Majed Amini, et al.
Publicado: (2026)
Low-Resource Vision Challenges for Foundation Models
por: Zhang, Yunhua, et al.
Publicado: (2024)
por: Zhang, Yunhua, et al.
Publicado: (2024)
Commonsense Video Question Answering through Video-Grounded Entailment Tree Reasoning
por: Liu, Huabin, et al.
Publicado: (2025)
por: Liu, Huabin, et al.
Publicado: (2025)
IPO: Interpretable Prompt Optimization for Vision-Language Models
por: Du, Yingjun, et al.
Publicado: (2024)
por: Du, Yingjun, et al.
Publicado: (2024)
Learn to Categorize or Categorize to Learn? Self-Coding for Generalized Category Discovery
por: Rastegar, Sarah, et al.
Publicado: (2023)
por: Rastegar, Sarah, et al.
Publicado: (2023)
A Survey of Visual Attention Models
por: Seyyed Mohammad Reza Hashemi
Publicado: (2015)
por: Seyyed Mohammad Reza Hashemi
Publicado: (2015)
LocoMotion: Learning Motion-Focused Video-Language Representations
por: Doughty, Hazel, et al.
Publicado: (2024)
por: Doughty, Hazel, et al.
Publicado: (2024)
'Explaining RL Decisions with Trajectories': A Reproducibility Study
por: Sadek, Karim Abdel, et al.
Publicado: (2024)
por: Sadek, Karim Abdel, et al.
Publicado: (2024)
GeneralizeFormer: Layer-Adaptive Model Generation across Test-Time Distribution Shifts
por: Ambekar, Sameer, et al.
Publicado: (2025)
por: Ambekar, Sameer, et al.
Publicado: (2025)
Dual Guidance Semi-Supervised Action Detection
por: Singh, Ankit, et al.
Publicado: (2025)
por: Singh, Ankit, et al.
Publicado: (2025)
SuperDisco: Super-Class Discovery Improves Visual Recognition for the Long-Tail
por: Du, Yingjun, et al.
Publicado: (2023)
por: Du, Yingjun, et al.
Publicado: (2023)
GateRA: Token-Aware Modulation for Parameter-Efficient Fine-Tuning
por: Ou, Jie, et al.
Publicado: (2025)
por: Ou, Jie, et al.
Publicado: (2025)
Beyond Coarse-Grained Matching in Video-Text Retrieval
por: Chen, Aozhu, et al.
Publicado: (2024)
por: Chen, Aozhu, et al.
Publicado: (2024)
SimPLR: A Simple and Plain Transformer for Efficient Object Detection and Segmentation
por: Nguyen, Duy-Kien, et al.
Publicado: (2023)
por: Nguyen, Duy-Kien, et al.
Publicado: (2023)
The Sound of Water: Inferring Physical Properties from Pouring Liquids
por: Bagad, Piyush, et al.
Publicado: (2024)
por: Bagad, Piyush, et al.
Publicado: (2024)
In-Context Learning Improves Compositional Understanding of Vision-Language Models
por: Nulli, Matteo, et al.
Publicado: (2024)
por: Nulli, Matteo, et al.
Publicado: (2024)
Active Continual Learning: On Balancing Knowledge Retention and Learnability
por: Vu, Thuy-Trang, et al.
Publicado: (2023)
por: Vu, Thuy-Trang, et al.
Publicado: (2023)
El uso de las redes sociales y la cultura popular para una mejor comprensión intercultural
por: Sait Tuzel
Publicado: (2017)
por: Sait Tuzel
Publicado: (2017)
Evaluation of Attribution Bias in Generator-Aware Retrieval-Augmented Large Language Models
por: Abolghasemi, Amin, et al.
Publicado: (2024)
por: Abolghasemi, Amin, et al.
Publicado: (2024)
What Layers When: Learning to Skip Compute in LLMs with Residual Gates
por: Laitenberger, Filipe, et al.
Publicado: (2025)
por: Laitenberger, Filipe, et al.
Publicado: (2025)
Segment Any 3D-Part in a Scene from a Sentence
por: Wu, Hongyu, et al.
Publicado: (2025)
por: Wu, Hongyu, et al.
Publicado: (2025)
Union-over-Intersections: Object Detection beyond Winner-Takes-All
por: Bhowmik, Aritra, et al.
Publicado: (2023)
por: Bhowmik, Aritra, et al.
Publicado: (2023)
PIN: Positional Insert Unlocks Object Localisation Abilities in VLMs
por: Dorkenwald, Michael, et al.
Publicado: (2024)
por: Dorkenwald, Michael, et al.
Publicado: (2024)
NeoBabel: A Multilingual Open Tower for Visual Generation
por: Derakhshani, Mohammad Mahdi, et al.
Publicado: (2025)
por: Derakhshani, Mohammad Mahdi, et al.
Publicado: (2025)
Dynamic Vocabulary Pruning in Early-Exit LLMs
por: Vincenti, Jort, et al.
Publicado: (2024)
por: Vincenti, Jort, et al.
Publicado: (2024)
Towards consistency of rule-based explainer and black box model -- fusion of rule induction and XAI-based feature importance
por: Kozielski, Michał, et al.
Publicado: (2024)
por: Kozielski, Michał, et al.
Publicado: (2024)
Atomistic investigation of deformation and fracture of individual structural components of metal matrix composites
por: Maździarz, Marcin, et al.
Publicado: (2024)
por: Maździarz, Marcin, et al.
Publicado: (2024)
PrAViC: Probabilistic Adaptation Framework for Real-Time Video Classification
por: Trędowicz, Magdalena, et al.
Publicado: (2024)
por: Trędowicz, Magdalena, et al.
Publicado: (2024)
Ejemplares similares
-
Domain Adaptation of Foundation LLMs for e-Commerce
por: Herold, Christian, et al.
Publicado: (2025) -
Vocabulary Customization for Efficient Domain-Specific LLM Deployment
por: Herold, Christian, et al.
Publicado: (2025) -
LiLiuM: eBay's Large Language Models for e-commerce
por: Herold, Christian, et al.
Publicado: (2024) -
ClusComp: A Simple Paradigm for Model Compression and Efficient Finetuning
por: Liao, Baohao, et al.
Publicado: (2025) -
Unilogit: Robust Machine Unlearning for LLMs Using Uniform-Target Self-Distillation
por: Vasilev, Stefan, et al.
Publicado: (2025)