TriAlignGR: Triangular Multitask Alignment with Multimodal Deep Interest Mining for Generative Recommendation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zeng, Yangchen, Peng, Hao, Guo, Rongfeng, Yu, Zhenyu, Hu, Zhiyuan, Wang, Jinze |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DeepInterestGR: Mining Deep Multi-Interest Using Multi-Modal LLMs for Generative Recommendation
von: Zeng, Yangchen, et al.
Veröffentlicht: (2026)
von: Zeng, Yangchen, et al.
Veröffentlicht: (2026)
Deep Interest Mining for Intent-Enriched Semantic IDs in Multimodal Generative Recommendation
von: Zeng, Yangchen, et al.
Veröffentlicht: (2026)
von: Zeng, Yangchen, et al.
Veröffentlicht: (2026)
Leveraging OpenFlamingo for Multimodal Embedding Analysis of C2C Car Parts Data
von: Rashid, Maisha Binte, et al.
Veröffentlicht: (2025)
von: Rashid, Maisha Binte, et al.
Veröffentlicht: (2025)
CourseTimeQA: A Lecture-Video Benchmark and a Latency-Constrained Cross-Modal Fusion Method for Timestamped QA
von: Kovalev, Vsevolod, et al.
Veröffentlicht: (2025)
von: Kovalev, Vsevolod, et al.
Veröffentlicht: (2025)
ReCoVR: Closing the Loop in Interactive Composed Video Retrieval
von: Zhang, Bingqing, et al.
Veröffentlicht: (2026)
von: Zhang, Bingqing, et al.
Veröffentlicht: (2026)
A Grounded Memory System For Smart Personal Assistants
von: Ocker, Felix, et al.
Veröffentlicht: (2025)
von: Ocker, Felix, et al.
Veröffentlicht: (2025)
ORPHEAS: A Cross-Lingual Greek-English Embedding Model for Retrieval-Augmented Generation
von: Livieris, Ioannis E., et al.
Veröffentlicht: (2026)
von: Livieris, Ioannis E., et al.
Veröffentlicht: (2026)
Evaluating Perspectival Biases in Cross-Modal Retrieval
von: Saengsukhiran, Teerapol, et al.
Veröffentlicht: (2025)
von: Saengsukhiran, Teerapol, et al.
Veröffentlicht: (2025)
Large Language Model for Qualitative Research -- A Systematic Mapping Study
von: Barros, Cauã Ferreira, et al.
Veröffentlicht: (2024)
von: Barros, Cauã Ferreira, et al.
Veröffentlicht: (2024)
Active Data
von: Arthur, Richard, et al.
Veröffentlicht: (2026)
von: Arthur, Richard, et al.
Veröffentlicht: (2026)
AI vs. Human Moderators: A Comparative Evaluation of Multimodal LLMs in Content Moderation for Brand Safety
von: Levi, Adi, et al.
Veröffentlicht: (2025)
von: Levi, Adi, et al.
Veröffentlicht: (2025)
Permanent Data Encoding (PDE): A Visual Language for Semantic Compression and Knowledge Preservation in 3-Character Units
von: Tsuyuki, Yoshiharu, et al.
Veröffentlicht: (2025)
von: Tsuyuki, Yoshiharu, et al.
Veröffentlicht: (2025)
AI-Friendly LaTeX: Using LaTeX Code as a Knowledge Source for Retrieval-Augmented Generation
von: Verhoeff, Tom
Veröffentlicht: (2026)
von: Verhoeff, Tom
Veröffentlicht: (2026)
MeVer at CheckThat! 2026: Cluster-Aware Hard-Negative Mining for Multilingual Scientific-Source Retrieval
von: Bakagianni, Juli, et al.
Veröffentlicht: (2026)
von: Bakagianni, Juli, et al.
Veröffentlicht: (2026)
A Case Study of Balanced Query Recommendation on Wikipedia
von: Mishra, Harshit, et al.
Veröffentlicht: (2025)
von: Mishra, Harshit, et al.
Veröffentlicht: (2025)
IMDMR: An Intelligent Multi-Dimensional Memory Retrieval System for Enhanced Conversational AI
von: Pawar, Tejas, et al.
Veröffentlicht: (2025)
von: Pawar, Tejas, et al.
Veröffentlicht: (2025)
Experimenting active and sequential learning in a medieval music manuscript
von: Sharma, Sachin, et al.
Veröffentlicht: (2025)
von: Sharma, Sachin, et al.
Veröffentlicht: (2025)
Loom: Hybrid Retrieval-Scoring Outfit Recommendation with Semantic Material Compatibility and Occasion-Aware Embedding Priors
von: Berlia, Anushree
Veröffentlicht: (2026)
von: Berlia, Anushree
Veröffentlicht: (2026)
LinkedOut: Linking World Knowledge Representation Out of Video LLM for Next-Generation Video Recommendation
von: Zhang, Haichao, et al.
Veröffentlicht: (2025)
von: Zhang, Haichao, et al.
Veröffentlicht: (2025)
DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark
von: Hu, Ruofan, et al.
Veröffentlicht: (2026)
von: Hu, Ruofan, et al.
Veröffentlicht: (2026)
Language processing in humans and computers
von: Pavlovic, Dusko
Veröffentlicht: (2024)
von: Pavlovic, Dusko
Veröffentlicht: (2024)
A Scalable and High Availability Solution for Recommending Resolutions to Problem Tickets
von: Saragadam, Harish, et al.
Veröffentlicht: (2025)
von: Saragadam, Harish, et al.
Veröffentlicht: (2025)
Optimizing Retrieval-Augmented Generation (RAG) for Colloquial Cantonese: A LoRA-Based Systematic Review
von: Calonge, David Santandreu, et al.
Veröffentlicht: (2025)
von: Calonge, David Santandreu, et al.
Veröffentlicht: (2025)
EdgeJury: Cross-Reviewed Small-Model Ensembles for Truthful Question Answering on Serverless Edge Inference
von: Kumar, Aayush
Veröffentlicht: (2025)
von: Kumar, Aayush
Veröffentlicht: (2025)
Perception-Aware Bias Detection for Query Suggestions
von: Haak, Fabian, et al.
Veröffentlicht: (2026)
von: Haak, Fabian, et al.
Veröffentlicht: (2026)
VOGUE: A Multimodal Dataset for Conversational Recommendation in Fashion
von: Guo, David, et al.
Veröffentlicht: (2025)
von: Guo, David, et al.
Veröffentlicht: (2025)
Assisted morbidity coding: the SISCO.web use case for identifying the main diagnosis in Hospital Discharge Records
von: Cardillo, Elena, et al.
Veröffentlicht: (2024)
von: Cardillo, Elena, et al.
Veröffentlicht: (2024)
Likert or Not: LLM Absolute Relevance Judgments on Fine-Grained Ordinal Scales
von: Godfrey, Charles, et al.
Veröffentlicht: (2025)
von: Godfrey, Charles, et al.
Veröffentlicht: (2025)
Complexity as Design Material
von: Windhager, Florian, et al.
Veröffentlicht: (2024)
von: Windhager, Florian, et al.
Veröffentlicht: (2024)
SPARTA: Scalable and Principled Benchmark of Tree-Structured Multi-hop QA over Text and Tables
von: Park, Sungho, et al.
Veröffentlicht: (2026)
von: Park, Sungho, et al.
Veröffentlicht: (2026)
Structured Interfaces for Automated Reasoning with 3D Scene Graphs
von: Ray, Aaron, et al.
Veröffentlicht: (2025)
von: Ray, Aaron, et al.
Veröffentlicht: (2025)
ByteRover: Agent-Native Memory Through LLM-Curated Hierarchical Context
von: Nguyen, Andy, et al.
Veröffentlicht: (2026)
von: Nguyen, Andy, et al.
Veröffentlicht: (2026)
IndoBERT-Sentiment: Context-Conditioned Sentiment Classification for Indonesian Text
von: Saputra, Muhammad Apriandito Arya, et al.
Veröffentlicht: (2026)
von: Saputra, Muhammad Apriandito Arya, et al.
Veröffentlicht: (2026)
IndoBERT-Relevancy: A Context-Conditioned Relevancy Classifier for Indonesian Text
von: Saputra, Muhammad Apriandito Arya, et al.
Veröffentlicht: (2026)
von: Saputra, Muhammad Apriandito Arya, et al.
Veröffentlicht: (2026)
Leveraging Lightweight Entity Extraction for Scalable Event-Based Image Retrieval
von: Minh, Dao Sy Duy, et al.
Veröffentlicht: (2025)
von: Minh, Dao Sy Duy, et al.
Veröffentlicht: (2025)
Quantifying and Narrowing the Unknown: Interactive Text-to-Video Retrieval via Uncertainty Minimization
von: Zhang, Bingqing, et al.
Veröffentlicht: (2025)
von: Zhang, Bingqing, et al.
Veröffentlicht: (2025)
DEUCE: Dual-diversity Enhancement and Uncertainty-awareness for Cold-start Active Learning
von: Guo, Jiaxin, et al.
Veröffentlicht: (2025)
von: Guo, Jiaxin, et al.
Veröffentlicht: (2025)
AVATAAR: Agentic Video Answering via Temporal Adaptive Alignment and Reasoning
von: Patel, Urjitkumar, et al.
Veröffentlicht: (2025)
von: Patel, Urjitkumar, et al.
Veröffentlicht: (2025)
Method for Aggregating Unstructured Data Using Large Language Models
von: Lazebnyi, Vsevolod, et al.
Veröffentlicht: (2026)
von: Lazebnyi, Vsevolod, et al.
Veröffentlicht: (2026)
Spatially-Grounded Document Retrieval via Patch-to-Region Relevance Propagation
von: Georgiou, Athos
Veröffentlicht: (2025)
von: Georgiou, Athos
Veröffentlicht: (2025)
Ähnliche Einträge
-
DeepInterestGR: Mining Deep Multi-Interest Using Multi-Modal LLMs for Generative Recommendation
von: Zeng, Yangchen, et al.
Veröffentlicht: (2026) -
Deep Interest Mining for Intent-Enriched Semantic IDs in Multimodal Generative Recommendation
von: Zeng, Yangchen, et al.
Veröffentlicht: (2026) -
Leveraging OpenFlamingo for Multimodal Embedding Analysis of C2C Car Parts Data
von: Rashid, Maisha Binte, et al.
Veröffentlicht: (2025) -
CourseTimeQA: A Lecture-Video Benchmark and a Latency-Constrained Cross-Modal Fusion Method for Timestamped QA
von: Kovalev, Vsevolod, et al.
Veröffentlicht: (2025) -
ReCoVR: Closing the Loop in Interactive Composed Video Retrieval
von: Zhang, Bingqing, et al.
Veröffentlicht: (2026)