COURIER: Contrastive User Intention Reconstruction for Large-Scale Visual Recommendation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Jia-Qi, Dai, Chenglei, OU, Dan, Li, Dongshuai, Huang, Ju, Zhan, De-Chuan, Zeng, Xiaoyi, Yang, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Revisiting Content-Based Music Recommendation: Efficient Feature Aggregation from Large-Scale Music Models
von: Zhou, Yizhi, et al.
Veröffentlicht: (2026)
von: Zhou, Yizhi, et al.
Veröffentlicht: (2026)
Guiding Diffusion-based Reconstruction with Contrastive Signals for Balanced Visual Representation
von: Han, Boyu, et al.
Veröffentlicht: (2026)
von: Han, Boyu, et al.
Veröffentlicht: (2026)
Predicting User Grasp Intentions in Virtual Reality
von: Zeng, Linghao
Veröffentlicht: (2025)
von: Zeng, Linghao
Veröffentlicht: (2025)
PM25Vision: A Large-Scale Benchmark Dataset for Visual Estimation of Air Quality
von: Han, Yang
Veröffentlicht: (2025)
von: Han, Yang
Veröffentlicht: (2025)
Learning Group Interactions and Semantic Intentions for Multi-Object Trajectory Prediction
von: Qi, Mengshi, et al.
Veröffentlicht: (2024)
von: Qi, Mengshi, et al.
Veröffentlicht: (2024)
Improving LLMs for Recommendation with Out-Of-Vocabulary Tokens
von: Huang, Ting-Ji, et al.
Veröffentlicht: (2024)
von: Huang, Ting-Ji, et al.
Veröffentlicht: (2024)
Mutual Information guided Visual Contrastive Learning
von: Chen, Hanyang, et al.
Veröffentlicht: (2025)
von: Chen, Hanyang, et al.
Veröffentlicht: (2025)
CACE-Net: Co-guidance Attention and Contrastive Enhancement for Effective Audio-Visual Event Localization
von: He, Xiang, et al.
Veröffentlicht: (2024)
von: He, Xiang, et al.
Veröffentlicht: (2024)
EfficientGS: Streamlining Gaussian Splatting for Large-Scale High-Resolution Scene Representation
von: Liu, Wenkai, et al.
Veröffentlicht: (2024)
von: Liu, Wenkai, et al.
Veröffentlicht: (2024)
eMotions: A Large-Scale Dataset and Audio-Visual Fusion Network for Emotion Analysis in Short-form Videos
von: Wu, Xuecheng, et al.
Veröffentlicht: (2025)
von: Wu, Xuecheng, et al.
Veröffentlicht: (2025)
Free Geometry: Refining 3D Reconstruction from Longer Versions of Itself
von: Dai, Yuhang, et al.
Veröffentlicht: (2026)
von: Dai, Yuhang, et al.
Veröffentlicht: (2026)
MosaicDoc: A Large-Scale Bilingual Benchmark for Visually Rich Document Understanding
von: Chen, Ketong, et al.
Veröffentlicht: (2025)
von: Chen, Ketong, et al.
Veröffentlicht: (2025)
EMIE-MAP: Large-Scale Road Surface Reconstruction Based on Explicit Mesh and Implicit Encoding
von: Wu, Wenhua, et al.
Veröffentlicht: (2024)
von: Wu, Wenhua, et al.
Veröffentlicht: (2024)
3DGS-DET: Empower 3D Gaussian Splatting with Boundary Guidance and Box-Focused Sampling for Indoor 3D Object Detection
von: Cao, Yang, et al.
Veröffentlicht: (2024)
von: Cao, Yang, et al.
Veröffentlicht: (2024)
CHASD: Language Increment-Calibrated Contrastive Decoding against Hallucination in LVLMs
von: Huang, Xiaoyi, et al.
Veröffentlicht: (2026)
von: Huang, Xiaoyi, et al.
Veröffentlicht: (2026)
VIAssist: Adapting Multi-modal Large Language Models for Users with Visual Impairments
von: Yang, Bufang, et al.
Veröffentlicht: (2024)
von: Yang, Bufang, et al.
Veröffentlicht: (2024)
PromptForge-350k: A Large-Scale Dataset and Contrastive Framework for Prompt-Based AI Image Forgery Localization
von: Wang, Jianpeng, et al.
Veröffentlicht: (2026)
von: Wang, Jianpeng, et al.
Veröffentlicht: (2026)
Understanding Bias in Large-Scale Visual Datasets
von: Zeng, Boya, et al.
Veröffentlicht: (2024)
von: Zeng, Boya, et al.
Veröffentlicht: (2024)
Enhancing Bandit Algorithms with LLMs for Time-varying User Preferences in Streaming Recommendations
von: Shen, Chenglei, et al.
Veröffentlicht: (2026)
von: Shen, Chenglei, et al.
Veröffentlicht: (2026)
3CAD: A Large-Scale Real-World 3C Product Dataset for Unsupervised Anomaly
von: Yang, Enquan, et al.
Veröffentlicht: (2025)
von: Yang, Enquan, et al.
Veröffentlicht: (2025)
CityGaussianV2: Efficient and Geometrically Accurate Reconstruction for Large-Scale Scenes
von: Liu, Yang, et al.
Veröffentlicht: (2024)
von: Liu, Yang, et al.
Veröffentlicht: (2024)
Pyramid Feature Attention Network for Monocular Depth Prediction
von: Xu, Yifang, et al.
Veröffentlicht: (2024)
von: Xu, Yifang, et al.
Veröffentlicht: (2024)
Fitting Different Interactive Information: Joint Classification of Emotion and Intention
von: Li, Xinger, et al.
Veröffentlicht: (2025)
von: Li, Xinger, et al.
Veröffentlicht: (2025)
TV100: A TV Series Dataset that Pre-Trained CLIP Has Not Seen
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024)
von: Zhou, Da-Wei, et al.
Veröffentlicht: (2024)
Polaris: Scaling Up Instruction-Guided Image Generation Towards Millions of Personalized Style Needs
von: Chen, Zhi-Kai, et al.
Veröffentlicht: (2026)
von: Chen, Zhi-Kai, et al.
Veröffentlicht: (2026)
Contrastive Conditional Alignment based on Label Shift Calibration for Imbalanced Domain Adaptation
von: Sun, Xiaona, et al.
Veröffentlicht: (2024)
von: Sun, Xiaona, et al.
Veröffentlicht: (2024)
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
von: Si, Chenglei, et al.
Veröffentlicht: (2024)
von: Si, Chenglei, et al.
Veröffentlicht: (2024)
Total-Decom: Decomposed 3D Scene Reconstruction with Minimal Interaction
von: Lyu, Xiaoyang, et al.
Veröffentlicht: (2024)
von: Lyu, Xiaoyang, et al.
Veröffentlicht: (2024)
Jointly Understand Your Command and Intention:Reciprocal Co-Evolution between Scene-Aware 3D Human Motion Synthesis and Analysis
von: Gao, Xuehao, et al.
Veröffentlicht: (2025)
von: Gao, Xuehao, et al.
Veröffentlicht: (2025)
Understanding the Implicit User Intention via Reasoning with Large Language Model for Image Editing
von: Wang, Yijia, et al.
Veröffentlicht: (2025)
von: Wang, Yijia, et al.
Veröffentlicht: (2025)
PhysPatch: A Physically Realizable and Transferable Adversarial Patch Attack for Multimodal Large Language Models-based Autonomous Driving Systems
von: Guo, Qi, et al.
Veröffentlicht: (2025)
von: Guo, Qi, et al.
Veröffentlicht: (2025)
Efficient Large Multi-modal Models via Visual Context Compression
von: Chen, Jieneng, et al.
Veröffentlicht: (2024)
von: Chen, Jieneng, et al.
Veröffentlicht: (2024)
VideoUFO: A Million-Scale User-Focused Dataset for Text-to-Video Generation
von: Wang, Wenhao, et al.
Veröffentlicht: (2025)
von: Wang, Wenhao, et al.
Veröffentlicht: (2025)
3D Vision and Language Pretraining with Large-Scale Synthetic Data
von: Yang, Dejie, et al.
Veröffentlicht: (2024)
von: Yang, Dejie, et al.
Veröffentlicht: (2024)
VRU-CIPI: Crossing Intention Prediction at Intersections for Improving Vulnerable Road Users Safety
von: Abdelrahman, Ahmed S., et al.
Veröffentlicht: (2025)
von: Abdelrahman, Ahmed S., et al.
Veröffentlicht: (2025)
Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering
von: Si, Chenglei, et al.
Veröffentlicht: (2024)
von: Si, Chenglei, et al.
Veröffentlicht: (2024)
M4Human: A Large-Scale Multimodal mmWave Radar Benchmark for Human Mesh Reconstruction
von: Fan, Junqiao, et al.
Veröffentlicht: (2025)
von: Fan, Junqiao, et al.
Veröffentlicht: (2025)
PCLVis: Visual Analytics of Process Communication Latency in Large-Scale Simulation
von: Bi, Chongke, et al.
Veröffentlicht: (2025)
von: Bi, Chongke, et al.
Veröffentlicht: (2025)
STORM: Spatio-Temporal Reconstruction Model for Large-Scale Outdoor Scenes
von: Yang, Jiawei, et al.
Veröffentlicht: (2024)
von: Yang, Jiawei, et al.
Veröffentlicht: (2024)
RoMe: Towards Large Scale Road Surface Reconstruction via Mesh Representation
von: Mei, Ruohong, et al.
Veröffentlicht: (2023)
von: Mei, Ruohong, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Revisiting Content-Based Music Recommendation: Efficient Feature Aggregation from Large-Scale Music Models
von: Zhou, Yizhi, et al.
Veröffentlicht: (2026) -
Guiding Diffusion-based Reconstruction with Contrastive Signals for Balanced Visual Representation
von: Han, Boyu, et al.
Veröffentlicht: (2026) -
Predicting User Grasp Intentions in Virtual Reality
von: Zeng, Linghao
Veröffentlicht: (2025) -
PM25Vision: A Large-Scale Benchmark Dataset for Visual Estimation of Air Quality
von: Han, Yang
Veröffentlicht: (2025) -
Learning Group Interactions and Semantic Intentions for Multi-Object Trajectory Prediction
von: Qi, Mengshi, et al.
Veröffentlicht: (2024)