Bounded-Compute Multimodal Regression for Product-Rating Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Leach, William, He, Ru, Ma, Sizhuo, Jia, Yizhen, Cao, Min, Wang, Jian, Cao, Rick |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
gQIR: Generative Quanta Image Reconstruction
by: Garg, Aryan, et al.
Published: (2026)
by: Garg, Aryan, et al.
Published: (2026)
Get Your Embedding Space in Order: Domain-Adaptive Regression for Forest Monitoring
by: Li, Sizhuo, et al.
Published: (2024)
by: Li, Sizhuo, et al.
Published: (2024)
InstantRestore: Single-Step Personalized Face Restoration with Shared-Image Attention
by: Zhang, Howard, et al.
Published: (2024)
by: Zhang, Howard, et al.
Published: (2024)
Engagement Prediction of Short Videos with Large Multimodal Models
by: Sun, Wei, et al.
Published: (2025)
by: Sun, Wei, et al.
Published: (2025)
Privacy-Preserving Visual Localization with Event Cameras
by: Kim, Junho, et al.
Published: (2022)
by: Kim, Junho, et al.
Published: (2022)
Delving Deep into Engagement Prediction of Short Videos
by: Li, Dasong, et al.
Published: (2024)
by: Li, Dasong, et al.
Published: (2024)
STORM: Benchmarking Visual Rating of MLLMs with a Comprehensive Ordinal Regression Dataset
by: Wang, Jinhong, et al.
Published: (2025)
by: Wang, Jinhong, et al.
Published: (2025)
Velocity Disambiguation for Video Frame Interpolation
by: Zhong, Zhihang, et al.
Published: (2023)
by: Zhong, Zhihang, et al.
Published: (2023)
GraphRevisedIE: Multimodal Information Extraction with Graph-Revised Network
by: Cao, Panfeng, et al.
Published: (2024)
by: Cao, Panfeng, et al.
Published: (2024)
Towards Lightweight Super-Resolution with Dual Regression Learning
by: Guo, Yong, et al.
Published: (2022)
by: Guo, Yong, et al.
Published: (2022)
Technical Report for CVPR 2024 WeatherProof Dataset Challenge: Semantic Segmentation on Paired Real Data
by: Cao, Guojin, et al.
Published: (2024)
by: Cao, Guojin, et al.
Published: (2024)
GuideSR: Rethinking Guidance for One-Step High-Fidelity Diffusion-Based Super-Resolution
by: Arora, Aditya, et al.
Published: (2025)
by: Arora, Aditya, et al.
Published: (2025)
Faster and Better 3D Splatting via Group Training
by: Wang, Chengbo, et al.
Published: (2024)
by: Wang, Chengbo, et al.
Published: (2024)
CineMatte: Background Matting for Virtual Production and Beyond
by: He, Yuanjian, et al.
Published: (2026)
by: He, Yuanjian, et al.
Published: (2026)
MMAR: Towards Lossless Multi-Modal Auto-Regressive Probabilistic Modeling
by: Yang, Jian, et al.
Published: (2024)
by: Yang, Jian, et al.
Published: (2024)
Test-Time Computing for Referring Multimodal Large Language Models
by: Wu, Mingrui, et al.
Published: (2026)
by: Wu, Mingrui, et al.
Published: (2026)
FPDIoU Loss: A Loss Function for Efficient Bounding Box Regression of Rotated Object Detection
by: Ma, Siliang, et al.
Published: (2024)
by: Ma, Siliang, et al.
Published: (2024)
See Less, See Right: Bi-directional Perceptual Shaping For Multimodal Reasoning
by: Zhang, Shuoshuo, et al.
Published: (2025)
by: Zhang, Shuoshuo, et al.
Published: (2025)
Towards Redundancy Reduction in Diffusion Models for Efficient Video Super-Resolution
by: Guo, Jinpei, et al.
Published: (2025)
by: Guo, Jinpei, et al.
Published: (2025)
MPDIoU: A Loss for Efficient and Accurate Bounding Box Regression
by: Ma, Siliang, et al.
Published: (2023)
by: Ma, Siliang, et al.
Published: (2023)
MA-LMM: Memory-Augmented Large Multimodal Model for Long-Term Video Understanding
by: He, Bo, et al.
Published: (2024)
by: He, Bo, et al.
Published: (2024)
Learning Domain Knowledge in Multimodal Large Language Models through Reinforcement Fine-Tuning
by: Cao, Qinglong, et al.
Published: (2026)
by: Cao, Qinglong, et al.
Published: (2026)
CREMD: Crowd-Sourced Emotional Multimodal Dogs Dataset
by: Baek, Jinho, et al.
Published: (2026)
by: Baek, Jinho, et al.
Published: (2026)
SuperMat: Physically Consistent PBR Material Estimation at Interactive Rates
by: Hong, Yijia, et al.
Published: (2024)
by: Hong, Yijia, et al.
Published: (2024)
SVC 2025: the First Multimodal Deception Detection Challenge
by: Lin, Xun, et al.
Published: (2025)
by: Lin, Xun, et al.
Published: (2025)
Decoupling the Image Perception and Multimodal Reasoning for Reasoning Segmentation with Digital Twin Representations
by: Li, Yizhen, et al.
Published: (2025)
by: Li, Yizhen, et al.
Published: (2025)
Reprojection Errors as Prompts for Efficient Scene Coordinate Regression
by: Liu, Ting-Ru, et al.
Published: (2024)
by: Liu, Ting-Ru, et al.
Published: (2024)
A Multimodal In-Context Tuning Approach for E-Commerce Product Description Generation
by: Li, Yunxin, et al.
Published: (2024)
by: Li, Yunxin, et al.
Published: (2024)
ARMOR: Empowering Multimodal Understanding Model with Interleaved Multimodal Generation Capability
by: Sun, Jianwen, et al.
Published: (2025)
by: Sun, Jianwen, et al.
Published: (2025)
Nested AutoRegressive Models
by: Wu, Hongyu, et al.
Published: (2025)
by: Wu, Hongyu, et al.
Published: (2025)
A Survey on Ordinal Regression: Applications, Advances and Prospects
by: Wang, Jinhong, et al.
Published: (2025)
by: Wang, Jinhong, et al.
Published: (2025)
HetSSNet: Spatial-Spectral Heterogeneous Graph Learning Network for Panchromatic and Multispectral Images Fusion
by: Ma, Mengting, et al.
Published: (2025)
by: Ma, Mengting, et al.
Published: (2025)
Spatial-Spectral Binarized Neural Network for Panchromatic and Multi-spectral Images Fusion
by: Jiang, Yizhen, et al.
Published: (2025)
by: Jiang, Yizhen, et al.
Published: (2025)
LAIP: Learning Local Alignment from Image-Phrase Modeling for Text-based Person Search
by: Wang, Haiguang, et al.
Published: (2024)
by: Wang, Haiguang, et al.
Published: (2024)
Lung Nodule Image Synthesis Driven by Two-Stage Generative Adversarial Networks
by: Cao, Lu, et al.
Published: (2026)
by: Cao, Lu, et al.
Published: (2026)
A Novel Bounding Box Regression Method for Single Object Tracking
by: Abdelaziz, Omar, et al.
Published: (2024)
by: Abdelaziz, Omar, et al.
Published: (2024)
WeCKD: Weakly-supervised Chained Distillation Network for Efficient Multimodal Medical Imaging
by: Rahman, Md. Abdur, et al.
Published: (2025)
by: Rahman, Md. Abdur, et al.
Published: (2025)
TIER: Text-Image Encoder-based Regression for AIGC Image Quality Assessment
by: Yuan, Jiquan, et al.
Published: (2024)
by: Yuan, Jiquan, et al.
Published: (2024)
PlanGen: Towards Unified Layout Planning and Image Generation in Auto-Regressive Vision Language Models
by: He, Runze, et al.
Published: (2025)
by: He, Runze, et al.
Published: (2025)
Resurrect Mask AutoRegressive Modeling for Efficient and Scalable Image Generation
by: Xin, Yi, et al.
Published: (2025)
by: Xin, Yi, et al.
Published: (2025)
Similar Items
-
gQIR: Generative Quanta Image Reconstruction
by: Garg, Aryan, et al.
Published: (2026) -
Get Your Embedding Space in Order: Domain-Adaptive Regression for Forest Monitoring
by: Li, Sizhuo, et al.
Published: (2024) -
InstantRestore: Single-Step Personalized Face Restoration with Shared-Image Attention
by: Zhang, Howard, et al.
Published: (2024) -
Engagement Prediction of Short Videos with Large Multimodal Models
by: Sun, Wei, et al.
Published: (2025) -
Privacy-Preserving Visual Localization with Event Cameras
by: Kim, Junho, et al.
Published: (2022)