PARSE: Part-Aware Relational Spatial Modeling
Fuente:
arXiv
Guardado en:
| Autores principales: | Bai, Yinuo, Xu, Peijun, Shao, Kuixiang, Jiao, Yuyang, Zhang, Jingxuan, Yao, Kaixin, Gu, Jiayuan, Yu, Jingyi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
XDen-1K: A Density Field Dataset of Real-World Objects
por: Zhang, Jingxuan, et al.
Publicado: (2025)
por: Zhang, Jingxuan, et al.
Publicado: (2025)
SPREAD: Spatial-Physical REasoning via geometry Aware Diffusion
por: Li, Minzhang, et al.
Publicado: (2026)
por: Li, Minzhang, et al.
Publicado: (2026)
CAST: Component-Aligned 3D Scene Reconstruction from an RGB Image
por: Yao, Kaixin, et al.
Publicado: (2025)
por: Yao, Kaixin, et al.
Publicado: (2025)
PartNeXt: A Next-Generation Dataset for Fine-Grained and Hierarchical 3D Part Understanding
por: Wang, Penghao, et al.
Publicado: (2025)
por: Wang, Penghao, et al.
Publicado: (2025)
DressCode: Autoregressively Sewing and Generating Garments from Text Guidance
por: He, Kai, et al.
Publicado: (2024)
por: He, Kai, et al.
Publicado: (2024)
TAPESTRY: From Geometry to Appearance via Consistent Turntable Videos
por: Zeng, Yan, et al.
Publicado: (2026)
por: Zeng, Yan, et al.
Publicado: (2026)
Efficient automatic segmentation for multi-level pulmonary arteries: The PARSE challenge
por: Luo, Gongning, et al.
Publicado: (2023)
por: Luo, Gongning, et al.
Publicado: (2023)
DualMamba: A Lightweight Spectral-Spatial Mamba-Convolution Network for Hyperspectral Image Classification
por: Sheng, Jiamu, et al.
Publicado: (2024)
por: Sheng, Jiamu, et al.
Publicado: (2024)
BG-Triangle: Bézier Gaussian Triangle for 3D Vectorization and Rendering
por: Wu, Minye, et al.
Publicado: (2025)
por: Wu, Minye, et al.
Publicado: (2025)
MouseGPT: A Large-scale Vision-Language Model for Mouse Behavior Analysis
por: Xu, Teng, et al.
Publicado: (2025)
por: Xu, Teng, et al.
Publicado: (2025)
Unfolding Spatial Cognition: Evaluating Multimodal Models on Visual Simulations
por: Li, Linjie, et al.
Publicado: (2025)
por: Li, Linjie, et al.
Publicado: (2025)
4DGen: Grounded 4D Content Generation with Spatial-temporal Consistency
por: Yin, Yuyang, et al.
Publicado: (2023)
por: Yin, Yuyang, et al.
Publicado: (2023)
V^3: Viewing Volumetric Videos on Mobiles via Streamable 2D Dynamic Gaussians
por: Wang, Penghao, et al.
Publicado: (2024)
por: Wang, Penghao, et al.
Publicado: (2024)
Explainable Part-Based Vehicle Classifier with Spatial Awareness
por: Caduff, Andreas, et al.
Publicado: (2026)
por: Caduff, Andreas, et al.
Publicado: (2026)
DMSSN: Distilled Mixed Spectral-Spatial Network for Hyperspectral Salient Object Detection
por: Qin, Haolin, et al.
Publicado: (2024)
por: Qin, Haolin, et al.
Publicado: (2024)
Refining CLIP's Spatial Awareness: A Visual-Centric Perspective
por: Qiu, Congpei, et al.
Publicado: (2025)
por: Qiu, Congpei, et al.
Publicado: (2025)
PARSE-Ego4D: Personal Action Recommendation Suggestions for Egocentric Videos
por: Abreu, Steven, et al.
Publicado: (2024)
por: Abreu, Steven, et al.
Publicado: (2024)
SAPNet++: Evolving Point-Prompted Instance Segmentation with Semantic and Spatial Awareness
por: Wei, Zhaoyang, et al.
Publicado: (2026)
por: Wei, Zhaoyang, et al.
Publicado: (2026)
Talking Head Generation via AU-Guided Landmark Prediction
por: Chang, Shao-Yu, et al.
Publicado: (2025)
por: Chang, Shao-Yu, et al.
Publicado: (2025)
Diffusion4D: Fast Spatial-temporal Consistent 4D Generation via Video Diffusion Models
por: Liang, Hanwen, et al.
Publicado: (2024)
por: Liang, Hanwen, et al.
Publicado: (2024)
SAKED: Mitigating Hallucination in Large Vision-Language Models via Stability-Aware Knowledge Enhanced Decoding
por: Li, Zhaoxu, et al.
Publicado: (2026)
por: Li, Zhaoxu, et al.
Publicado: (2026)
Relation-Aware Diffusion Model for Controllable Poster Layout Generation
por: Li, Fengheng, et al.
Publicado: (2023)
por: Li, Fengheng, et al.
Publicado: (2023)
Decoding with Structured Awareness: Integrating Directional, Frequency-Spatial, and Structural Attention for Medical Image Segmentation
por: Zhang, Fan, et al.
Publicado: (2025)
por: Zhang, Fan, et al.
Publicado: (2025)
Intervention-Aware Multiscale Representation Learning from Imaging Phenomics and Perturbation Transcriptomics
por: Chen, Jiayuan, et al.
Publicado: (2026)
por: Chen, Jiayuan, et al.
Publicado: (2026)
MambaMoE: Mixture-of-Spectral-Spatial-Experts State Space Model for Hyperspectral Image Classification
por: Xu, Yichu, et al.
Publicado: (2025)
por: Xu, Yichu, et al.
Publicado: (2025)
DUE: Dynamic Uncertainty-Aware Explanation Supervision via 3D Imputation
por: Zhao, Qilong, et al.
Publicado: (2024)
por: Zhao, Qilong, et al.
Publicado: (2024)
MADiff: Motion-Aware Mamba Diffusion Models for Hand Trajectory Prediction on Egocentric Videos
por: Ma, Junyi, et al.
Publicado: (2024)
por: Ma, Junyi, et al.
Publicado: (2024)
Grounded 3D-Aware Spatial Vision-Language Modeling
por: Cheng, An-Chieh, et al.
Publicado: (2026)
por: Cheng, An-Chieh, et al.
Publicado: (2026)
Pixel Perfect: Relational Image Quality Assessment with Spatially-Aware Distortions
por: Khan, Fadeel Sher, et al.
Publicado: (2026)
por: Khan, Fadeel Sher, et al.
Publicado: (2026)
Spatial-ORMLLM: Improve Spatial Relation Understanding in the Operating Room with Multimodal Large Language Model
por: He, Peiqi, et al.
Publicado: (2025)
por: He, Peiqi, et al.
Publicado: (2025)
TurboReg: TurboClique for Robust and Efficient Point Cloud Registration
por: Yan, Shaocheng, et al.
Publicado: (2025)
por: Yan, Shaocheng, et al.
Publicado: (2025)
RelMap: Enhancing Online Map Construction with Class-Aware Spatial Relation and Semantic Priors
por: Cai, Tianhui, et al.
Publicado: (2025)
por: Cai, Tianhui, et al.
Publicado: (2025)
SurgOnAir: Hierarchy-Aware Real-Time Surgical Video Commentary
por: He, Jingyi, et al.
Publicado: (2026)
por: He, Jingyi, et al.
Publicado: (2026)
THOR: Text to Human-Object Interaction Diffusion via Relation Intervention
por: Wu, Qianyang, et al.
Publicado: (2024)
por: Wu, Qianyang, et al.
Publicado: (2024)
InternSpatial: A Comprehensive Dataset for Spatial Reasoning in Vision-Language Models
por: Deng, Nianchen, et al.
Publicado: (2025)
por: Deng, Nianchen, et al.
Publicado: (2025)
Rethinking Where to Edit: Task-Aware Localization for Instruction-Based Image Editing
por: He, Jingxuan, et al.
Publicado: (2026)
por: He, Jingxuan, et al.
Publicado: (2026)
SimBase: A Simple Baseline for Temporal Video Grounding
por: Bao, Peijun, et al.
Publicado: (2024)
por: Bao, Peijun, et al.
Publicado: (2024)
FAFA: Frequency-Aware Flow-Aided Self-Supervision for Underwater Object Pose Estimation
por: Tang, Jingyi, et al.
Publicado: (2024)
por: Tang, Jingyi, et al.
Publicado: (2024)
Circuit Mechanisms for Spatial Relation Generation in Diffusion Transformers
por: Wang, Binxu, et al.
Publicado: (2026)
por: Wang, Binxu, et al.
Publicado: (2026)
Ego to World: Collaborative Spatial Reasoning in Embodied Systems via Reinforcement Learning
por: Zhou, Heng, et al.
Publicado: (2026)
por: Zhou, Heng, et al.
Publicado: (2026)
Ejemplares similares
-
XDen-1K: A Density Field Dataset of Real-World Objects
por: Zhang, Jingxuan, et al.
Publicado: (2025) -
SPREAD: Spatial-Physical REasoning via geometry Aware Diffusion
por: Li, Minzhang, et al.
Publicado: (2026) -
CAST: Component-Aligned 3D Scene Reconstruction from an RGB Image
por: Yao, Kaixin, et al.
Publicado: (2025) -
PartNeXt: A Next-Generation Dataset for Fine-Grained and Hierarchical 3D Part Understanding
por: Wang, Penghao, et al.
Publicado: (2025) -
DressCode: Autoregressively Sewing and Generating Garments from Text Guidance
por: He, Kai, et al.
Publicado: (2024)