Saved in:
| Main Authors: | Yin, Qian, Wen, Di, Peng, Kunyu, Schneider, David, Zhong, Zeyun, Jaus, Alexander, Marinov, Zdravko, Wei, Jiale, Liu, Ruiping, Zheng, Junwei, Chen, Yufan, Zhang, Chen, Qi, Lei, Stiefelhagen, Rainer |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.01668 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
IMPACT-CYCLE: A Contract-Based Multi-Agent System for Claim-Level Supervisory Correction of Long-Video Semantic Memory
by: Kong, Weitong, et al.
Published: (2026)
by: Kong, Weitong, et al.
Published: (2026)
IMPACT-HOI: Supervisory Control for Onset-Anchored Partial HOI Event Construction
by: Zhang, Haoshen, et al.
Published: (2026)
by: Zhang, Haoshen, et al.
Published: (2026)
OmniFall: From Staged Through Synthetic to Wild, A Unified Multi-Domain Dataset for Robust Fall Detection
by: Schneider, David, et al.
Published: (2025)
by: Schneider, David, et al.
Published: (2025)
LIMIS: Towards Language-based Interactive Medical Image Segmentation
by: Heinemann, Lena, et al.
Published: (2024)
by: Heinemann, Lena, et al.
Published: (2024)
Rethinking Annotator Simulation: Realistic Evaluation of Whole-Body PET Lesion Interactive Segmentation Methods
by: Marinov, Zdravko, et al.
Published: (2024)
by: Marinov, Zdravko, et al.
Published: (2024)
Good Enough: Is it Worth Improving your Label Quality?
by: Jaus, Alexander, et al.
Published: (2025)
by: Jaus, Alexander, et al.
Published: (2025)
Deep Interactive Segmentation of Medical Images: A Systematic Review and Taxonomy
by: Marinov, Zdravko, et al.
Published: (2023)
by: Marinov, Zdravko, et al.
Published: (2023)
Open Panoramic Segmentation
by: Zheng, Junwei, et al.
Published: (2024)
by: Zheng, Junwei, et al.
Published: (2024)
Every Component Counts: Rethinking the Measure of Success for Medical Semantic Segmentation in Multi-Instance Segmentation Tasks
by: Jaus, Alexander, et al.
Published: (2024)
by: Jaus, Alexander, et al.
Published: (2024)
Is Visual in-Context Learning for Compositional Medical Tasks within Reach?
by: Reiß, Simon, et al.
Published: (2025)
by: Reiß, Simon, et al.
Published: (2025)
Exploring Self-supervised Skeleton-based Action Recognition in Occluded Environments
by: Chen, Yifei, et al.
Published: (2023)
by: Chen, Yifei, et al.
Published: (2023)
RoDLA: Benchmarking the Robustness of Document Layout Analysis Models
by: Chen, Yufan, et al.
Published: (2024)
by: Chen, Yufan, et al.
Published: (2024)
HybriDLA: Hybrid Generation for Document Layout Analysis
by: Chen, Yufan, et al.
Published: (2025)
by: Chen, Yufan, et al.
Published: (2025)
Graph-based Document Structure Analysis
by: Chen, Yufan, et al.
Published: (2025)
by: Chen, Yufan, et al.
Published: (2025)
MICA: Multi-Agent Industrial Coordination Assistant
by: Wen, Di, et al.
Published: (2025)
by: Wen, Di, et al.
Published: (2025)
Towards Multi-Source Domain Generalization for Sleep Staging with Noisy Labels
by: Wang, Kening, et al.
Published: (2026)
by: Wang, Kening, et al.
Published: (2026)
Elevating Skeleton-Based Action Recognition with Efficient Multi-Modality Self-Supervision
by: Wei, Yiping, et al.
Published: (2023)
by: Wei, Yiping, et al.
Published: (2023)
Snap, Segment, Deploy: A Visual Data and Detection Pipeline for Wearable Industrial Assistants
by: Wen, Di, et al.
Published: (2025)
by: Wen, Di, et al.
Published: (2025)
Fourier Prompt Tuning for Modality-Incomplete Scene Segmentation
by: Liu, Ruiping, et al.
Published: (2024)
by: Liu, Ruiping, et al.
Published: (2024)
Skeleton-Based Human Action Recognition with Noisy Labels
by: Xu, Yi, et al.
Published: (2024)
by: Xu, Yi, et al.
Published: (2024)
IMPACT: A Dataset for Multi-Granularity Human Procedural Action Understanding in Industrial Assembly
by: Wen, Di, et al.
Published: (2026)
by: Wen, Di, et al.
Published: (2026)
Rethinking Video Human-Object Interaction: Set Prediction over Time for Unified Detection and Anticipation
by: Luo, Yuanhao, et al.
Published: (2026)
by: Luo, Yuanhao, et al.
Published: (2026)
SGR3 Model: Scene Graph Retrieval-Reasoning Model in 3D
by: Wang, Zirui, et al.
Published: (2026)
by: Wang, Zirui, et al.
Published: (2026)
Referring Atomic Video Action Recognition
by: Peng, Kunyu, et al.
Published: (2024)
by: Peng, Kunyu, et al.
Published: (2024)
GRASPing Anatomy to Improve Pathology Segmentation
by: Li, Keyi, et al.
Published: (2025)
by: Li, Keyi, et al.
Published: (2025)
Situat3DChange: Situated 3D Change Understanding Dataset for Multimodal Large Language Model
by: Liu, Ruiping, et al.
Published: (2025)
by: Liu, Ruiping, et al.
Published: (2025)
RoHOI: Robustness Benchmark for Human-Object Interaction Detection
by: Wen, Di, et al.
Published: (2025)
by: Wen, Di, et al.
Published: (2025)
Data Diet: Can Trimming PET/CT Datasets Enhance Lesion Segmentation?
by: Jaus, Alexander, et al.
Published: (2024)
by: Jaus, Alexander, et al.
Published: (2024)
Scene-agnostic Pose Regression for Visual Localization
by: Zheng, Junwei, et al.
Published: (2025)
by: Zheng, Junwei, et al.
Published: (2025)
Semantic Segmentation for Preoperative Planning in Transcatheter Aortic Valve Replacement
by: Zöllner, Cedric, et al.
Published: (2025)
by: Zöllner, Cedric, et al.
Published: (2025)
HopaDIFF: Holistic-Partial Aware Fourier Conditioned Diffusion for Referring Human Action Segmentation in Multi-Person Scenarios
by: Peng, Kunyu, et al.
Published: (2025)
by: Peng, Kunyu, et al.
Published: (2025)
EgoExoMem: Cross-View Memory Reasoning over Synchronized Egocentric and Exocentric Videos
by: Liu, Ruiping, et al.
Published: (2026)
by: Liu, Ruiping, et al.
Published: (2026)
RHO: Robust Holistic OSM-Based Metric Cross-View Geo-Localization
by: Zheng, Junwei, et al.
Published: (2026)
by: Zheng, Junwei, et al.
Published: (2026)
OneBEV: Using One Panoramic Image for Bird's-Eye-View Semantic Mapping
by: Wei, Jiale, et al.
Published: (2024)
by: Wei, Jiale, et al.
Published: (2024)
RefAtomNet++: Advancing Referring Atomic Video Action Recognition using Semantic Retrieval based Multi-Trajectory Mamba
by: Peng, Kunyu, et al.
Published: (2025)
by: Peng, Kunyu, et al.
Published: (2025)
Deformable Mamba for Wide Field of View Segmentation
by: Hu, Jie, et al.
Published: (2024)
by: Hu, Jie, et al.
Published: (2024)
Learning to Look Closer: A New Instance-Wise Loss for Small Cerebral Lesion Segmentation
by: Bouteille, Luc, et al.
Published: (2025)
by: Bouteille, Luc, et al.
Published: (2025)
More than the Sum: Panorama-Language Models for Adverse Omni-Scenes
by: Fan, Weijia, et al.
Published: (2026)
by: Fan, Weijia, et al.
Published: (2026)
Exploring Video-Based Driver Activity Recognition under Noisy Labels
by: Fan, Linjuan, et al.
Published: (2025)
by: Fan, Linjuan, et al.
Published: (2025)
What if? Emulative Simulation with World Models for Situated Reasoning
by: Liu, Ruiping, et al.
Published: (2026)
by: Liu, Ruiping, et al.
Published: (2026)
Similar Items
-
IMPACT-CYCLE: A Contract-Based Multi-Agent System for Claim-Level Supervisory Correction of Long-Video Semantic Memory
by: Kong, Weitong, et al.
Published: (2026) -
IMPACT-HOI: Supervisory Control for Onset-Anchored Partial HOI Event Construction
by: Zhang, Haoshen, et al.
Published: (2026) -
OmniFall: From Staged Through Synthetic to Wild, A Unified Multi-Domain Dataset for Robust Fall Detection
by: Schneider, David, et al.
Published: (2025) -
LIMIS: Towards Language-based Interactive Medical Image Segmentation
by: Heinemann, Lena, et al.
Published: (2024) -
Rethinking Annotator Simulation: Realistic Evaluation of Whole-Body PET Lesion Interactive Segmentation Methods
by: Marinov, Zdravko, et al.
Published: (2024)