Saved in:
| Main Author: | Odem, Tom |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2511.01728 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Visually Interpretable Subtask Reasoning for Visual Question Answering
by: Cheng, Yu, et al.
Published: (2025)
by: Cheng, Yu, et al.
Published: (2025)
AbracADDbra: Touch-Guided Object Addition by Decoupling Placement and Editing Subtasks
by: Swami, Kunal, et al.
Published: (2026)
by: Swami, Kunal, et al.
Published: (2026)
Subtask-Aware Visual Reward Learning from Segmented Demonstrations
by: Kim, Changyeon, et al.
Published: (2025)
by: Kim, Changyeon, et al.
Published: (2025)
Backdoor Unlearning by Linear Task Decomposition
by: Abdelraheem, Amel, et al.
Published: (2025)
by: Abdelraheem, Amel, et al.
Published: (2025)
LifeCLEF Plant Identification Task 2014
by: Goeau, Herve, et al.
Published: (2025)
by: Goeau, Herve, et al.
Published: (2025)
LifeCLEF Plant Identification Task 2015
by: Goeau, Herve, et al.
Published: (2025)
by: Goeau, Herve, et al.
Published: (2025)
A Preliminary Exploration Towards General Image Restoration
by: Kong, Xiangtao, et al.
Published: (2024)
by: Kong, Xiangtao, et al.
Published: (2024)
Towards Lossless Implicit Neural Representation via Bit Plane Decomposition
by: Han, Woo Kyoung, et al.
Published: (2025)
by: Han, Woo Kyoung, et al.
Published: (2025)
Qwen-Image-Layered: Towards Inherent Editability via Layer Decomposition
by: Yin, Shengming, et al.
Published: (2025)
by: Yin, Shengming, et al.
Published: (2025)
Quantified Task Misalignment to Inform PEFT: An Exploration of Domain Generalization and Catastrophic Forgetting in CLIP
by: Niss, Laura, et al.
Published: (2024)
by: Niss, Laura, et al.
Published: (2024)
ODYSSEY: Open-World Quadrupeds Exploration and Manipulation for Long-Horizon Tasks
by: Wang, Kaijun, et al.
Published: (2025)
by: Wang, Kaijun, et al.
Published: (2025)
Toward a Holistic Evaluation of Robustness in CLIP Models
by: Tu, Weijie, et al.
Published: (2024)
by: Tu, Weijie, et al.
Published: (2024)
Improving Bird's Eye View Semantic Segmentation by Task Decomposition
by: Zhao, Tianhao, et al.
Published: (2024)
by: Zhao, Tianhao, et al.
Published: (2024)
Towards All-in-One Medical Image Re-Identification
by: Tian, Yuan, et al.
Published: (2025)
by: Tian, Yuan, et al.
Published: (2025)
Towards Reliable Identification of Diffusion-based Image Manipulations
by: Costanzino, Alex, et al.
Published: (2025)
by: Costanzino, Alex, et al.
Published: (2025)
Enhancing Abnormality Identification: Robust Out-of-Distribution Strategies for Deepfake Detection
by: Maiano, Luca, et al.
Published: (2025)
by: Maiano, Luca, et al.
Published: (2025)
Learning to See and Act: Task-Aware Virtual View Exploration for Robotic Manipulation
by: Bai, Yongjie, et al.
Published: (2025)
by: Bai, Yongjie, et al.
Published: (2025)
Task-Driven Exploration: Decoupling and Inter-Task Feedback for Joint Moment Retrieval and Highlight Detection
by: Yang, Jin, et al.
Published: (2024)
by: Yang, Jin, et al.
Published: (2024)
MagicWorld: Towards Long-Horizon Stability for Interactive Video World Exploration
by: Li, Guangyuan, et al.
Published: (2025)
by: Li, Guangyuan, et al.
Published: (2025)
QDFormer: Towards Robust Audiovisual Segmentation in Complex Environments with Quantization-based Semantic Decomposition
by: Li, Xiang, et al.
Published: (2023)
by: Li, Xiang, et al.
Published: (2023)
MASIV: Toward Material-Agnostic System Identification from Videos
by: Zhao, Yizhou, et al.
Published: (2025)
by: Zhao, Yizhou, et al.
Published: (2025)
An Independent Discriminant Network Towards Identification of Counterfeit Images and Videos
by: Kar, Shayantani, et al.
Published: (2025)
by: Kar, Shayantani, et al.
Published: (2025)
Bridging Data Trials and Task Barriers: A Unified Framework for Sketch Biometric Identification
by: Liu, Decheng, et al.
Published: (2026)
by: Liu, Decheng, et al.
Published: (2026)
Decompose and Compare Consistency: Measuring VLMs' Answer Reliability via Task-Decomposition Consistency Comparison
by: Yang, Qian, et al.
Published: (2024)
by: Yang, Qian, et al.
Published: (2024)
From Visual to Multimodal: Systematic Ablation of Encoders and Fusion Strategies in Animal Identification
by: Kudryavtsev, Vasiliy, et al.
Published: (2026)
by: Kudryavtsev, Vasiliy, et al.
Published: (2026)
Does the Data Processing Inequality Reflect Practice? On the Utility of Low-Level Tasks
by: Turgeman, Roy, et al.
Published: (2025)
by: Turgeman, Roy, et al.
Published: (2025)
STiL: Semi-supervised Tabular-Image Learning for Comprehensive Task-Relevant Information Exploration in Multimodal Classification
by: Du, Siyi, et al.
Published: (2025)
by: Du, Siyi, et al.
Published: (2025)
Towards Robust Sequential Decomposition for Complex Image Editing
by: Zeng, Zilai, et al.
Published: (2026)
by: Zeng, Zilai, et al.
Published: (2026)
InstaScene: Towards Complete 3D Instance Decomposition and Reconstruction from Cluttered Scenes
by: Yang, Zesong, et al.
Published: (2025)
by: Yang, Zesong, et al.
Published: (2025)
DeMoGen: Towards Decompositional Human Motion Generation with Energy-Based Diffusion Models
by: Zhang, Jianrong, et al.
Published: (2025)
by: Zhang, Jianrong, et al.
Published: (2025)
NoiseController: Towards Consistent Multi-view Video Generation via Noise Decomposition and Collaboration
by: Dong, Haotian, et al.
Published: (2025)
by: Dong, Haotian, et al.
Published: (2025)
Towards Anytime Retrieval: A Benchmark for Anytime Person Re-Identification
by: Li, Xulin, et al.
Published: (2025)
by: Li, Xulin, et al.
Published: (2025)
MULTIFLOW: Shifting Towards Task-Agnostic Vision-Language Pruning
by: Farina, Matteo, et al.
Published: (2024)
by: Farina, Matteo, et al.
Published: (2024)
Toward a Diffusion-Based Generalist for Dense Vision Tasks
by: Fan, Yue, et al.
Published: (2024)
by: Fan, Yue, et al.
Published: (2024)
SCOUT+: Towards Practical Task-Driven Drivers' Gaze Prediction
by: Kotseruba, Iuliia, et al.
Published: (2024)
by: Kotseruba, Iuliia, et al.
Published: (2024)
SPAgent: Adaptive Task Decomposition and Model Selection for General Video Generation and Editing
by: Tu, Rong-Cheng, et al.
Published: (2024)
by: Tu, Rong-Cheng, et al.
Published: (2024)
Finding Waldo: Towards Efficient Exploration of NeRF Scene Spaces
by: Skartados, Evangelos, et al.
Published: (2024)
by: Skartados, Evangelos, et al.
Published: (2024)
UniD-Shift: Towards Unified Semantic Segmentation via Interpretable Share-Private Multimodal Decomposition
by: Zhang, Shuai, et al.
Published: (2026)
by: Zhang, Shuai, et al.
Published: (2026)
AERR-Nav: Adaptive Exploration-Recovery-Reminiscing Strategy for Zero-Shot Object Navigation
by: Huang, Jingzhi, et al.
Published: (2026)
by: Huang, Jingzhi, et al.
Published: (2026)
Towards Continuous Home Cage Monitoring: An Evaluation of Tracking and Identification Strategies for Laboratory Mice
by: Oberhauser, Juan Pablo, et al.
Published: (2025)
by: Oberhauser, Juan Pablo, et al.
Published: (2025)
Similar Items
-
Visually Interpretable Subtask Reasoning for Visual Question Answering
by: Cheng, Yu, et al.
Published: (2025) -
AbracADDbra: Touch-Guided Object Addition by Decoupling Placement and Editing Subtasks
by: Swami, Kunal, et al.
Published: (2026) -
Subtask-Aware Visual Reward Learning from Segmented Demonstrations
by: Kim, Changyeon, et al.
Published: (2025) -
Backdoor Unlearning by Linear Task Decomposition
by: Abdelraheem, Amel, et al.
Published: (2025) -
LifeCLEF Plant Identification Task 2014
by: Goeau, Herve, et al.
Published: (2025)