Revisiting Disentanglement in Downstream Tasks: A Study on Its Necessity for Abstract Visual Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Nai, Ruiqian, Wen, Zixin, Li, Ji, Li, Yuanzhi, Gao, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Advancing Generalization Across a Variety of Abstract Visual Reasoning Tasks
by: Małkiński, Mikołaj, et al.
Published: (2025)
by: Małkiński, Mikołaj, et al.
Published: (2025)
Slot Abstractors: Toward Scalable Abstract Visual Reasoning
by: Mondal, Shanka Subhra, et al.
Published: (2024)
by: Mondal, Shanka Subhra, et al.
Published: (2024)
Transferable Adversarial Attacks on SAM and Its Downstream Models
by: Xia, Song, et al.
Published: (2024)
by: Xia, Song, et al.
Published: (2024)
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks
by: Ding, Yuhe, et al.
Published: (2024)
by: Ding, Yuhe, et al.
Published: (2024)
Synthesizer Based Efficient Self-Attention for Vision Tasks
by: Zhu, Guangyang, et al.
Published: (2022)
by: Zhu, Guangyang, et al.
Published: (2022)
Beyond Visual Fidelity: Benchmarking Super-Resolution Models for Large-Scale Remote Sensing Imagery via Downstream Task Integration
by: Li, Zhili, et al.
Published: (2026)
by: Li, Zhili, et al.
Published: (2026)
Downstream Task Guided Masking Learning in Masked Autoencoders Using Multi-Level Optimization
by: Guo, Han, et al.
Published: (2024)
by: Guo, Han, et al.
Published: (2024)
A Unified View of Abstract Visual Reasoning Problems
by: Małkiński, Mikołaj, et al.
Published: (2024)
by: Małkiński, Mikołaj, et al.
Published: (2024)
Multi-Label Contrastive Learning for Abstract Visual Reasoning
by: Małkiński, Mikołaj, et al.
Published: (2020)
by: Małkiński, Mikołaj, et al.
Published: (2020)
Learning Differentiable Logic Programs for Abstract Visual Reasoning
by: Shindo, Hikaru, et al.
Published: (2023)
by: Shindo, Hikaru, et al.
Published: (2023)
VisReason: A Large-Scale Dataset for Visual Chain-of-Thought Reasoning
by: Li, Lingxiao, et al.
Published: (2025)
by: Li, Lingxiao, et al.
Published: (2025)
Multi-instance Learning as Downstream Task of Self-Supervised Learning-based Pre-trained Model
by: Matsuishi, Koki, et al.
Published: (2025)
by: Matsuishi, Koki, et al.
Published: (2025)
Stable Diffusion Dataset Generation for Downstream Classification Tasks
by: Lomurno, Eugenio, et al.
Published: (2024)
by: Lomurno, Eugenio, et al.
Published: (2024)
Multi-Task Learning for Visually Grounded Reasoning in Gastrointestinal VQA
by: Safwan, Itbaan, et al.
Published: (2025)
by: Safwan, Itbaan, et al.
Published: (2025)
Learning Visual Abstract Reasoning through Dual-Stream Networks
by: Zhao, Kai, et al.
Published: (2024)
by: Zhao, Kai, et al.
Published: (2024)
Understanding and Mitigating the Label Noise in Pre-training on Downstream Tasks
by: Chen, Hao, et al.
Published: (2023)
by: Chen, Hao, et al.
Published: (2023)
Look-Back: Implicit Visual Re-focusing in MLLM Reasoning
by: Yang, Shuo, et al.
Published: (2025)
by: Yang, Shuo, et al.
Published: (2025)
On the Out-of-Distribution Generalization of Reasoning in Multimodal LLMs for Simple Visual Planning Tasks
by: Neuhaus, Yannic, et al.
Published: (2026)
by: Neuhaus, Yannic, et al.
Published: (2026)
Role of Locality and Weight Sharing in Image-Based Tasks: A Sample Complexity Separation between CNNs, LCNs, and FCNs
by: Lahoti, Aakash, et al.
Published: (2024)
by: Lahoti, Aakash, et al.
Published: (2024)
Revisiting the Power of Prompt for Visual Tuning
by: Wang, Yuzhu, et al.
Published: (2024)
by: Wang, Yuzhu, et al.
Published: (2024)
One Self-Configurable Model to Solve Many Abstract Visual Reasoning Problems
by: Małkiński, Mikołaj, et al.
Published: (2023)
by: Małkiński, Mikołaj, et al.
Published: (2023)
Prompt Disentanglement via Language Guidance and Representation Alignment for Domain Generalization
by: Cheng, De, et al.
Published: (2025)
by: Cheng, De, et al.
Published: (2025)
Logo-VGR: Visual Grounded Reasoning for Open-world Logo Recognition
by: Liang, Zichen, et al.
Published: (2025)
by: Liang, Zichen, et al.
Published: (2025)
DRESS: Disentangled Representation-based Self-Supervised Meta-Learning for Diverse Tasks
by: Cui, Wei, et al.
Published: (2025)
by: Cui, Wei, et al.
Published: (2025)
Johnny: Structuring Representation Space to Enhance Machine Abstract Reasoning Ability
by: Song, Ruizhuo, et al.
Published: (2025)
by: Song, Ruizhuo, et al.
Published: (2025)
Aggregate-and-Adapt Natural Language Prompts for Downstream Generalization of CLIP
by: Huang, Chen, et al.
Published: (2024)
by: Huang, Chen, et al.
Published: (2024)
Quantum-Gated Task-interaction Knowledge Distillation for Pre-trained Model-based Class-Incremental Learning
by: Li, Linjie, et al.
Published: (2026)
by: Li, Linjie, et al.
Published: (2026)
ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs
by: Wang, Xiyao, et al.
Published: (2025)
by: Wang, Xiyao, et al.
Published: (2025)
Revisit Visual Prompt Tuning: The Expressiveness of Prompt Experts
by: Le, Minh, et al.
Published: (2025)
by: Le, Minh, et al.
Published: (2025)
A-I-RAVEN and I-RAVEN-Mesh: Two New Benchmarks for Abstract Visual Reasoning
by: Małkiński, Mikołaj, et al.
Published: (2024)
by: Małkiński, Mikołaj, et al.
Published: (2024)
DeepLatent: Think with Images via Parallel Latent Visual Reasoning
by: Lu, Dongchen, et al.
Published: (2026)
by: Lu, Dongchen, et al.
Published: (2026)
Multi-Task Model Merging via Adaptive Weight Disentanglement
by: Xiong, Feng, et al.
Published: (2024)
by: Xiong, Feng, et al.
Published: (2024)
Closed-Loop Unsupervised Representation Disentanglement with $β$-VAE Distillation and Diffusion Probabilistic Feedback
by: Jin, Xin, et al.
Published: (2024)
by: Jin, Xin, et al.
Published: (2024)
OmniPrism: Learning Disentangled Visual Concept for Image Generation
by: Li, Yangyang, et al.
Published: (2024)
by: Li, Yangyang, et al.
Published: (2024)
Masked Autoencoders for Ultrasound Signals: Robust Representation Learning for Downstream Applications
by: Roßteutscher, Immanuel, et al.
Published: (2025)
by: Roßteutscher, Immanuel, et al.
Published: (2025)
ComplicitSplat: Downstream Models are Vulnerable to Blackbox Attacks by 3D Gaussian Splat Camouflages
by: Hull, Matthew, et al.
Published: (2025)
by: Hull, Matthew, et al.
Published: (2025)
Merging and Disentangling Views in Visual Reinforcement Learning for Robotic Manipulation
by: Almuzairee, Abdulaziz, et al.
Published: (2025)
by: Almuzairee, Abdulaziz, et al.
Published: (2025)
Imagine while Reasoning in Space: Multimodal Visualization-of-Thought
by: Li, Chengzu, et al.
Published: (2025)
by: Li, Chengzu, et al.
Published: (2025)
ResAD: A Simple Framework for Class Generalizable Anomaly Detection
by: Yao, Xincheng, et al.
Published: (2024)
by: Yao, Xincheng, et al.
Published: (2024)
Better than Average: Spatially-Aware Aggregation of Segmentation Uncertainty Improves Downstream Performance
by: Guarino, Vanessa Emanuela, et al.
Published: (2026)
by: Guarino, Vanessa Emanuela, et al.
Published: (2026)
Similar Items
-
Advancing Generalization Across a Variety of Abstract Visual Reasoning Tasks
by: Małkiński, Mikołaj, et al.
Published: (2025) -
Slot Abstractors: Toward Scalable Abstract Visual Reasoning
by: Mondal, Shanka Subhra, et al.
Published: (2024) -
Transferable Adversarial Attacks on SAM and Its Downstream Models
by: Xia, Song, et al.
Published: (2024) -
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks
by: Ding, Yuhe, et al.
Published: (2024) -
Synthesizer Based Efficient Self-Attention for Vision Tasks
by: Zhu, Guangyang, et al.
Published: (2022)