Pix2Code: Learning to Compose Neural Visual Concepts as Programs
Fuente:
arXiv
Saved in:
| Main Authors: | Wüst, Antonia, Stammer, Wolfgang, Delfosse, Quentin, Dhami, Devendra Singh, Kersting, Kristian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
V-LoL: A Diagnostic Dataset for Visual Logical Learning
by: Helff, Lukas, et al.
Published: (2023)
by: Helff, Lukas, et al.
Published: (2023)
Synthesizing Visual Concepts as Vision-Language Programs
by: Wüst, Antonia, et al.
Published: (2025)
by: Wüst, Antonia, et al.
Published: (2025)
Learning Differentiable Logic Programs for Abstract Visual Reasoning
by: Shindo, Hikaru, et al.
Published: (2023)
by: Shindo, Hikaru, et al.
Published: (2023)
Neural Concept Binder
by: Stammer, Wolfgang, et al.
Published: (2024)
by: Stammer, Wolfgang, et al.
Published: (2024)
BlendRL: A Framework for Merging Symbolic and Neural Policy Learning
by: Shindo, Hikaru, et al.
Published: (2024)
by: Shindo, Hikaru, et al.
Published: (2024)
Object Centric Concept Bottlenecks
by: Steinmann, David, et al.
Published: (2025)
by: Steinmann, David, et al.
Published: (2025)
Boosting Object Representation Learning via Motion and Object Continuity
by: Delfosse, Quentin, et al.
Published: (2022)
by: Delfosse, Quentin, et al.
Published: (2022)
OCAtari: Object-Centric Atari 2600 Reinforcement Learning Environments
by: Delfosse, Quentin, et al.
Published: (2023)
by: Delfosse, Quentin, et al.
Published: (2023)
DeiSAM: Segment Anything with Deictic Prompting
by: Shindo, Hikaru, et al.
Published: (2024)
by: Shindo, Hikaru, et al.
Published: (2024)
EXPIL: Explanatory Predicate Invention for Learning in Games
by: Sha, Jingyuan, et al.
Published: (2024)
by: Sha, Jingyuan, et al.
Published: (2024)
CycliST: A Video Language Model Benchmark for Reasoning on Cyclical State Transitions
by: Kohaut, Simon, et al.
Published: (2025)
by: Kohaut, Simon, et al.
Published: (2025)
Bongard in Wonderland: Visual Puzzles that Still Make AI Go Mad?
by: Wüst, Antonia, et al.
Published: (2024)
by: Wüst, Antonia, et al.
Published: (2024)
Multimodal Crowd Counting with Pix2Pix GANs
by: Khan, Muhammad Asif, et al.
Published: (2024)
by: Khan, Muhammad Asif, et al.
Published: (2024)
SRU-Pix2Pix: A Fusion-Driven Generator Network for Medical Image Translation with Few-Shot Learning
by: Qiu, Xihe, et al.
Published: (2026)
by: Qiu, Xihe, et al.
Published: (2026)
Learning to Intervene on Concept Bottlenecks
by: Steinmann, David, et al.
Published: (2023)
by: Steinmann, David, et al.
Published: (2023)
GRAIL: Autonomous Concept Grounding for Neuro-Symbolic Reinforcement Learning
by: Shindo, Hikaru, et al.
Published: (2026)
by: Shindo, Hikaru, et al.
Published: (2026)
Interpretable end-to-end Neurosymbolic Reinforcement Learning agents
by: Grandien, Nils, et al.
Published: (2024)
by: Grandien, Nils, et al.
Published: (2024)
Composing Concepts from Images and Videos via Concept-prompt Binding
by: Kong, Xianghao, et al.
Published: (2025)
by: Kong, Xianghao, et al.
Published: (2025)
Pix2NPHM: Learning to Regress NPHM Reconstructions From a Single Image
by: Giebenhain, Simon, et al.
Published: (2025)
by: Giebenhain, Simon, et al.
Published: (2025)
Core Tokensets for Data-efficient Sequential Training of Transformers
by: Paul, Subarnaduti, et al.
Published: (2024)
by: Paul, Subarnaduti, et al.
Published: (2024)
Pix2Gif: Motion-Guided Diffusion for GIF Generation
by: Kandala, Hitesh, et al.
Published: (2024)
by: Kandala, Hitesh, et al.
Published: (2024)
Fodor and Pylyshyn's Legacy: Still No Human-like Systematic Compositionality in Neural Networks
by: Woydt, Tim, et al.
Published: (2025)
by: Woydt, Tim, et al.
Published: (2025)
Causal Explanations Over Time: Articulated Reasoning for Interactive Environments
by: Rödling, Sebastian, et al.
Published: (2025)
by: Rödling, Sebastian, et al.
Published: (2025)
SnapPix: Efficient-Coding--Inspired In-Sensor Compression for Edge Vision
by: Lin, Weikai, et al.
Published: (2025)
by: Lin, Weikai, et al.
Published: (2025)
Pixelis: Reasoning in Pixels, from Seeing to Acting
by: Zhou, Yunpeng
Published: (2026)
by: Zhou, Yunpeng
Published: (2026)
Pix2Cap-COCO: Advancing Visual Comprehension via Pixel-Level Captioning
by: You, Zuyao, et al.
Published: (2025)
by: You, Zuyao, et al.
Published: (2025)
STORM: Segment, Track, and Object Re-Localization from a Single Image
by: Deng, Yu, et al.
Published: (2025)
by: Deng, Yu, et al.
Published: (2025)
Learning to Infer Generative Template Programs for Visual Concepts
by: Jones, R. Kenny, et al.
Published: (2024)
by: Jones, R. Kenny, et al.
Published: (2024)
An Organism Starts with a Single Pix-Cell: A Neural Cellular Diffusion for High-Resolution Image Synthesis
by: Elbatel, Marawan, et al.
Published: (2024)
by: Elbatel, Marawan, et al.
Published: (2024)
Pix2Next: Leveraging Vision Foundation Models for RGB to NIR Image Translation
by: Jin, Youngwan, et al.
Published: (2024)
by: Jin, Youngwan, et al.
Published: (2024)
ART: Adaptive Relation Tuning for Generalized Relation Prediction
by: Sudhakaran, Gopika, et al.
Published: (2025)
by: Sudhakaran, Gopika, et al.
Published: (2025)
Rethinking Concept Bottleneck Models: From Pitfalls to Solutions
by: Tapli, Merve, et al.
Published: (2026)
by: Tapli, Merve, et al.
Published: (2026)
PixLore: A Dataset-driven Approach to Rich Image Captioning
by: Bonilla-Salvador, Diego, et al.
Published: (2023)
by: Bonilla-Salvador, Diego, et al.
Published: (2023)
LlavaGuard: An Open VLM-based Framework for Safeguarding Vision Datasets and Models
by: Helff, Lukas, et al.
Published: (2024)
by: Helff, Lukas, et al.
Published: (2024)
TurtleBench: A Visual Programming Benchmark in Turtle Geometry
by: Rismanchian, Sina, et al.
Published: (2024)
by: Rismanchian, Sina, et al.
Published: (2024)
Discovering Hidden Visual Concepts Beyond Linguistic Input in Infant Learning
by: Ke, Xueyi, et al.
Published: (2025)
by: Ke, Xueyi, et al.
Published: (2025)
ConceptScope: Characterizing Dataset Bias via Disentangled Visual Concepts
by: Choi, Jinho, et al.
Published: (2025)
by: Choi, Jinho, et al.
Published: (2025)
Interpretable Concept Bottlenecks to Align Reinforcement Learning Agents
by: Delfosse, Quentin, et al.
Published: (2024)
by: Delfosse, Quentin, et al.
Published: (2024)
HackAtari: Atari Learning Environments for Robust and Continual Reinforcement Learning
by: Delfosse, Quentin, et al.
Published: (2024)
by: Delfosse, Quentin, et al.
Published: (2024)
Visual Delta Generator with Large Multi-modal Models for Semi-supervised Composed Image Retrieval
by: Jang, Young Kyun, et al.
Published: (2024)
by: Jang, Young Kyun, et al.
Published: (2024)
Similar Items
-
V-LoL: A Diagnostic Dataset for Visual Logical Learning
by: Helff, Lukas, et al.
Published: (2023) -
Synthesizing Visual Concepts as Vision-Language Programs
by: Wüst, Antonia, et al.
Published: (2025) -
Learning Differentiable Logic Programs for Abstract Visual Reasoning
by: Shindo, Hikaru, et al.
Published: (2023) -
Neural Concept Binder
by: Stammer, Wolfgang, et al.
Published: (2024) -
BlendRL: A Framework for Merging Symbolic and Neural Policy Learning
by: Shindo, Hikaru, et al.
Published: (2024)