Deep Neural Networks Can Learn Generalizable Same-Different Visual Relations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tartaglini, Alexa R., Feucht, Sheridan, Lepori, Michael A., Vong, Wai Keen, Lovering, Charles, Lake, Brenden M., Pavlick, Ellie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Beyond the Doors of Perception: Vision Transformers Represent Relations Between Objects
von: Lepori, Michael A., et al.
Veröffentlicht: (2024)
von: Lepori, Michael A., et al.
Veröffentlicht: (2024)
On the robustness of modeling grounded word learning through a child's egocentric input
von: Vong, Wai Keen, et al.
Veröffentlicht: (2025)
von: Vong, Wai Keen, et al.
Veröffentlicht: (2025)
H-ARC: A Robust Estimate of Human Performance on the Abstraction and Reasoning Corpus Benchmark
von: LeGris, Solim, et al.
Veröffentlicht: (2024)
von: LeGris, Solim, et al.
Veröffentlicht: (2024)
I Walk the Line: Examining the Role of Gestalt Continuity in Object Binding for Vision Transformers
von: Tartaglini, Alexa R., et al.
Veröffentlicht: (2026)
von: Tartaglini, Alexa R., et al.
Veröffentlicht: (2026)
Instilling Inductive Biases with Subnetworks
von: Zhang, Enyan, et al.
Veröffentlicht: (2023)
von: Zhang, Enyan, et al.
Veröffentlicht: (2023)
Uncovering Intermediate Variables in Transformers using Circuit Probing
von: Lepori, Michael A., et al.
Veröffentlicht: (2023)
von: Lepori, Michael A., et al.
Veröffentlicht: (2023)
The Same But Different: Structural Similarities and Differences in Multilingual Language Modeling
von: Zhang, Ruochen, et al.
Veröffentlicht: (2024)
von: Zhang, Ruochen, et al.
Veröffentlicht: (2024)
Dual Process Learning: Controlling Use of In-Context vs. In-Weights Strategies with Weight Forgetting
von: Anand, Suraj, et al.
Veröffentlicht: (2024)
von: Anand, Suraj, et al.
Veröffentlicht: (2024)
Convolutional Neural Networks Can (Meta-)Learn the Same-Different Relation
von: Gupta, Max, et al.
Veröffentlicht: (2025)
von: Gupta, Max, et al.
Veröffentlicht: (2025)
Whither symbols in the era of advanced neural networks?
von: Griffiths, Thomas L., et al.
Veröffentlicht: (2025)
von: Griffiths, Thomas L., et al.
Veröffentlicht: (2025)
Are LLMs Models of Distributional Semantics? A Case Study on Quantifiers
von: Enyan, Zhang, et al.
Veröffentlicht: (2024)
von: Enyan, Zhang, et al.
Veröffentlicht: (2024)
From Prediction to Understanding: Will AI Foundation Models Transform Brain Science?
von: Serre, Thomas, et al.
Veröffentlicht: (2025)
von: Serre, Thomas, et al.
Veröffentlicht: (2025)
Does Training on Synthetic Data Make Models Less Robust?
von: Zhang, Lingze, et al.
Veröffentlicht: (2025)
von: Zhang, Lingze, et al.
Veröffentlicht: (2025)
How Do Language Models Compose Functions?
von: Khandelwal, Apoorv, et al.
Veröffentlicht: (2025)
von: Khandelwal, Apoorv, et al.
Veröffentlicht: (2025)
Is This Just Fantasy? Language Model Representations Reflect Human Judgments of Event Plausibility
von: Lepori, Michael A., et al.
Veröffentlicht: (2025)
von: Lepori, Michael A., et al.
Veröffentlicht: (2025)
Vector Arithmetic in Concept and Token Subspaces
von: Feucht, Sheridan, et al.
Veröffentlicht: (2025)
von: Feucht, Sheridan, et al.
Veröffentlicht: (2025)
The dynamic interplay between in-context and in-weight learning in humans and neural networks
von: Russin, Jacob, et al.
Veröffentlicht: (2024)
von: Russin, Jacob, et al.
Veröffentlicht: (2024)
LLMs model how humans induce logically structured rules
von: Loo, Alyssa, et al.
Veröffentlicht: (2025)
von: Loo, Alyssa, et al.
Veröffentlicht: (2025)
mOthello: When Do Cross-Lingual Representation Alignment and Cross-Lingual Transfer Emerge in Multilingual Models?
von: Hua, Tianze, et al.
Veröffentlicht: (2024)
von: Hua, Tianze, et al.
Veröffentlicht: (2024)
What is an "Abstract Reasoner"? Revisiting Experiments and Arguments about Large Language Models
von: Yun, Tian, et al.
Veröffentlicht: (2025)
von: Yun, Tian, et al.
Veröffentlicht: (2025)
Handling and Interpreting Missing Modalities in Patient Clinical Trajectories via Autoregressive Sequence Modeling
von: Wang, Andrew, et al.
Veröffentlicht: (2026)
von: Wang, Andrew, et al.
Veröffentlicht: (2026)
Talking Heads: Understanding Inter-layer Communication in Transformer Language Models
von: Merullo, Jack, et al.
Veröffentlicht: (2024)
von: Merullo, Jack, et al.
Veröffentlicht: (2024)
How Do Vision-Language Models Process Conflicting Information Across Modalities?
von: Hua, Tianze, et al.
Veröffentlicht: (2025)
von: Hua, Tianze, et al.
Veröffentlicht: (2025)
A Knapsack by Any Other Name: Presentation impacts LLM performance on NP-hard problems
von: Duchnowski, Alex, et al.
Veröffentlicht: (2025)
von: Duchnowski, Alex, et al.
Veröffentlicht: (2025)
Circuit Component Reuse Across Tasks in Transformer Language Models
von: Merullo, Jack, et al.
Veröffentlicht: (2023)
von: Merullo, Jack, et al.
Veröffentlicht: (2023)
Language Models Implement Simple Word2Vec-style Vector Arithmetic
von: Merullo, Jack, et al.
Veröffentlicht: (2023)
von: Merullo, Jack, et al.
Veröffentlicht: (2023)
Overcoming classic challenges for artificial neural networks by providing incentives and practice
von: Irie, Kazuki, et al.
Veröffentlicht: (2024)
von: Irie, Kazuki, et al.
Veröffentlicht: (2024)
Are they human? Detecting large language models by probing human memory constraints
von: Schug, Simon, et al.
Veröffentlicht: (2026)
von: Schug, Simon, et al.
Veröffentlicht: (2026)
Not How Many, But Which: Parameter Placement in Low-Rank Adaptation
von: Sehanobish, Arijit, et al.
Veröffentlicht: (2026)
von: Sehanobish, Arijit, et al.
Veröffentlicht: (2026)
Diagnosing Bottlenecks in Data Visualization Understanding by Vision-Language Models
von: Tartaglini, Alexa R., et al.
Veröffentlicht: (2025)
von: Tartaglini, Alexa R., et al.
Veröffentlicht: (2025)
Source-Modality Monitoring in Vision-Language Models
von: Hua, Etha Tianze, et al.
Veröffentlicht: (2026)
von: Hua, Etha Tianze, et al.
Veröffentlicht: (2026)
The Dual-Route Model of Induction
von: Feucht, Sheridan, et al.
Veröffentlicht: (2025)
von: Feucht, Sheridan, et al.
Veröffentlicht: (2025)
Erasing Conceptual Knowledge from Language Models
von: Gandikota, Rohit, et al.
Veröffentlicht: (2024)
von: Gandikota, Rohit, et al.
Veröffentlicht: (2024)
Token Erasure as a Footprint of Implicit Vocabulary Items in LLMs
von: Feucht, Sheridan, et al.
Veröffentlicht: (2024)
von: Feucht, Sheridan, et al.
Veröffentlicht: (2024)
Transformer Mechanisms Mimic Frontostriatal Gating Operations When Trained on Human Working Memory Tasks
von: Traylor, Aaron, et al.
Veröffentlicht: (2024)
von: Traylor, Aaron, et al.
Veröffentlicht: (2024)
An explainable transformer circuit for compositional generalization
von: Tang, Cheng, et al.
Veröffentlicht: (2025)
von: Tang, Cheng, et al.
Veröffentlicht: (2025)
CoLLEGe: Concept Embedding Generation for Large Language Models
von: Teehan, Ryan, et al.
Veröffentlicht: (2024)
von: Teehan, Ryan, et al.
Veröffentlicht: (2024)
Video Finetuning Improves Reasoning Between Frames
von: Yang, Ruiqi, et al.
Veröffentlicht: (2025)
von: Yang, Ruiqi, et al.
Veröffentlicht: (2025)
Transferring Linear Features Across Language Models With Model Stitching
von: Chen, Alan, et al.
Veröffentlicht: (2025)
von: Chen, Alan, et al.
Veröffentlicht: (2025)
LLMs as Models for Analogical Reasoning
von: Musker, Sam, et al.
Veröffentlicht: (2024)
von: Musker, Sam, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Beyond the Doors of Perception: Vision Transformers Represent Relations Between Objects
von: Lepori, Michael A., et al.
Veröffentlicht: (2024) -
On the robustness of modeling grounded word learning through a child's egocentric input
von: Vong, Wai Keen, et al.
Veröffentlicht: (2025) -
H-ARC: A Robust Estimate of Human Performance on the Abstraction and Reasoning Corpus Benchmark
von: LeGris, Solim, et al.
Veröffentlicht: (2024) -
I Walk the Line: Examining the Role of Gestalt Continuity in Object Binding for Vision Transformers
von: Tartaglini, Alexa R., et al.
Veröffentlicht: (2026) -
Instilling Inductive Biases with Subnetworks
von: Zhang, Enyan, et al.
Veröffentlicht: (2023)