ActivationReasoning: Logical Reasoning in Latent Activation Spaces
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Helff, Lukas, Härle, Ruben, Stammer, Wolfgang, Friedrich, Felix, Brack, Manuel, Wüst, Antonia, Shindo, Hikaru, Schramowski, Patrick, Kersting, Kristian |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SLR: Automated Synthesis for Scalable Logical Reasoning
von: Helff, Lukas, et al.
Veröffentlicht: (2025)
von: Helff, Lukas, et al.
Veröffentlicht: (2025)
LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking
von: Helff, Lukas, et al.
Veröffentlicht: (2026)
von: Helff, Lukas, et al.
Veröffentlicht: (2026)
SocialGrid: A Benchmark for Planning and Social Reasoning in Embodied Multi-Agent Systems
von: Shindo, Hikaru, et al.
Veröffentlicht: (2026)
von: Shindo, Hikaru, et al.
Veröffentlicht: (2026)
Synthesizing Visual Concepts as Vision-Language Programs
von: Wüst, Antonia, et al.
Veröffentlicht: (2025)
von: Wüst, Antonia, et al.
Veröffentlicht: (2025)
V-LoL: A Diagnostic Dataset for Visual Logical Learning
von: Helff, Lukas, et al.
Veröffentlicht: (2023)
von: Helff, Lukas, et al.
Veröffentlicht: (2023)
LlavaGuard: An Open VLM-based Framework for Safeguarding Vision Datasets and Models
von: Helff, Lukas, et al.
Veröffentlicht: (2024)
von: Helff, Lukas, et al.
Veröffentlicht: (2024)
Learning by Self-Explaining
von: Stammer, Wolfgang, et al.
Veröffentlicht: (2023)
von: Stammer, Wolfgang, et al.
Veröffentlicht: (2023)
A Typology for Exploring the Mitigation of Shortcut Behavior
von: Friedrich, Felix, et al.
Veröffentlicht: (2022)
von: Friedrich, Felix, et al.
Veröffentlicht: (2022)
SCAR: Sparse Conditioned Autoencoders for Concept Detection and Steering in LLMs
von: Härle, Ruben, et al.
Veröffentlicht: (2024)
von: Härle, Ruben, et al.
Veröffentlicht: (2024)
DeiSAM: Segment Anything with Deictic Prompting
von: Shindo, Hikaru, et al.
Veröffentlicht: (2024)
von: Shindo, Hikaru, et al.
Veröffentlicht: (2024)
Fodor and Pylyshyn's Legacy: Still No Human-like Systematic Compositionality in Neural Networks
von: Woydt, Tim, et al.
Veröffentlicht: (2025)
von: Woydt, Tim, et al.
Veröffentlicht: (2025)
Object Centric Concept Bottlenecks
von: Steinmann, David, et al.
Veröffentlicht: (2025)
von: Steinmann, David, et al.
Veröffentlicht: (2025)
Neural Concept Binder
von: Stammer, Wolfgang, et al.
Veröffentlicht: (2024)
von: Stammer, Wolfgang, et al.
Veröffentlicht: (2024)
Learning Differentiable Logic Programs for Abstract Visual Reasoning
von: Shindo, Hikaru, et al.
Veröffentlicht: (2023)
von: Shindo, Hikaru, et al.
Veröffentlicht: (2023)
Measuring and Guiding Monosemanticity
von: Härle, Ruben, et al.
Veröffentlicht: (2025)
von: Härle, Ruben, et al.
Veröffentlicht: (2025)
Exploiting Cultural Biases via Homoglyphs in Text-to-Image Synthesis
von: Struppek, Lukas, et al.
Veröffentlicht: (2022)
von: Struppek, Lukas, et al.
Veröffentlicht: (2022)
Bongard in Wonderland: Visual Puzzles that Still Make AI Go Mad?
von: Wüst, Antonia, et al.
Veröffentlicht: (2024)
von: Wüst, Antonia, et al.
Veröffentlicht: (2024)
No Safe Dose: How Training Data Drives Unsafe Image Generation
von: Friedrich, Felix, et al.
Veröffentlicht: (2026)
von: Friedrich, Felix, et al.
Veröffentlicht: (2026)
Core Tokensets for Data-efficient Sequential Training of Transformers
von: Paul, Subarnaduti, et al.
Veröffentlicht: (2024)
von: Paul, Subarnaduti, et al.
Veröffentlicht: (2024)
T-FREE: Subword Tokenizer-Free Generative LLMs via Sparse Representations for Memory-Efficient Embeddings
von: Deiseroth, Björn, et al.
Veröffentlicht: (2024)
von: Deiseroth, Björn, et al.
Veröffentlicht: (2024)
Pix2Code: Learning to Compose Neural Visual Concepts as Programs
von: Wüst, Antonia, et al.
Veröffentlicht: (2024)
von: Wüst, Antonia, et al.
Veröffentlicht: (2024)
ART: Adaptive Relation Tuning for Generalized Relation Prediction
von: Sudhakaran, Gopika, et al.
Veröffentlicht: (2025)
von: Sudhakaran, Gopika, et al.
Veröffentlicht: (2025)
How to Train your Text-to-Image Model: Evaluating Design Choices for Synthetic Training Captions
von: Brack, Manuel, et al.
Veröffentlicht: (2025)
von: Brack, Manuel, et al.
Veröffentlicht: (2025)
Does CLIP Know My Face?
von: Hintersdorf, Dominik, et al.
Veröffentlicht: (2022)
von: Hintersdorf, Dominik, et al.
Veröffentlicht: (2022)
LEDITS++: Limitless Image Editing using Text-to-Image Models
von: Brack, Manuel, et al.
Veröffentlicht: (2023)
von: Brack, Manuel, et al.
Veröffentlicht: (2023)
Learning to Intervene on Concept Bottlenecks
von: Steinmann, David, et al.
Veröffentlicht: (2023)
von: Steinmann, David, et al.
Veröffentlicht: (2023)
AtMan: Understanding Transformer Predictions Through Memory Efficient Attention Manipulation
von: Deiseroth, Björn, et al.
Veröffentlicht: (2023)
von: Deiseroth, Björn, et al.
Veröffentlicht: (2023)
Beyond Overcorrection: Evaluating Diversity in T2I Models with DivBench
von: Friedrich, Felix, et al.
Veröffentlicht: (2025)
von: Friedrich, Felix, et al.
Veröffentlicht: (2025)
LIME: Making LLM Data More Efficient with Linguistic Metadata Embeddings
von: Sztwiertnia, Sebastian, et al.
Veröffentlicht: (2025)
von: Sztwiertnia, Sebastian, et al.
Veröffentlicht: (2025)
Navigating Shortcuts, Spurious Correlations, and Confounders: From Origins via Detection to Mitigation
von: Steinmann, David, et al.
Veröffentlicht: (2024)
von: Steinmann, David, et al.
Veröffentlicht: (2024)
GRAIL: Autonomous Concept Grounding for Neuro-Symbolic Reinforcement Learning
von: Shindo, Hikaru, et al.
Veröffentlicht: (2026)
von: Shindo, Hikaru, et al.
Veröffentlicht: (2026)
CHRONOBERG: Capturing Language Evolution and Temporal Awareness in Foundation Models
von: Hegde, Niharika, et al.
Veröffentlicht: (2025)
von: Hegde, Niharika, et al.
Veröffentlicht: (2025)
BlendRL: A Framework for Merging Symbolic and Neural Policy Learning
von: Shindo, Hikaru, et al.
Veröffentlicht: (2024)
von: Shindo, Hikaru, et al.
Veröffentlicht: (2024)
Multilingual Text-to-Image Generation Magnifies Gender Stereotypes and Prompt Engineering May Not Help You
von: Friedrich, Felix, et al.
Veröffentlicht: (2024)
von: Friedrich, Felix, et al.
Veröffentlicht: (2024)
Adaptive Rational Activations to Boost Deep Reinforcement Learning
von: Delfosse, Quentin, et al.
Veröffentlicht: (2021)
von: Delfosse, Quentin, et al.
Veröffentlicht: (2021)
Making deep neural networks right for the right scientific reasons by interacting with their explanations
von: Schramowski, Patrick, et al.
Veröffentlicht: (2020)
von: Schramowski, Patrick, et al.
Veröffentlicht: (2020)
EXPIL: Explanatory Predicate Invention for Learning in Games
von: Sha, Jingyuan, et al.
Veröffentlicht: (2024)
von: Sha, Jingyuan, et al.
Veröffentlicht: (2024)
LLMs Lost in Translation: M-ALERT uncovers Cross-Linguistic Safety Inconsistencies
von: Friedrich, Felix, et al.
Veröffentlicht: (2024)
von: Friedrich, Felix, et al.
Veröffentlicht: (2024)
Right in Time: Reactive Reasoning in Regulated Traffic Spaces
von: Kohaut, Simon, et al.
Veröffentlicht: (2026)
von: Kohaut, Simon, et al.
Veröffentlicht: (2026)
Depth-Recurrent Attention Mixtures: Giving Latent Reasoning the Attention it Deserves
von: Knupp, Jonas, et al.
Veröffentlicht: (2026)
von: Knupp, Jonas, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
SLR: Automated Synthesis for Scalable Logical Reasoning
von: Helff, Lukas, et al.
Veröffentlicht: (2025) -
LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking
von: Helff, Lukas, et al.
Veröffentlicht: (2026) -
SocialGrid: A Benchmark for Planning and Social Reasoning in Embodied Multi-Agent Systems
von: Shindo, Hikaru, et al.
Veröffentlicht: (2026) -
Synthesizing Visual Concepts as Vision-Language Programs
von: Wüst, Antonia, et al.
Veröffentlicht: (2025) -
V-LoL: A Diagnostic Dataset for Visual Logical Learning
von: Helff, Lukas, et al.
Veröffentlicht: (2023)