Saved in:
| Main Authors: | Wattenberg, Martin, Viégas, Fernanda B. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2407.14662 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dialogue Action Tokens: Steering Language Models in Goal-Directed Dialogue with a Multi-Turn Planner
by: Li, Kenneth, et al.
Published: (2024)
by: Li, Kenneth, et al.
Published: (2024)
When Bad Data Leads to Good Models
by: Li, Kenneth, et al.
Published: (2025)
by: Li, Kenneth, et al.
Published: (2025)
The Geometry of Self-Verification in a Task-Specific Reasoning Model
by: Lee, Andrew, et al.
Published: (2025)
by: Lee, Andrew, et al.
Published: (2025)
Inference-Time Intervention: Eliciting Truthful Answers from a Language Model
by: Li, Kenneth, et al.
Published: (2023)
by: Li, Kenneth, et al.
Published: (2023)
What Does it Mean for a Neural Network to Learn a "World Model"?
by: Li, Kenneth, et al.
Published: (2025)
by: Li, Kenneth, et al.
Published: (2025)
Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task
by: Li, Kenneth, et al.
Published: (2022)
by: Li, Kenneth, et al.
Published: (2022)
Measuring and Controlling Instruction (In)Stability in Language Model Dialogs
by: Li, Kenneth, et al.
Published: (2024)
by: Li, Kenneth, et al.
Published: (2024)
Why Can't Transformers Learn Multiplication? Reverse-Engineering Reveals Long-Range Dependency Pitfalls
by: Bai, Xiaoyan, et al.
Published: (2025)
by: Bai, Xiaoyan, et al.
Published: (2025)
Tensor Product Representation Probes Reveal Shared Structure Across Linear Directions
by: Lee, Andrew, et al.
Published: (2026)
by: Lee, Andrew, et al.
Published: (2026)
Decomposing Query-Key Feature Interactions Using Contrastive Covariances
by: Lee, Andrew, et al.
Published: (2026)
by: Lee, Andrew, et al.
Published: (2026)
Shared Global and Local Geometry of Language Model Embeddings
by: Lee, Andrew, et al.
Published: (2025)
by: Lee, Andrew, et al.
Published: (2025)
Does visualization help AI understand data?
by: Li, Victoria R., et al.
Published: (2025)
by: Li, Victoria R., et al.
Published: (2025)
Generalization in Neural Networks: A Broad Survey
by: Rohlfs, Chris
Published: (2022)
by: Rohlfs, Chris
Published: (2022)
A Survey on Knowledge Editing of Neural Networks
by: Mazzia, Vittorio, et al.
Published: (2023)
by: Mazzia, Vittorio, et al.
Published: (2023)
Call Me Maybe: Enhancing JavaScript Call Graph Construction using Graph Neural Networks
by: Bhuiyan, Masudul Hasan Masud, et al.
Published: (2025)
by: Bhuiyan, Masudul Hasan Masud, et al.
Published: (2025)
A Comprehensive Survey on Self-Interpretable Neural Networks
by: Ji, Yang, et al.
Published: (2025)
by: Ji, Yang, et al.
Published: (2025)
Adaptive Physics-informed Neural Networks: A Survey
by: Torres, Edgar, et al.
Published: (2025)
by: Torres, Edgar, et al.
Published: (2025)
Can Interpretation Predict Behavior on Unseen Data?
by: Li, Victoria R., et al.
Published: (2025)
by: Li, Victoria R., et al.
Published: (2025)
Survey on Generalization Theory for Graph Neural Networks
by: Vasileiou, Antonis, et al.
Published: (2025)
by: Vasileiou, Antonis, et al.
Published: (2025)
To Call or Not to Call: Diagnosing Intrinsic Over-Calling Bias in LLM Agents
by: Shi, Wei, et al.
Published: (2026)
by: Shi, Wei, et al.
Published: (2026)
A Relational Inductive Bias for Dimensional Abstraction in Neural Networks
by: Campbell, Declan, et al.
Published: (2024)
by: Campbell, Declan, et al.
Published: (2024)
Ethical AI for Young Digital Citizens: A Call to Action on Privacy Governance
by: Shouli, Austin, et al.
Published: (2025)
by: Shouli, Austin, et al.
Published: (2025)
A Survey on Deep Neural Networks in Collaborative Filtering Recommendation Systems
by: Li, Pang, et al.
Published: (2024)
by: Li, Pang, et al.
Published: (2024)
Oversmoothing Alleviation in Graph Neural Networks: A Survey and Unified View
by: Jin, Yufei, et al.
Published: (2024)
by: Jin, Yufei, et al.
Published: (2024)
Deep Model Merging: The Sister of Neural Network Interpretability -- A Survey
by: Khan, Arham, et al.
Published: (2024)
by: Khan, Arham, et al.
Published: (2024)
Graph Neural Networks for Job Shop Scheduling Problems: A Survey
by: Smit, Igor G., et al.
Published: (2024)
by: Smit, Igor G., et al.
Published: (2024)
Low-bit Model Quantization for Deep Neural Networks: A Survey
by: Liu, Kai, et al.
Published: (2025)
by: Liu, Kai, et al.
Published: (2025)
Systematic Relational Reasoning With Epistemic Graph Neural Networks
by: Khalid, Irtaza, et al.
Published: (2024)
by: Khalid, Irtaza, et al.
Published: (2024)
From Language to Action in Arabic: Reliable Structured Tool Calling via Data-Centric Fine-Tuning
by: Nacar, Omer, et al.
Published: (2026)
by: Nacar, Omer, et al.
Published: (2026)
Recommender Systems for Good (RS4Good): Survey of Use Cases and a Call to Action for Research that Matters
by: Jannach, Dietmar, et al.
Published: (2024)
by: Jannach, Dietmar, et al.
Published: (2024)
Compositional Function Networks: A High-Performance Alternative to Deep Neural Networks with Built-in Interpretability
by: Li, Fang
Published: (2025)
by: Li, Fang
Published: (2025)
A Survey on Graph Neural Networks for Fraud Detection in Ride Hailing Platforms
by: Hewageegana, Kanishka, et al.
Published: (2025)
by: Hewageegana, Kanishka, et al.
Published: (2025)
Graph Neural Networks in Multi-Omics Cancer Research: A Structured Survey
by: Zohari, Payam, et al.
Published: (2025)
by: Zohari, Payam, et al.
Published: (2025)
A Theoretical Analysis of Compositional Generalization in Neural Networks: A Necessary and Sufficient Condition
by: Li, Yuanpeng
Published: (2025)
by: Li, Yuanpeng
Published: (2025)
SRNN: Spatiotemporal Relational Neural Network for Intuitive Physics Understanding
by: Yang, Fei
Published: (2025)
by: Yang, Fei
Published: (2025)
From Algorithm to Hardware: A Survey on Efficient and Safe Deployment of Deep Neural Networks
by: Geng, Xue, et al.
Published: (2024)
by: Geng, Xue, et al.
Published: (2024)
A Call to Action for a Secure-by-Design Generative AI Paradigm
by: Alharthi, Dalal, et al.
Published: (2025)
by: Alharthi, Dalal, et al.
Published: (2025)
Multi-Relational Graph Neural Network for Out-of-Domain Link Prediction
by: Sattar, Asma, et al.
Published: (2024)
by: Sattar, Asma, et al.
Published: (2024)
Hyperbolic Hypergraph Neural Networks for Multi-Relational Knowledge Hypergraph Representation
by: Li, Mengfan, et al.
Published: (2024)
by: Li, Mengfan, et al.
Published: (2024)
What Planning Problems Can A Relational Neural Network Solve?
by: Mao, Jiayuan, et al.
Published: (2023)
by: Mao, Jiayuan, et al.
Published: (2023)
Similar Items
-
Dialogue Action Tokens: Steering Language Models in Goal-Directed Dialogue with a Multi-Turn Planner
by: Li, Kenneth, et al.
Published: (2024) -
When Bad Data Leads to Good Models
by: Li, Kenneth, et al.
Published: (2025) -
The Geometry of Self-Verification in a Task-Specific Reasoning Model
by: Lee, Andrew, et al.
Published: (2025) -
Inference-Time Intervention: Eliciting Truthful Answers from a Language Model
by: Li, Kenneth, et al.
Published: (2023) -
What Does it Mean for a Neural Network to Learn a "World Model"?
by: Li, Kenneth, et al.
Published: (2025)