The Propensity for Density in Feed-forward Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Schoots, Nandi, Jackson, Alex, Kholmovaia, Ali, McBurney, Peter, Shanahan, Murray |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Training Neural Networks for Modularity aids Interpretability
von: Golechha, Satvik, et al.
Veröffentlicht: (2024)
von: Golechha, Satvik, et al.
Veröffentlicht: (2024)
Decoding Communications with Partial Information
von: Cope, Dylan, et al.
Veröffentlicht: (2025)
von: Cope, Dylan, et al.
Veröffentlicht: (2025)
Learning Translations: Emergent Communication Pretraining for Cooperative Language Acquisition
von: Cope, Dylan, et al.
Veröffentlicht: (2024)
von: Cope, Dylan, et al.
Veröffentlicht: (2024)
Still "Talking About Large Language Models": Some Clarifications
von: Shanahan, Murray
Veröffentlicht: (2024)
von: Shanahan, Murray
Veröffentlicht: (2024)
Relating Piecewise Linear Kolmogorov Arnold Networks to ReLU Networks
von: Schoots, Nandi, et al.
Veröffentlicht: (2025)
von: Schoots, Nandi, et al.
Veröffentlicht: (2025)
The Topos of Transformer Networks
von: Villani, Mattia Jacopo, et al.
Veröffentlicht: (2024)
von: Villani, Mattia Jacopo, et al.
Veröffentlicht: (2024)
Studying Cross-cluster Modularity in Neural Networks
von: Golechha, Satvik, et al.
Veröffentlicht: (2025)
von: Golechha, Satvik, et al.
Veröffentlicht: (2025)
Existential Conversations with Large Language Models: Content, Community, and Culture
von: Shanahan, Murray, et al.
Veröffentlicht: (2024)
von: Shanahan, Murray, et al.
Veröffentlicht: (2024)
Extending Activation Steering to Broad Skills and Multiple Behaviours
von: van der Weij, Teun, et al.
Veröffentlicht: (2024)
von: van der Weij, Teun, et al.
Veröffentlicht: (2024)
Soft Contamination Means Benchmarks Test Shallow Generalization
von: Spiesberger, Ari, et al.
Veröffentlicht: (2026)
von: Spiesberger, Ari, et al.
Veröffentlicht: (2026)
A Closed-form Solution for Weight Optimization in Fully-connected Feed-forward Neural Networks
von: Tomic, Slavisa, et al.
Veröffentlicht: (2024)
von: Tomic, Slavisa, et al.
Veröffentlicht: (2024)
Dissecting Language Models: Machine Unlearning via Selective Pruning
von: Pochinkov, Nicholas, et al.
Veröffentlicht: (2024)
von: Pochinkov, Nicholas, et al.
Veröffentlicht: (2024)
PropensityBench: Evaluating Latent Safety Risks in Large Language Models via an Agentic Approach
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2025)
von: Sehwag, Udari Madhushani, et al.
Veröffentlicht: (2025)
Distributional Associations vs In-Context Reasoning: A Study of Feed-forward and Attention Layers
von: Chen, Lei, et al.
Veröffentlicht: (2024)
von: Chen, Lei, et al.
Veröffentlicht: (2024)
Transformers Use Causal World Models in Maze-Solving Tasks
von: Spies, Alex F., et al.
Veröffentlicht: (2024)
von: Spies, Alex F., et al.
Veröffentlicht: (2024)
Capabilities Ain't All You Need: Measuring Propensities in AI
von: Romero-Alvarado, Daniel, et al.
Veröffentlicht: (2026)
von: Romero-Alvarado, Daniel, et al.
Veröffentlicht: (2026)
On the Role of Transformer Feed-Forward Layers in Nonlinear In-Context Learning
von: Sun, Haoyuan, et al.
Veröffentlicht: (2025)
von: Sun, Haoyuan, et al.
Veröffentlicht: (2025)
Epistemically-guided forward-backward exploration
von: Urpí, Núria Armengol, et al.
Veröffentlicht: (2025)
von: Urpí, Núria Armengol, et al.
Veröffentlicht: (2025)
One Mask to Rule Them All: On Hidden Facts after Editing and How to Find Them
von: Holmov, Ali, et al.
Veröffentlicht: (2026)
von: Holmov, Ali, et al.
Veröffentlicht: (2026)
On the generalization of language models from in-context learning and finetuning: a controlled study
von: Lampinen, Andrew K., et al.
Veröffentlicht: (2025)
von: Lampinen, Andrew K., et al.
Veröffentlicht: (2025)
Preferential subspace identification (PSID) with forward-backward smoothing
von: Sani, Omid G., et al.
Veröffentlicht: (2025)
von: Sani, Omid G., et al.
Veröffentlicht: (2025)
Lookbehind-SAM: k steps back, 1 step forward
von: Mordido, Gonçalo, et al.
Veröffentlicht: (2023)
von: Mordido, Gonçalo, et al.
Veröffentlicht: (2023)
Quotient DAGs for Off-Policy Evaluation:Forward-Flow Importance Sampling and Exact Slate Propensities
von: Xie, Ziwen, et al.
Veröffentlicht: (2026)
von: Xie, Ziwen, et al.
Veröffentlicht: (2026)
Optimizing Dense Feed-Forward Neural Networks
von: Balderas, Luis, et al.
Veröffentlicht: (2023)
von: Balderas, Luis, et al.
Veröffentlicht: (2023)
Simulacra as Conscious Exotica
von: Shanahan, Murray
Veröffentlicht: (2024)
von: Shanahan, Murray
Veröffentlicht: (2024)
Palatable Conceptions of Disembodied Being
von: Shanahan, Murray
Veröffentlicht: (2025)
von: Shanahan, Murray
Veröffentlicht: (2025)
Enhancing Customer Service Chatbots with Context-Aware NLU through Selective Attention and Multi-task Learning
von: Nandi, Subhadip, et al.
Veröffentlicht: (2025)
von: Nandi, Subhadip, et al.
Veröffentlicht: (2025)
Mimicry and the Emergence of Cooperative Communication
von: Cope, Dylan, et al.
Veröffentlicht: (2024)
von: Cope, Dylan, et al.
Veröffentlicht: (2024)
Feed-Forward Optimization With Delayed Feedback for Neural Network Training
von: Flügel, Katharina, et al.
Veröffentlicht: (2023)
von: Flügel, Katharina, et al.
Veröffentlicht: (2023)
Sobolev Space Regularised Pre Density Models
von: Kozdoba, Mark, et al.
Veröffentlicht: (2023)
von: Kozdoba, Mark, et al.
Veröffentlicht: (2023)
Generative Modeling with Flow-Guided Density Ratio Learning
von: Heng, Alvin, et al.
Veröffentlicht: (2023)
von: Heng, Alvin, et al.
Veröffentlicht: (2023)
Enhanced Generative Model Evaluation with Clipped Density and Coverage
von: Salvy, Nicolas, et al.
Veröffentlicht: (2025)
von: Salvy, Nicolas, et al.
Veröffentlicht: (2025)
Beyond Demand Estimation: Consumer Surplus Evaluation via Cumulative Propensity Weights
von: Bian, Zeyu, et al.
Veröffentlicht: (2026)
von: Bian, Zeyu, et al.
Veröffentlicht: (2026)
Deep-layer limit and stability analysis of the basic forward-backward-splitting induced network (II): learning problems
von: Lin, Xuan, et al.
Veröffentlicht: (2026)
von: Lin, Xuan, et al.
Veröffentlicht: (2026)
Adaptive Transformer Modelling of Density Function for Nonparametric Survival Analysis
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
von: Zhang, Xin, et al.
Veröffentlicht: (2024)
Contrastive Entropy Bounds for Density and Conditional Density Decomposition
von: Hu, Bo, et al.
Veröffentlicht: (2025)
von: Hu, Bo, et al.
Veröffentlicht: (2025)
FLoRA: Fused forward-backward adapters for parameter efficient fine-tuning and reducing inference-time latencies of LLMs
von: Gowda, Dhananjaya, et al.
Veröffentlicht: (2025)
von: Gowda, Dhananjaya, et al.
Veröffentlicht: (2025)
Enhancing Fast Feed Forward Networks with Load Balancing and a Master Leaf Node
von: Charalampopoulos, Andreas, et al.
Veröffentlicht: (2024)
von: Charalampopoulos, Andreas, et al.
Veröffentlicht: (2024)
Dynamic sparsity in tree-structured feed-forward layers at scale
von: Sedghi, Reza, et al.
Veröffentlicht: (2026)
von: Sedghi, Reza, et al.
Veröffentlicht: (2026)
An Overview and Discussion of the Suitability of Existing Speech Datasets to Train Machine Learning Models for Collective Problem Solving
von: Villuri, Gnaneswar, et al.
Veröffentlicht: (2024)
von: Villuri, Gnaneswar, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Training Neural Networks for Modularity aids Interpretability
von: Golechha, Satvik, et al.
Veröffentlicht: (2024) -
Decoding Communications with Partial Information
von: Cope, Dylan, et al.
Veröffentlicht: (2025) -
Learning Translations: Emergent Communication Pretraining for Cooperative Language Acquisition
von: Cope, Dylan, et al.
Veröffentlicht: (2024) -
Still "Talking About Large Language Models": Some Clarifications
von: Shanahan, Murray
Veröffentlicht: (2024) -
Relating Piecewise Linear Kolmogorov Arnold Networks to ReLU Networks
von: Schoots, Nandi, et al.
Veröffentlicht: (2025)