Persistent Topological Features in Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Gardinazzi, Yuri, Viswanathan, Karthik, Panerai, Giada, Ansuini, Alessio, Cazzaniga, Alberto, Biagetti, Matteo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Geometry of Tokens in Internal Representations of Large Language Models
di: Viswanathan, Karthik, et al.
Pubblicazione: (2025)
di: Viswanathan, Karthik, et al.
Pubblicazione: (2025)
The representation landscape of few-shot learning and fine-tuning in large language models
di: Doimo, Diego, et al.
Pubblicazione: (2024)
di: Doimo, Diego, et al.
Pubblicazione: (2024)
Zigzag Persistence of Neural Responses to Time-Varying Stimuli
di: Gardinazzi, Yuri, et al.
Pubblicazione: (2026)
di: Gardinazzi, Yuri, et al.
Pubblicazione: (2026)
Emergent representations in networks trained with the Forward-Forward algorithm
di: Tosato, Niccolò, et al.
Pubblicazione: (2023)
di: Tosato, Niccolò, et al.
Pubblicazione: (2023)
Probing Geometry of Next Token Prediction Using Cumulant Expansion of the Softmax Entropy
di: Viswanathan, Karthik, et al.
Pubblicazione: (2025)
di: Viswanathan, Karthik, et al.
Pubblicazione: (2025)
The Flood Complex: Large-Scale Persistent Homology on Millions of Points
di: Graf, Florian, et al.
Pubblicazione: (2025)
di: Graf, Florian, et al.
Pubblicazione: (2025)
Bias after Prompting: Persistent Discrimination in Large Language Models
di: Sivakumar, Nivedha, et al.
Pubblicazione: (2025)
di: Sivakumar, Nivedha, et al.
Pubblicazione: (2025)
The Narrow Gate: Localized Image-Text Communication in Native Multimodal Models
di: Serra, Alessandro Pietro, et al.
Pubblicazione: (2024)
di: Serra, Alessandro Pietro, et al.
Pubblicazione: (2024)
Algorithm for Interpretable Graph Features via Motivic Persistent Cohomology
di: Maruyama, Yoshihiro
Pubblicazione: (2025)
di: Maruyama, Yoshihiro
Pubblicazione: (2025)
Topological Autoencoders++: Fast and Accurate Cycle-Aware Dimensionality Reduction
di: Clémot, Mattéo, et al.
Pubblicazione: (2025)
di: Clémot, Mattéo, et al.
Pubblicazione: (2025)
Rotary Offset Features in Large Language Models
di: Jonasson, André
Pubblicazione: (2025)
di: Jonasson, André
Pubblicazione: (2025)
Robust Barycenters of Persistence Diagrams
di: Sisouk, Keanu, et al.
Pubblicazione: (2025)
di: Sisouk, Keanu, et al.
Pubblicazione: (2025)
Understanding and inverse design of implicit bias in stochastic learning: a geometric perspective
di: Aladrah, Nicola, et al.
Pubblicazione: (2026)
di: Aladrah, Nicola, et al.
Pubblicazione: (2026)
PHLP: Sole Persistent Homology for Link Prediction - Interpretable Feature Extraction
di: You, Junwon, et al.
Pubblicazione: (2024)
di: You, Junwon, et al.
Pubblicazione: (2024)
Automatically Interpreting Millions of Features in Large Language Models
di: Paulo, Gonçalo, et al.
Pubblicazione: (2024)
di: Paulo, Gonçalo, et al.
Pubblicazione: (2024)
Semantic Structure of Feature Space in Large Language Models
di: Kozlowski, Austin C., et al.
Pubblicazione: (2026)
di: Kozlowski, Austin C., et al.
Pubblicazione: (2026)
Latent Feature Mining for Predictive Model Enhancement with Large Language Models
di: Li, Bingxuan, et al.
Pubblicazione: (2024)
di: Li, Bingxuan, et al.
Pubblicazione: (2024)
GeoLAN: Geometric Learning of Latent Explanatory Directions in Large Language Models
di: Pan, Tianyu Bell, et al.
Pubblicazione: (2026)
di: Pan, Tianyu Bell, et al.
Pubblicazione: (2026)
The Curved Spacetime of Transformer Architectures
di: Di Sipio, Riccardo, et al.
Pubblicazione: (2025)
di: Di Sipio, Riccardo, et al.
Pubblicazione: (2025)
Feature Alignment and Representation Transfer in Knowledge Distillation for Large Language Models
di: Yang, Junjie, et al.
Pubblicazione: (2025)
di: Yang, Junjie, et al.
Pubblicazione: (2025)
In-Context Symbolic Regression: Leveraging Large Language Models for Function Discovery
di: Merler, Matteo, et al.
Pubblicazione: (2024)
di: Merler, Matteo, et al.
Pubblicazione: (2024)
Head Pursuit: Probing Attention Specialization in Multimodal Transformers
di: Basile, Lorenzo, et al.
Pubblicazione: (2025)
di: Basile, Lorenzo, et al.
Pubblicazione: (2025)
Persistent Homology via Finite Topological Spaces
di: Kayacan, Selçuk
Pubblicazione: (2025)
di: Kayacan, Selçuk
Pubblicazione: (2025)
Topological Machine Learning with Unreduced Persistence Diagrams
di: Abreu, Nicole, et al.
Pubblicazione: (2025)
di: Abreu, Nicole, et al.
Pubblicazione: (2025)
Wu's Method can Boost Symbolic AI to Rival Silver Medalists and AlphaGeometry to Outperform Gold Medalists at IMO Geometry
di: Sinha, Shiven, et al.
Pubblicazione: (2024)
di: Sinha, Shiven, et al.
Pubblicazione: (2024)
Pre-trained Large Language Models Use Fourier Features to Compute Addition
di: Zhou, Tianyi, et al.
Pubblicazione: (2024)
di: Zhou, Tianyi, et al.
Pubblicazione: (2024)
Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models
di: Deng, Boyi, et al.
Pubblicazione: (2026)
di: Deng, Boyi, et al.
Pubblicazione: (2026)
Topological Spatial Graph Coarsening
di: Calissano, Anna, et al.
Pubblicazione: (2025)
di: Calissano, Anna, et al.
Pubblicazione: (2025)
Cover Learning for Large-Scale Topology Representation
di: Scoccola, Luis, et al.
Pubblicazione: (2025)
di: Scoccola, Luis, et al.
Pubblicazione: (2025)
Narrative Feature or Structured Feature? A Study of Large Language Models to Identify Cancer Patients at Risk of Heart Failure
di: Chen, Ziyi, et al.
Pubblicazione: (2024)
di: Chen, Ziyi, et al.
Pubblicazione: (2024)
NeuroPrune: A Neuro-inspired Topological Sparse Training Algorithm for Large Language Models
di: Dhurandhar, Amit, et al.
Pubblicazione: (2024)
di: Dhurandhar, Amit, et al.
Pubblicazione: (2024)
MoSE: Mixture of Slimmable Experts for Efficient and Adaptive Language Models
di: Tastan, Nurbek, et al.
Pubblicazione: (2026)
di: Tastan, Nurbek, et al.
Pubblicazione: (2026)
The Birth of Knowledge: Emergent Features across Time, Space, and Scale in Large Language Models
di: Sawmya, Shashata, et al.
Pubblicazione: (2025)
di: Sawmya, Shashata, et al.
Pubblicazione: (2025)
Certifying Robustness via Topological Representations
di: Agerberg, Jens, et al.
Pubblicazione: (2025)
di: Agerberg, Jens, et al.
Pubblicazione: (2025)
Adaptive Feature-based Low-Rank Compression of Large Language Models via Bayesian Optimization
di: Ji, Yixin, et al.
Pubblicazione: (2024)
di: Ji, Yixin, et al.
Pubblicazione: (2024)
LLM-Select: Feature Selection with Large Language Models
di: Jeong, Daniel P., et al.
Pubblicazione: (2024)
di: Jeong, Daniel P., et al.
Pubblicazione: (2024)
Quantum Attention by Overlap Interference: Predicting Sequences from Classical and Many-Body Quantum Data
di: Pecilli, Alessio, et al.
Pubblicazione: (2026)
di: Pecilli, Alessio, et al.
Pubblicazione: (2026)
Rapid and Precise Topological Comparison with Merge Tree Neural Networks
di: Qin, Yu, et al.
Pubblicazione: (2024)
di: Qin, Yu, et al.
Pubblicazione: (2024)
Knowledge-Driven Feature Selection and Engineering for Genotype Data with Large Language Models
di: Lee, Joseph, et al.
Pubblicazione: (2024)
di: Lee, Joseph, et al.
Pubblicazione: (2024)
Param$Δ$ for Direct Weight Mixing: Post-Train Large Language Model at Zero Cost
di: Cao, Sheng, et al.
Pubblicazione: (2025)
di: Cao, Sheng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
The Geometry of Tokens in Internal Representations of Large Language Models
di: Viswanathan, Karthik, et al.
Pubblicazione: (2025) -
The representation landscape of few-shot learning and fine-tuning in large language models
di: Doimo, Diego, et al.
Pubblicazione: (2024) -
Zigzag Persistence of Neural Responses to Time-Varying Stimuli
di: Gardinazzi, Yuri, et al.
Pubblicazione: (2026) -
Emergent representations in networks trained with the Forward-Forward algorithm
di: Tosato, Niccolò, et al.
Pubblicazione: (2023) -
Probing Geometry of Next Token Prediction Using Cumulant Expansion of the Softmax Entropy
di: Viswanathan, Karthik, et al.
Pubblicazione: (2025)