Modeling citation worthiness by using attention-based bidirectional long short-term memory networks and interpretable models
Fuente:
arXiv
Saved in:
| Main Authors: | Zeng, Tong, Acuna, Daniel E. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dataset Mention Extraction in Scientific Articles Using Bi-LSTM-CRF Model
by: Zeng, Tong, et al.
Published: (2024)
by: Zeng, Tong, et al.
Published: (2024)
A hybrid transformer and attention based recurrent neural network for robust and interpretable sentiment analysis of tweets
by: Jahin, Md Abrar, et al.
Published: (2024)
by: Jahin, Md Abrar, et al.
Published: (2024)
Human-interpretable clustering of short-text using large language models
by: Miller, Justin K., et al.
Published: (2024)
by: Miller, Justin K., et al.
Published: (2024)
Zero-shot data citation function classification using transformer-based large language models (LLMs)
by: Byers, Neil, et al.
Published: (2025)
by: Byers, Neil, et al.
Published: (2025)
TransformerFAM: Feedback attention is working memory
by: Hwang, Dongseong, et al.
Published: (2024)
by: Hwang, Dongseong, et al.
Published: (2024)
Prompt reinforcing for long-term planning of large language models
by: Lin, Hsien-Chin, et al.
Published: (2025)
by: Lin, Hsien-Chin, et al.
Published: (2025)
Forma mentis networks predict creativity ratings of short texts via interpretable artificial intelligence in human and GPT-simulated raters
by: Haim, Edith, et al.
Published: (2024)
by: Haim, Edith, et al.
Published: (2024)
Simple linear attention language models balance the recall-throughput tradeoff
by: Arora, Simran, et al.
Published: (2024)
by: Arora, Simran, et al.
Published: (2024)
Relational inductive biases on attention mechanisms
by: Mijangos, Víctor, et al.
Published: (2025)
by: Mijangos, Víctor, et al.
Published: (2025)
Pretraining with hierarchical memories: separating long-tail and common knowledge
by: Pouransari, Hadi, et al.
Published: (2025)
by: Pouransari, Hadi, et al.
Published: (2025)
Integrating a Heterogeneous Graph with Entity-aware Self-attention using Relative Position Labels for Reading Comprehension Model
by: Foolad, Shima, et al.
Published: (2023)
by: Foolad, Shima, et al.
Published: (2023)
From communities to interpretable network and word embedding: an unified approach
by: Prouteau, Thibault, et al.
Published: (2024)
by: Prouteau, Thibault, et al.
Published: (2024)
LLM-based feature generation from text for interpretable machine learning
by: Balek, Vojtěch, et al.
Published: (2024)
by: Balek, Vojtěch, et al.
Published: (2024)
Enhancing Sindhi Word Segmentation using Subword Representation Learning and Position-aware Self-attention
by: Ali, Wazir, et al.
Published: (2020)
by: Ali, Wazir, et al.
Published: (2020)
The study of short texts in digital politics: Document aggregation for topic modeling
by: Nakka, Nitheesha, et al.
Published: (2025)
by: Nakka, Nitheesha, et al.
Published: (2025)
Combining topic modelling and citation network analysis to study case law from the European Court on Human Rights on the right to respect for private and family life
by: Mohammadi, M., et al.
Published: (2024)
by: Mohammadi, M., et al.
Published: (2024)
AtteSTNet -- An attention and subword tokenization based approach for code-switched text hate speech detection
by: Shingi, Geet, et al.
Published: (2021)
by: Shingi, Geet, et al.
Published: (2021)
Boosting classification reliability of NLP transformer models in the long run
by: Kmetty, Zoltán, et al.
Published: (2023)
by: Kmetty, Zoltán, et al.
Published: (2023)
Fresh in memory: Training-order recency is linearly encoded in language model activations
by: Krasheninnikov, Dmitrii, et al.
Published: (2025)
by: Krasheninnikov, Dmitrii, et al.
Published: (2025)
Comparison of different Unique hard attention transformer models by the formal languages they can recognize
by: Ryvkin, Leonid
Published: (2025)
by: Ryvkin, Leonid
Published: (2025)
Probing self-attention in self-supervised speech models for cross-linguistic differences
by: Gopinath, Sai, et al.
Published: (2024)
by: Gopinath, Sai, et al.
Published: (2024)
Surrogate modeling for interpreting black-box LLMs in medical predictions
by: Han, Changho, et al.
Published: (2026)
by: Han, Changho, et al.
Published: (2026)
Can sparse autoencoders be used to decompose and interpret steering vectors?
by: Mayne, Harry, et al.
Published: (2024)
by: Mayne, Harry, et al.
Published: (2024)
Research on the Application of Deep Learning-based BERT Model in Sentiment Analysis
by: Wu, Yichao, et al.
Published: (2024)
by: Wu, Yichao, et al.
Published: (2024)
Are LLM-based methods good enough for detecting unfair terms of service?
by: Frasheri, Mirgita, et al.
Published: (2024)
by: Frasheri, Mirgita, et al.
Published: (2024)
Thermodynamics of bidirectional associative memories
by: Barra, Adriano, et al.
Published: (2022)
by: Barra, Adriano, et al.
Published: (2022)
Dissecting Long-Chain-of-Thought Reasoning Models: An Empirical Study
by: Mu, Yongyu, et al.
Published: (2025)
by: Mu, Yongyu, et al.
Published: (2025)
AdaLomo: Low-memory Optimization with Adaptive Learning Rate
by: Lv, Kai, et al.
Published: (2023)
by: Lv, Kai, et al.
Published: (2023)
ChatGPT for automated grading of short answer questions in mechanical ventilation
by: Jade, Tejas, et al.
Published: (2025)
by: Jade, Tejas, et al.
Published: (2025)
Taming Knowledge Conflicts in Language Models
by: Li, Gaotang, et al.
Published: (2025)
by: Li, Gaotang, et al.
Published: (2025)
Phase transition on a context-sensitive random language model with short range interactions
by: Toji, Yuma, et al.
Published: (2026)
by: Toji, Yuma, et al.
Published: (2026)
Offline Exploration-Aware Fine-Tuning for Long-Chain Mathematical Reasoning
by: Mu, Yongyu, et al.
Published: (2026)
by: Mu, Yongyu, et al.
Published: (2026)
From Accuracy to Robustness: A Study of Rule- and Model-based Verifiers in Mathematical Reasoning
by: Huang, Yuzhen, et al.
Published: (2025)
by: Huang, Yuzhen, et al.
Published: (2025)
Discovering influential text using convolutional neural networks
by: Ayers, Megan, et al.
Published: (2024)
by: Ayers, Megan, et al.
Published: (2024)
The more polypersonal the better -- a short look on space geometry of fine-tuned layers
by: Kudriashov, Sergei, et al.
Published: (2025)
by: Kudriashov, Sergei, et al.
Published: (2025)
What are you sinking? A geometric approach on attention sink
by: Ruscio, Valeria, et al.
Published: (2025)
by: Ruscio, Valeria, et al.
Published: (2025)
Detection of developmental language disorder in Cypriot Greek children using a neural network algorithm
by: Georgiou, Georgios P., et al.
Published: (2023)
by: Georgiou, Georgios P., et al.
Published: (2023)
Latent learning: episodic memory complements parametric learning by enabling flexible reuse of experiences
by: Lampinen, Andrew Kyle, et al.
Published: (2025)
by: Lampinen, Andrew Kyle, et al.
Published: (2025)
GotFunding: A grant recommendation system based on scientific articles
by: Zeng, Tong, et al.
Published: (2024)
by: Zeng, Tong, et al.
Published: (2024)
Adjoint sharding for very long context training of state space models
by: Xu, Xingzi, et al.
Published: (2025)
by: Xu, Xingzi, et al.
Published: (2025)
Similar Items
-
Dataset Mention Extraction in Scientific Articles Using Bi-LSTM-CRF Model
by: Zeng, Tong, et al.
Published: (2024) -
A hybrid transformer and attention based recurrent neural network for robust and interpretable sentiment analysis of tweets
by: Jahin, Md Abrar, et al.
Published: (2024) -
Human-interpretable clustering of short-text using large language models
by: Miller, Justin K., et al.
Published: (2024) -
Zero-shot data citation function classification using transformer-based large language models (LLMs)
by: Byers, Neil, et al.
Published: (2025) -
TransformerFAM: Feedback attention is working memory
by: Hwang, Dongseong, et al.
Published: (2024)