Breaking Free Transformer Models: Task-specific Context Attribution Promises Improved Generalizability Without Fine-tuning Pre-trained LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tytarenko, Stepan, Amin, Mohammad Ruhul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Kastor: Fine-tuned Small Language Models for Shape-based Active Relation Extraction
von: Celian, Ringwald, et al.
Veröffentlicht: (2025)
von: Celian, Ringwald, et al.
Veröffentlicht: (2025)
Improving LLMs with a knowledge from databases
von: Máša, Petr
Veröffentlicht: (2025)
von: Máša, Petr
Veröffentlicht: (2025)
Pre-trained Language Model with Prompts for Temporal Knowledge Graph Completion
von: Xu, Wenjie, et al.
Veröffentlicht: (2023)
von: Xu, Wenjie, et al.
Veröffentlicht: (2023)
From Fake Focus to Real Precision: Confusion-Driven Adversarial Attention Learning in Transformers
von: Liu, Yawei
Veröffentlicht: (2025)
von: Liu, Yawei
Veröffentlicht: (2025)
Unveiling the Secret Recipe: A Guide For Supervised Fine-Tuning Small LLMs
von: Pareja, Aldo, et al.
Veröffentlicht: (2024)
von: Pareja, Aldo, et al.
Veröffentlicht: (2024)
Mixture of Experts Approaches in Dense Retrieval Tasks
von: Sokli, Effrosyni, et al.
Veröffentlicht: (2025)
von: Sokli, Effrosyni, et al.
Veröffentlicht: (2025)
Learning, Fast and Slow: Towards LLMs That Adapt Continually
von: Tiwari, Rishabh, et al.
Veröffentlicht: (2026)
von: Tiwari, Rishabh, et al.
Veröffentlicht: (2026)
Cross-Lingual Generalization and Compression: From Language-Specific to Shared Neurons
von: Riemenschneider, Frederick, et al.
Veröffentlicht: (2025)
von: Riemenschneider, Frederick, et al.
Veröffentlicht: (2025)
CR-LT-KGQA: A Knowledge Graph Question Answering Dataset Requiring Commonsense Reasoning and Long-Tail Knowledge
von: Guo, Willis, et al.
Veröffentlicht: (2024)
von: Guo, Willis, et al.
Veröffentlicht: (2024)
Overcoming the Generalization Limits of SLM Finetuning for Shape-Based Extraction of Datatype and Object Properties
von: Ringwald, Célian, et al.
Veröffentlicht: (2025)
von: Ringwald, Célian, et al.
Veröffentlicht: (2025)
RAPTOR-AI for Disaster OODA Loop: Hierarchical Multimodal RAG with Experience-Driven Agentic Decision-Making
von: Yasuno, Takato
Veröffentlicht: (2026)
von: Yasuno, Takato
Veröffentlicht: (2026)
Automatic Prompt Optimization for Knowledge Graph Construction: Insights from an Empirical Study
von: Mihindukulasooriya, Nandana, et al.
Veröffentlicht: (2025)
von: Mihindukulasooriya, Nandana, et al.
Veröffentlicht: (2025)
SemanticAgent: A Semantics-Aware Framework for Text-to-SQL Data Synthesis
von: Gao, Qiang, et al.
Veröffentlicht: (2026)
von: Gao, Qiang, et al.
Veröffentlicht: (2026)
Large Language Models as Oracles for Ontology Alignment
von: Lushnei, Sviatoslav, et al.
Veröffentlicht: (2025)
von: Lushnei, Sviatoslav, et al.
Veröffentlicht: (2025)
Multi-Paradigm Agent Interaction in Practice:A Systematic Analysis of Generator-Evaluator, ReAct Loop,and Adversarial Evaluation in the buddyMe Framework
von: Wang, Xiaohua, et al.
Veröffentlicht: (2026)
von: Wang, Xiaohua, et al.
Veröffentlicht: (2026)
OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models
von: Cotti, Luca, et al.
Veröffentlicht: (2025)
von: Cotti, Luca, et al.
Veröffentlicht: (2025)
HIP Network: Historical Information Passing Network for Extrapolation Reasoning on Temporal Knowledge Graph
von: He, Yongquan, et al.
Veröffentlicht: (2024)
von: He, Yongquan, et al.
Veröffentlicht: (2024)
SAGE: A Strategy-Aware Graph-Enhanced Generation Framework For Online Counseling
von: Aharon, Eliya Naomi, et al.
Veröffentlicht: (2026)
von: Aharon, Eliya Naomi, et al.
Veröffentlicht: (2026)
LLM-Assisted Formalization Enables Deterministic Detection of Statutory Inconsistency in the Internal Revenue Code
von: Yadamsuren, Borchuluun, et al.
Veröffentlicht: (2025)
von: Yadamsuren, Borchuluun, et al.
Veröffentlicht: (2025)
MIRAGE: Scaling Test-Time Inference with Parallel Graph-Retrieval-Augmented Reasoning Chains
von: Wei, Kaiwen, et al.
Veröffentlicht: (2025)
von: Wei, Kaiwen, et al.
Veröffentlicht: (2025)
Retrieval-Based Multi-Label Legal Annotation: Extensible, Data-Efficient and Hallucination-Free
von: Zhang, Li, et al.
Veröffentlicht: (2026)
von: Zhang, Li, et al.
Veröffentlicht: (2026)
A systematic review of relation extraction task since the emergence of Transformers
von: Celian, Ringwald, et al.
Veröffentlicht: (2025)
von: Celian, Ringwald, et al.
Veröffentlicht: (2025)
ROZA Graphs: Self-Improving Near-Deterministic RAG through Evidence-Centric Feedback
von: Penaroza, Matthew
Veröffentlicht: (2026)
von: Penaroza, Matthew
Veröffentlicht: (2026)
LLMs for Game Theory: Entropy-Guided In-Context Learning and Adaptive CoT Reasoning
von: Banfi, Tommaso Felice, et al.
Veröffentlicht: (2026)
von: Banfi, Tommaso Felice, et al.
Veröffentlicht: (2026)
GraphWalk: Enabling Reasoning in Large Language Models through Tool-Based Graph Navigation
von: Ghandi, Taraneh, et al.
Veröffentlicht: (2026)
von: Ghandi, Taraneh, et al.
Veröffentlicht: (2026)
A Domain-Agnostic Neurosymbolic Approach for Big Social Data Analysis: Evaluating Mental Health Sentiment on Social Media during COVID-19
von: Khandelwal, Vedant, et al.
Veröffentlicht: (2024)
von: Khandelwal, Vedant, et al.
Veröffentlicht: (2024)
Automated Circuit Interpretation via Probe Prompting
von: Birardi, Giuseppe
Veröffentlicht: (2025)
von: Birardi, Giuseppe
Veröffentlicht: (2025)
A Pluggable Common Sense-Enhanced Framework for Knowledge Graph Completion
von: Niu, Guanglin, et al.
Veröffentlicht: (2024)
von: Niu, Guanglin, et al.
Veröffentlicht: (2024)
The Limits of Obliviate: Evaluating Unlearning in LLMs via Stimulus-Knowledge Entanglement-Behavior Framework
von: Shah, Aakriti, et al.
Veröffentlicht: (2025)
von: Shah, Aakriti, et al.
Veröffentlicht: (2025)
An Explainable Collaborative Dialogue System using a Theory of Mind
von: Cohen, Philip R., et al.
Veröffentlicht: (2023)
von: Cohen, Philip R., et al.
Veröffentlicht: (2023)
Learning What Matters: Probabilistic Task Selection via Mutual Information for Model Finetuning
von: Chanda, Prateek, et al.
Veröffentlicht: (2025)
von: Chanda, Prateek, et al.
Veröffentlicht: (2025)
MeMo: Towards Language Models with Associative Memory Mechanisms
von: Zanzotto, Fabio Massimo, et al.
Veröffentlicht: (2025)
von: Zanzotto, Fabio Massimo, et al.
Veröffentlicht: (2025)
By Their Fruits You Will Know Them: Comparing Formalizations of Law by the Decisions They Encode
von: Vernie, Julius, et al.
Veröffentlicht: (2026)
von: Vernie, Julius, et al.
Veröffentlicht: (2026)
HyperPersona: A Multi-Level Hypergraph Framework for Text-Based Automatic Personality Prediction
von: Heydari, Sina, et al.
Veröffentlicht: (2026)
von: Heydari, Sina, et al.
Veröffentlicht: (2026)
Enabling Transparent Cyber Threat Intelligence Combining Large Language Models and Domain Ontologies
von: Cotti, Luca, et al.
Veröffentlicht: (2025)
von: Cotti, Luca, et al.
Veröffentlicht: (2025)
Attention-based sequential recommendation system using multimodal data
von: Oh, Hyungtaik, et al.
Veröffentlicht: (2024)
von: Oh, Hyungtaik, et al.
Veröffentlicht: (2024)
RefiningGPT: Specialized language Models for Automated Refinery Unit-level Process Diagram Synthesis
von: Liu, Dongxiao, et al.
Veröffentlicht: (2026)
von: Liu, Dongxiao, et al.
Veröffentlicht: (2026)
Do LLMs Truly Understand When a Precedent Is Overruled?
von: Zhang, Li, et al.
Veröffentlicht: (2025)
von: Zhang, Li, et al.
Veröffentlicht: (2025)
Ontology Learning with LLMs: A Benchmark Study on Axiom Identification
von: Bakker, Roos M., et al.
Veröffentlicht: (2025)
von: Bakker, Roos M., et al.
Veröffentlicht: (2025)
Robustness of Large Language Models to Perturbations in Text
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Kastor: Fine-tuned Small Language Models for Shape-based Active Relation Extraction
von: Celian, Ringwald, et al.
Veröffentlicht: (2025) -
Improving LLMs with a knowledge from databases
von: Máša, Petr
Veröffentlicht: (2025) -
Pre-trained Language Model with Prompts for Temporal Knowledge Graph Completion
von: Xu, Wenjie, et al.
Veröffentlicht: (2023) -
From Fake Focus to Real Precision: Confusion-Driven Adversarial Attention Learning in Transformers
von: Liu, Yawei
Veröffentlicht: (2025) -
Unveiling the Secret Recipe: A Guide For Supervised Fine-Tuning Small LLMs
von: Pareja, Aldo, et al.
Veröffentlicht: (2024)