Gespeichert in:
| Hauptverfasser: | Liong, Khai Jiet, Wu, Hongqiu, Zhao, Hai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2402.16470 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Empower Nested Boolean Logic via Self-Supervised Curriculum Learning
von: Wu, Hongqiu, et al.
Veröffentlicht: (2023)
von: Wu, Hongqiu, et al.
Veröffentlicht: (2023)
Chinese Spelling Correction as Rephrasing Language Model
von: Liu, Linfeng, et al.
Veröffentlicht: (2023)
von: Liu, Linfeng, et al.
Veröffentlicht: (2023)
X-TURING: Towards an Enhanced and Efficient Turing Test for Long-Term Dialogue Agents
von: Wu, Weiqi, et al.
Veröffentlicht: (2024)
von: Wu, Weiqi, et al.
Veröffentlicht: (2024)
Game Development as Human-LLM Interaction
von: Hong, Jiale, et al.
Veröffentlicht: (2024)
von: Hong, Jiale, et al.
Veröffentlicht: (2024)
OPEN-THEATRE: An Open-Source Toolkit for LLM-based Interactive Drama
von: Xu, Tianyang, et al.
Veröffentlicht: (2025)
von: Xu, Tianyang, et al.
Veröffentlicht: (2025)
Towards Enhanced Immersion and Agency for LLM-based Interactive Drama
von: Wu, Hongqiu, et al.
Veröffentlicht: (2025)
von: Wu, Hongqiu, et al.
Veröffentlicht: (2025)
BriLLM: Brain-inspired Large Language Model
von: Zhao, Hai, et al.
Veröffentlicht: (2025)
von: Zhao, Hai, et al.
Veröffentlicht: (2025)
A Coin Has Two Sides: A Novel Detector-Corrector Framework for Chinese Spelling Correction
von: Zeng, Xiangke, et al.
Veröffentlicht: (2024)
von: Zeng, Xiangke, et al.
Veröffentlicht: (2024)
From Role-Play to Drama-Interaction: An LLM Solution
von: Wu, Weiqi, et al.
Veröffentlicht: (2024)
von: Wu, Weiqi, et al.
Veröffentlicht: (2024)
Hidden Ghost Hand: Unveiling Backdoor Vulnerabilities in MLLM-Powered Mobile GUI Agents
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2025)
von: Cheng, Pengzhou, et al.
Veröffentlicht: (2025)
Head-wise Shareable Attention for Large Language Models
von: Cao, Zouying, et al.
Veröffentlicht: (2024)
von: Cao, Zouying, et al.
Veröffentlicht: (2024)
From Self-Attention to Markov Models: Unveiling the Dynamics of Generative Transformers
von: Ildiz, M. Emrullah, et al.
Veröffentlicht: (2024)
von: Ildiz, M. Emrullah, et al.
Veröffentlicht: (2024)
Astraea: A State-Aware Scheduling Engine for LLM-Powered Agents
von: Ni, Hongqiu, et al.
Veröffentlicht: (2025)
von: Ni, Hongqiu, et al.
Veröffentlicht: (2025)
IAM: Efficient Inference through Attention Mapping between Different-scale LLMs
von: Zhao, Yi, et al.
Veröffentlicht: (2025)
von: Zhao, Yi, et al.
Veröffentlicht: (2025)
CDER: Collaborative Evidence Retrieval for Document-level Relation Extraction
von: Tran, Khai Phan, et al.
Veröffentlicht: (2025)
von: Tran, Khai Phan, et al.
Veröffentlicht: (2025)
Breach in the Shield: Unveiling the Vulnerabilities of Large Language Models
von: Dai, Runpeng, et al.
Veröffentlicht: (2025)
von: Dai, Runpeng, et al.
Veröffentlicht: (2025)
DAC: A Dynamic Attention-aware Approach for Task-Agnostic Prompt Compression
von: Zhao, Yi, et al.
Veröffentlicht: (2025)
von: Zhao, Yi, et al.
Veröffentlicht: (2025)
Unveiling the Hidden Structure of Self-Attention via Kernel Principal Component Analysis
von: Teo, Rachel S. Y., et al.
Veröffentlicht: (2024)
von: Teo, Rachel S. Y., et al.
Veröffentlicht: (2024)
Intrinsic Model Weaknesses: How Priming Attacks Unveil Vulnerabilities in Large Language Models
von: Huang, Yuyi, et al.
Veröffentlicht: (2025)
von: Huang, Yuyi, et al.
Veröffentlicht: (2025)
VietMed: A Dataset and Benchmark for Automatic Speech Recognition of Vietnamese in the Medical Domain
von: Le-Duc, Khai
Veröffentlicht: (2024)
von: Le-Duc, Khai
Veröffentlicht: (2024)
Unfolding the Headline: Iterative Self-Questioning for News Retrieval and Timeline Summarization
von: Wu, Weiqi, et al.
Veröffentlicht: (2025)
von: Wu, Weiqi, et al.
Veröffentlicht: (2025)
Beyond Text: Unveiling Privacy Vulnerabilities in Multi-modal Retrieval-Augmented Generation
von: Zhang, Jiankun, et al.
Veröffentlicht: (2025)
von: Zhang, Jiankun, et al.
Veröffentlicht: (2025)
Textual-to-Visual Iterative Self-Verification for Slide Generation
von: Xu, Yunqing, et al.
Veröffentlicht: (2025)
von: Xu, Yunqing, et al.
Veröffentlicht: (2025)
Self-Judge: Selective Instruction Following with Alignment Self-Evaluation
von: Ye, Hai, et al.
Veröffentlicht: (2024)
von: Ye, Hai, et al.
Veröffentlicht: (2024)
Unveiling the Magic: Investigating Attention Distillation in Retrieval-augmented Generation
von: Li, Zizhong, et al.
Veröffentlicht: (2024)
von: Li, Zizhong, et al.
Veröffentlicht: (2024)
Unveiling Simplicities of Attention: Adaptive Long-Context Head Identification
von: Donhauser, Konstantin, et al.
Veröffentlicht: (2025)
von: Donhauser, Konstantin, et al.
Veröffentlicht: (2025)
Exclusive Self Attention
von: Zhai, Shuangfei
Veröffentlicht: (2026)
von: Zhai, Shuangfei
Veröffentlicht: (2026)
VaeDiff-DocRE: End-to-end Data Augmentation Framework for Document-level Relation Extraction
von: Tran, Khai Phan, et al.
Veröffentlicht: (2024)
von: Tran, Khai Phan, et al.
Veröffentlicht: (2024)
Anisotropy Is Inherent to Self-Attention in Transformers
von: Godey, Nathan, et al.
Veröffentlicht: (2024)
von: Godey, Nathan, et al.
Veröffentlicht: (2024)
Unveiling and Harnessing Hidden Attention Sinks: Enhancing Large Language Models without Training through Attention Calibration
von: Yu, Zhongzhi, et al.
Veröffentlicht: (2024)
von: Yu, Zhongzhi, et al.
Veröffentlicht: (2024)
Self-Prompting Large Language Models for Zero-Shot Open-Domain QA
von: Li, Junlong, et al.
Veröffentlicht: (2022)
von: Li, Junlong, et al.
Veröffentlicht: (2022)
Vulnerability of Text-to-Image Models to Prompt Template Stealing: A Differential Evolution Approach
von: Wu, Yurong, et al.
Veröffentlicht: (2025)
von: Wu, Yurong, et al.
Veröffentlicht: (2025)
Unveiling Knowledge Utilization Mechanisms in LLM-based Retrieval-Augmented Generation
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
von: Wang, Yuhao, et al.
Veröffentlicht: (2025)
Mitigating Misleading Chain-of-Thought Reasoning with Selective Filtering
von: Wu, Yexin, et al.
Veröffentlicht: (2024)
von: Wu, Yexin, et al.
Veröffentlicht: (2024)
Real-time Speech Summarization for Medical Conversations
von: Le-Duc, Khai, et al.
Veröffentlicht: (2024)
von: Le-Duc, Khai, et al.
Veröffentlicht: (2024)
Does Self-Attention Need Separate Weights in Transformers?
von: Kowsher, Md, et al.
Veröffentlicht: (2024)
von: Kowsher, Md, et al.
Veröffentlicht: (2024)
UniBias: Unveiling and Mitigating LLM Bias through Internal Attention and FFN Manipulation
von: Zhou, Hanzhang, et al.
Veröffentlicht: (2024)
von: Zhou, Hanzhang, et al.
Veröffentlicht: (2024)
Attention-Seeker: Dynamic Self-Attention Scoring for Unsupervised Keyphrase Extraction
von: Z., Erwin D. López, et al.
Veröffentlicht: (2024)
von: Z., Erwin D. López, et al.
Veröffentlicht: (2024)
A Multi-Expert Structural-Semantic Hybrid Framework for Unveiling Historical Patterns in Temporal Knowledge Graphs
von: Deng, Yimin, et al.
Veröffentlicht: (2025)
von: Deng, Yimin, et al.
Veröffentlicht: (2025)
Unveiling Linguistic Regions in Large Language Models
von: Zhang, Zhihao, et al.
Veröffentlicht: (2024)
von: Zhang, Zhihao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Empower Nested Boolean Logic via Self-Supervised Curriculum Learning
von: Wu, Hongqiu, et al.
Veröffentlicht: (2023) -
Chinese Spelling Correction as Rephrasing Language Model
von: Liu, Linfeng, et al.
Veröffentlicht: (2023) -
X-TURING: Towards an Enhanced and Efficient Turing Test for Long-Term Dialogue Agents
von: Wu, Weiqi, et al.
Veröffentlicht: (2024) -
Game Development as Human-LLM Interaction
von: Hong, Jiale, et al.
Veröffentlicht: (2024) -
OPEN-THEATRE: An Open-Source Toolkit for LLM-based Interactive Drama
von: Xu, Tianyang, et al.
Veröffentlicht: (2025)