Dial-MAE: ConTextual Masked Auto-Encoder for Retrieval-based Dialogue Systems
Fuente:
arXiv
Guardado en:
| Autores principales: | Su, Zhenpeng, Wu, Xing, Zhou, Wei, Ma, Guangyuan, Hu, Songlin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
HC3 Plus: A Semantic-Invariant Human ChatGPT Comparison Corpus
por: Su, Zhenpeng, et al.
Publicado: (2023)
por: Su, Zhenpeng, et al.
Publicado: (2023)
LightRetriever: A LLM-based Text Retrieval Architecture with Extremely Faster Query Inference
por: Ma, Guangyuan, et al.
Publicado: (2025)
por: Ma, Guangyuan, et al.
Publicado: (2025)
ConTextual: Improving Clinical Text Summarization in LLMs with Context-preserving Token Filtering and Knowledge Graphs
por: Piya, Fahmida Liza, et al.
Publicado: (2025)
por: Piya, Fahmida Liza, et al.
Publicado: (2025)
The 2nd FutureDial Challenge: Dialog Systems with Retrieval Augmented Generation (FutureDial-RAG)
por: Cai, Yucheng, et al.
Publicado: (2024)
por: Cai, Yucheng, et al.
Publicado: (2024)
ConTextual: Evaluating Context-Sensitive Text-Rich Visual Reasoning in Large Multimodal Models
por: Wadhawan, Rohan, et al.
Publicado: (2024)
por: Wadhawan, Rohan, et al.
Publicado: (2024)
DialSim: A Dialogue Simulator for Evaluating Long-Term Multi-Party Dialogue Understanding of Conversational Agents
por: Kim, Jiho, et al.
Publicado: (2024)
por: Kim, Jiho, et al.
Publicado: (2024)
MaskMoE: Boosting Token-Level Learning via Routing Mask in Mixture-of-Experts
por: Su, Zhenpeng, et al.
Publicado: (2024)
por: Su, Zhenpeng, et al.
Publicado: (2024)
Task-level Distributionally Robust Optimization for Large Language Model-based Dense Retrieval
por: Ma, Guangyuan, et al.
Publicado: (2024)
por: Ma, Guangyuan, et al.
Publicado: (2024)
BERnaT: Basque Encoders for Representing Natural Textual Diversity
por: Azurmendi, Ekhi, et al.
Publicado: (2025)
por: Azurmendi, Ekhi, et al.
Publicado: (2025)
Quest: Query-centric Data Synthesis Approach for Long-context Scaling of Large Language Model
por: Gao, Chaochen, et al.
Publicado: (2024)
por: Gao, Chaochen, et al.
Publicado: (2024)
Drop your Decoder: Pre-training with Bag-of-Word Prediction for Dense Passage Retrieval
por: Ma, Guangyuan, et al.
Publicado: (2024)
por: Ma, Guangyuan, et al.
Publicado: (2024)
ColorMAE: Exploring data-independent masking strategies in Masked AutoEncoders
por: Hinojosa, Carlos, et al.
Publicado: (2024)
por: Hinojosa, Carlos, et al.
Publicado: (2024)
Quasi-symbolic Semantic Geometry over Transformer-based Variational AutoEncoder
por: Zhang, Yingji, et al.
Publicado: (2022)
por: Zhang, Yingji, et al.
Publicado: (2022)
CIKMar: A Dual-Encoder Approach to Prompt-Based Reranking in Educational Dialogue Systems
por: Lopo, Joanito Agili, et al.
Publicado: (2024)
por: Lopo, Joanito Agili, et al.
Publicado: (2024)
HalluSAE: Detecting Hallucinations in Large Language Models via Sparse Auto-Encoders
por: Chen, Boshui, et al.
Publicado: (2026)
por: Chen, Boshui, et al.
Publicado: (2026)
Dial: A Knowledge-Grounded Dialect-Specific NL2SQL System
por: Zhang, Xiang, et al.
Publicado: (2026)
por: Zhang, Xiang, et al.
Publicado: (2026)
Mem2ActBench: A Benchmark for Evaluating Long-Term Memory Utilization in Task-Oriented Autonomous Agents
por: Shen, Yiting, et al.
Publicado: (2026)
por: Shen, Yiting, et al.
Publicado: (2026)
Beyond Imitation: Learning Key Reasoning Steps from Dual Chain-of-Thoughts in Reasoning Distillation
por: Dai, Chengwei, et al.
Publicado: (2024)
por: Dai, Chengwei, et al.
Publicado: (2024)
Improve Student's Reasoning Generalizability through Cascading Decomposed CoTs Distillation
por: Dai, Chengwei, et al.
Publicado: (2024)
por: Dai, Chengwei, et al.
Publicado: (2024)
NExtLong: Toward Effective Long-Context Training without Long Documents
por: Gao, Chaochen, et al.
Publicado: (2025)
por: Gao, Chaochen, et al.
Publicado: (2025)
LongMagpie: A Self-synthesis Method for Generating Large-scale Long-context Instructions
por: Gao, Chaochen, et al.
Publicado: (2025)
por: Gao, Chaochen, et al.
Publicado: (2025)
CodePMP: Scalable Preference Model Pretraining for Large Language Model Reasoning
por: Yu, Huimu, et al.
Publicado: (2024)
por: Yu, Huimu, et al.
Publicado: (2024)
Discrepancy-Aware Graph Mask Auto-Encoder
por: Zheng, Ziyu, et al.
Publicado: (2025)
por: Zheng, Ziyu, et al.
Publicado: (2025)
Sparse Auto-Encoders and Holism about Large Language Models
por: Grindrod, Jumbly
Publicado: (2026)
por: Grindrod, Jumbly
Publicado: (2026)
SafeDialBench: A Fine-Grained Safety Evaluation Benchmark for Large Language Models in Multi-Turn Dialogues with Diverse Jailbreak Attacks
por: Cao, Hongye, et al.
Publicado: (2025)
por: Cao, Hongye, et al.
Publicado: (2025)
BiCon-Gate: Consistency-Gated De-colloquialisation for Dialogue Fact-Checking
por: Park, Hyunkyung, et al.
Publicado: (2026)
por: Park, Hyunkyung, et al.
Publicado: (2026)
Transferring Structure Knowledge: A New Task to Fake news Detection Towards Cold-Start Propagation
por: Wei, Lingwei, et al.
Publicado: (2024)
por: Wei, Lingwei, et al.
Publicado: (2024)
Domain-Adapted Retrieval for In-Context Annotation of Pedagogical Dialogue Acts
por: Lee, Jinsook, et al.
Publicado: (2026)
por: Lee, Jinsook, et al.
Publicado: (2026)
Conceptual Contrastive Edits in Textual and Vision-Language Retrieval
por: Lymperaiou, Maria, et al.
Publicado: (2025)
por: Lymperaiou, Maria, et al.
Publicado: (2025)
Exploring Gradient-Guided Masked Language Model to Detect Textual Adversarial Attacks
por: Zhang, Xiaomei, et al.
Publicado: (2025)
por: Zhang, Xiaomei, et al.
Publicado: (2025)
Textual Self-attention Network: Test-Time Preference Optimization through Textual Gradient-based Attention
por: Mo, Shibing, et al.
Publicado: (2025)
por: Mo, Shibing, et al.
Publicado: (2025)
Diagnosing Retrieval Bias Under Multiple In-Context Knowledge Updates in Large Language Models
por: Qiao, Boyu, et al.
Publicado: (2026)
por: Qiao, Boyu, et al.
Publicado: (2026)
MedDialBench: Benchmarking LLM Diagnostic Robustness under Parametric Adversarial Patient Behaviors
por: Luo, Xiaotian, et al.
Publicado: (2026)
por: Luo, Xiaotian, et al.
Publicado: (2026)
DuConTE: Dual-Granularity Text Encoder with Topology-Constrained Attention for Text-attributed Graphs
por: Liang, Lexuan, et al.
Publicado: (2026)
por: Liang, Lexuan, et al.
Publicado: (2026)
Fine-Grained Behavior Simulation with Role-Playing Large Language Model on Social Media
por: Li, Kun, et al.
Publicado: (2024)
por: Li, Kun, et al.
Publicado: (2024)
Perspective Dial: Measuring Perspective of Text and Guiding LLM Outputs
por: Kim, Taejin, et al.
Publicado: (2025)
por: Kim, Taejin, et al.
Publicado: (2025)
"In Dialogues We Learn": Towards Personalized Dialogue Without Pre-defined Profiles through In-Dialogue Learning
por: Cheng, Chuanqi, et al.
Publicado: (2024)
por: Cheng, Chuanqi, et al.
Publicado: (2024)
NeRF-MAE: Masked AutoEncoders for Self-Supervised 3D Representation Learning for Neural Radiance Fields
por: Irshad, Muhammad Zubair, et al.
Publicado: (2024)
por: Irshad, Muhammad Zubair, et al.
Publicado: (2024)
Layer-Wise Evolution of Representations in Fine-Tuned Transformers: Insights from Sparse AutoEncoders
por: Nadipalli, Suneel
Publicado: (2025)
por: Nadipalli, Suneel
Publicado: (2025)
Enhancing Retrieval and Managing Retrieval: A Four-Module Synergy for Improved Quality and Efficiency in RAG Systems
por: Shi, Yunxiao, et al.
Publicado: (2024)
por: Shi, Yunxiao, et al.
Publicado: (2024)
Ejemplares similares
-
HC3 Plus: A Semantic-Invariant Human ChatGPT Comparison Corpus
por: Su, Zhenpeng, et al.
Publicado: (2023) -
LightRetriever: A LLM-based Text Retrieval Architecture with Extremely Faster Query Inference
por: Ma, Guangyuan, et al.
Publicado: (2025) -
ConTextual: Improving Clinical Text Summarization in LLMs with Context-preserving Token Filtering and Knowledge Graphs
por: Piya, Fahmida Liza, et al.
Publicado: (2025) -
The 2nd FutureDial Challenge: Dialog Systems with Retrieval Augmented Generation (FutureDial-RAG)
por: Cai, Yucheng, et al.
Publicado: (2024) -
ConTextual: Evaluating Context-Sensitive Text-Rich Visual Reasoning in Large Multimodal Models
por: Wadhawan, Rohan, et al.
Publicado: (2024)