DMA: Online RAG Alignment with Human Feedback
Fuente:
arXiv
Guardado en:
| Autores principales: | Bai, Yu, Miao, Yukai, Wang, Dawei, Chen, Li, Long, Fei, Zhai, Rundi, Li, Dan, Ren, Yanyu, Liu, Tianfeng, Xie, Hongtao, Yang, Ce, Cai, Xuhui |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Pistis-RAG: Enhancing Retrieval-Augmented Generation with Human Feedback
por: Bai, Yu, et al.
Publicado: (2024)
por: Bai, Yu, et al.
Publicado: (2024)
ProphetFuzz: Fully Automated Prediction and Fuzzing of High-Risk Option Combinations with Only Documentation via Large Language Model
por: Wang, Dawei, et al.
Publicado: (2024)
por: Wang, Dawei, et al.
Publicado: (2024)
PA-RAG: RAG Alignment via Multi-Perspective Preference Optimization
por: Wu, Jiayi, et al.
Publicado: (2024)
por: Wu, Jiayi, et al.
Publicado: (2024)
6DMA-Enabled ISAC for Low-Altitude Economy
por: Jiao, Yingchao, et al.
Publicado: (2026)
por: Jiao, Yingchao, et al.
Publicado: (2026)
Safe MPC Alignment with Human Directional Feedback
por: Xie, Zhixian, et al.
Publicado: (2024)
por: Xie, Zhixian, et al.
Publicado: (2024)
EntWorld: A Holistic Environment and Benchmark for Verifiable Enterprise GUI Agents
por: Mo, Ying, et al.
Publicado: (2026)
por: Mo, Ying, et al.
Publicado: (2026)
6DMA-Assisted Secure Wireless Communications
por: Qian, Yanzhi, et al.
Publicado: (2025)
por: Qian, Yanzhi, et al.
Publicado: (2025)
Understanding the Learning Dynamics of Alignment with Human Feedback
por: Im, Shawn, et al.
Publicado: (2024)
por: Im, Shawn, et al.
Publicado: (2024)
RaFe: Ranking Feedback Improves Query Rewriting for RAG
por: Mao, Shengyu, et al.
Publicado: (2024)
por: Mao, Shengyu, et al.
Publicado: (2024)
Uncertainty-Penalized Reinforcement Learning from Human Feedback with Diverse Reward LoRA Ensembles
por: Zhai, Yuanzhao, et al.
Publicado: (2023)
por: Zhai, Yuanzhao, et al.
Publicado: (2023)
DMA-Latte: Expanding the Reach of DMA Offloads to Latency-bound ML Communication
por: Pati, Suchita, et al.
Publicado: (2025)
por: Pati, Suchita, et al.
Publicado: (2025)
Preference Ranking Optimization for Human Alignment
por: Song, Feifan, et al.
Publicado: (2023)
por: Song, Feifan, et al.
Publicado: (2023)
Adapt as You Say: Online Interactive Bimanual Skill Adaptation via Human Language Feedback
por: Li, Zhuo, et al.
Publicado: (2026)
por: Li, Zhuo, et al.
Publicado: (2026)
VideoRAG: Retrieval-Augmented Generation with Extreme Long-Context Videos
por: Ren, Xubin, et al.
Publicado: (2025)
por: Ren, Xubin, et al.
Publicado: (2025)
Towards Efficient Online Exploration for Reinforcement Learning with Human Feedback
por: Li, Gen, et al.
Publicado: (2025)
por: Li, Gen, et al.
Publicado: (2025)
SmartRAG: Jointly Learn RAG-Related Tasks From the Environment Feedback
por: Gao, Jingsheng, et al.
Publicado: (2024)
por: Gao, Jingsheng, et al.
Publicado: (2024)
RegionRAG: Region-level Retrieval-Augmented Generation for Visual Document Understanding
por: Li, Yinglu, et al.
Publicado: (2025)
por: Li, Yinglu, et al.
Publicado: (2025)
Online-Optimized RAG for Tool Use and Function Calling
por: Pan, Yu, et al.
Publicado: (2025)
por: Pan, Yu, et al.
Publicado: (2025)
Feedback-Driven Vision-Language Alignment with Minimal Human Supervision
por: Giannone, Giorgio, et al.
Publicado: (2025)
por: Giannone, Giorgio, et al.
Publicado: (2025)
Locality Implies Complex Numbers in Quantum Mechanics
por: Feng, Tianfeng, et al.
Publicado: (2025)
por: Feng, Tianfeng, et al.
Publicado: (2025)
Flexible-Sector 6DMA Base Station: Modeling and Design
por: Li, Yunli, et al.
Publicado: (2025)
por: Li, Yunli, et al.
Publicado: (2025)
Throughput Maximization for Multiuser Communications with Flexible-Sector 6DMA
por: Shi, Xiaoming, et al.
Publicado: (2025)
por: Shi, Xiaoming, et al.
Publicado: (2025)
Out-of-Distribution Learning with Human Feedback
por: Bai, Haoyue, et al.
Publicado: (2024)
por: Bai, Haoyue, et al.
Publicado: (2024)
Quantum Correlation Sharing: A Review On Recent Progress From Nonlocality To Other Non-Classical Correlations
por: Cai, Zinuo, et al.
Publicado: (2024)
por: Cai, Zinuo, et al.
Publicado: (2024)
An Improved Synthesis of Ionizable Lipids D‐Lin‐KC2‐DMA and D‐Lin‐MC3‐DMA
por: Ardalan Nabi, et al.
Publicado: (2025)
por: Ardalan Nabi, et al.
Publicado: (2025)
Spatiotemporally Consistent Indoor Lighting Estimation with Diffusion Priors
por: Tong, Mutian, et al.
Publicado: (2025)
por: Tong, Mutian, et al.
Publicado: (2025)
Seeing is Improving: Visual Feedback for Iterative Text Layout Refinement
por: Guo, Junrong, et al.
Publicado: (2026)
por: Guo, Junrong, et al.
Publicado: (2026)
Distinguishing Backward Volume Magnetostatic Spin Wave Vectors via the Spin Wave Doppler Effect
por: Su, Xuhui, et al.
Publicado: (2024)
por: Su, Xuhui, et al.
Publicado: (2024)
LiFT: Leveraging Human Feedback for Text-to-Video Model Alignment
por: Wang, Yibin, et al.
Publicado: (2024)
por: Wang, Yibin, et al.
Publicado: (2024)
RLTHF: Targeted Human Feedback for LLM Alignment
por: Xu, Yifei, et al.
Publicado: (2025)
por: Xu, Yifei, et al.
Publicado: (2025)
Axioms for AI Alignment from Human Feedback
por: Ge, Luise, et al.
Publicado: (2024)
por: Ge, Luise, et al.
Publicado: (2024)
Deep return of straw improves topsoil health and maize productivity more efficiently than shallow return
por: Jin Li, et al.
Publicado: (2024)
por: Jin Li, et al.
Publicado: (2024)
FG-CLIP: Fine-Grained Visual and Textual Alignment
por: Xie, Chunyu, et al.
Publicado: (2025)
por: Xie, Chunyu, et al.
Publicado: (2025)
OnlineHOI: Towards Online Human-Object Interaction Generation and Perception
por: Ji, Yihong, et al.
Publicado: (2025)
por: Ji, Yihong, et al.
Publicado: (2025)
Purify-then-Align: Towards Robust Human Sensing under Modality Missing with Knowledge Distillation from Noisy Multimodal Teacher
por: Weng, Pengcheng, et al.
Publicado: (2026)
por: Weng, Pengcheng, et al.
Publicado: (2026)
Unsupervised Large Language Model Alignment for Information Retrieval via Contrastive Feedback
por: Dong, Qian, et al.
Publicado: (2023)
por: Dong, Qian, et al.
Publicado: (2023)
NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback
por: Gaikwad, Madhava, et al.
Publicado: (2025)
por: Gaikwad, Madhava, et al.
Publicado: (2025)
HumanVLM: Foundation for Human-Scene Vision-Language Model
por: Dai, Dawei, et al.
Publicado: (2024)
por: Dai, Dawei, et al.
Publicado: (2024)
Direct Language Model Alignment from Online AI Feedback
por: Guo, Shangmin, et al.
Publicado: (2024)
por: Guo, Shangmin, et al.
Publicado: (2024)
SAIL: Self-Amplified Iterative Learning for Diffusion Model Alignment with Minimal Human Feedback
por: He, Xiaoxuan, et al.
Publicado: (2026)
por: He, Xiaoxuan, et al.
Publicado: (2026)
Ejemplares similares
-
Pistis-RAG: Enhancing Retrieval-Augmented Generation with Human Feedback
por: Bai, Yu, et al.
Publicado: (2024) -
ProphetFuzz: Fully Automated Prediction and Fuzzing of High-Risk Option Combinations with Only Documentation via Large Language Model
por: Wang, Dawei, et al.
Publicado: (2024) -
PA-RAG: RAG Alignment via Multi-Perspective Preference Optimization
por: Wu, Jiayi, et al.
Publicado: (2024) -
6DMA-Enabled ISAC for Low-Altitude Economy
por: Jiao, Yingchao, et al.
Publicado: (2026) -
Safe MPC Alignment with Human Directional Feedback
por: Xie, Zhixian, et al.
Publicado: (2024)