Is There No Such Thing as a Bad Question? H4R: HalluciBot For Ratiocination, Rewriting, Ranking, and Routing
Fuente:
arXiv
Salvato in:
| Autori principali: | Watson, William, Cho, Nicole, Srishankar, Nishan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FISHNET: Financial Intelligence from Sub-querying, Harmonizing, Neural-Conditioning, Expert Swarms, and Task Planning
di: Cho, Nicole, et al.
Pubblicazione: (2024)
di: Cho, Nicole, et al.
Pubblicazione: (2024)
MultiQ&A: An Analysis in Measuring Robustness via Automated Crowdsourcing of Question Perturbations and Answers
di: Cho, Nicole, et al.
Pubblicazione: (2025)
di: Cho, Nicole, et al.
Pubblicazione: (2025)
QueryBandits for Hallucination Mitigation: Exploiting Semantic Features for No-Regret Rewriting
di: Cho, Nicole, et al.
Pubblicazione: (2025)
di: Cho, Nicole, et al.
Pubblicazione: (2025)
AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations
di: Verma, Gaurav, et al.
Pubblicazione: (2024)
di: Verma, Gaurav, et al.
Pubblicazione: (2024)
R-Bot: An LLM-based Query Rewrite System
di: Sun, Zhaoyan, et al.
Pubblicazione: (2024)
di: Sun, Zhaoyan, et al.
Pubblicazione: (2024)
No One Size Fits All: QueryBandits for Hallucination Mitigation
di: Cho, Nicole, et al.
Pubblicazione: (2026)
di: Cho, Nicole, et al.
Pubblicazione: (2026)
Grounded Relational Inference: Domain Knowledge Driven Explainable Autonomous Driving
di: Tang, Chen, et al.
Pubblicazione: (2021)
di: Tang, Chen, et al.
Pubblicazione: (2021)
LAW: Legal Agentic Workflows for Custody and Fund Services Contracts
di: Watson, William, et al.
Pubblicazione: (2024)
di: Watson, William, et al.
Pubblicazione: (2024)
SpectR: Dynamically Composing LM Experts with Spectral Routing
di: Fleshman, William, et al.
Pubblicazione: (2025)
di: Fleshman, William, et al.
Pubblicazione: (2025)
TASER: Table Agents for Schema-guided Extraction and Recommendation
di: Cho, Nicole, et al.
Pubblicazione: (2025)
di: Cho, Nicole, et al.
Pubblicazione: (2025)
ChartAgent: A Multimodal Agent for Visually Grounded Reasoning in Complex Chart Question Answering
di: Kaur, Rachneet, et al.
Pubblicazione: (2025)
di: Kaur, Rachneet, et al.
Pubblicazione: (2025)
MoR: Mixture of Ranks for Low-Rank Adaptation Tuning
di: Tang, Chuanyu, et al.
Pubblicazione: (2024)
di: Tang, Chuanyu, et al.
Pubblicazione: (2024)
PRewrite: Prompt Rewriting with Reinforcement Learning
di: Kong, Weize, et al.
Pubblicazione: (2024)
di: Kong, Weize, et al.
Pubblicazione: (2024)
On the Limitations of Rank-One Model Editing in Answering Multi-hop Questions
di: He, Zhiyuan, et al.
Pubblicazione: (2026)
di: He, Zhiyuan, et al.
Pubblicazione: (2026)
Short-form Text Rewriting with Phi Silica
di: Tadimeti, Divya, et al.
Pubblicazione: (2026)
di: Tadimeti, Divya, et al.
Pubblicazione: (2026)
Hypothesis-Conditioned Query Rewriting for Decision-Useful Retrieval
di: Chang, Hangeol, et al.
Pubblicazione: (2026)
di: Chang, Hangeol, et al.
Pubblicazione: (2026)
When Bad Data Leads to Good Models
di: Li, Kenneth, et al.
Pubblicazione: (2025)
di: Li, Kenneth, et al.
Pubblicazione: (2025)
FADE: Why Bad Descriptions Happen to Good Features
di: Puri, Bruno, et al.
Pubblicazione: (2025)
di: Puri, Bruno, et al.
Pubblicazione: (2025)
SEQR: Secure and Efficient QR-based LoRA Routing
di: Fleshman, William, et al.
Pubblicazione: (2025)
di: Fleshman, William, et al.
Pubblicazione: (2025)
Evidence-Focused Fact Summarization for Knowledge-Augmented Zero-Shot Question Answering
di: Ko, Sungho, et al.
Pubblicazione: (2024)
di: Ko, Sungho, et al.
Pubblicazione: (2024)
Preference Learning Algorithms Do Not Learn Preference Rankings
di: Chen, Angelica, et al.
Pubblicazione: (2024)
di: Chen, Angelica, et al.
Pubblicazione: (2024)
Sample-Efficient Online Learning in LM Agents via Hindsight Trajectory Rewriting
di: Hu, Michael Y., et al.
Pubblicazione: (2025)
di: Hu, Michael Y., et al.
Pubblicazione: (2025)
Predicting Compact Phrasal Rewrites with Large Language Models for ASR Post Editing
di: Zhang, Hao, et al.
Pubblicazione: (2025)
di: Zhang, Hao, et al.
Pubblicazione: (2025)
Assigning Distinct Roles to Quantized and Low-Rank Matrices Toward Optimal Weight Decomposition
di: Cho, Yoonjun, et al.
Pubblicazione: (2025)
di: Cho, Yoonjun, et al.
Pubblicazione: (2025)
Toxicity Detection Should Measure Contextual Harm, Not Text-Intrinsic Badness
di: Berezin, Sergei, et al.
Pubblicazione: (2025)
di: Berezin, Sergei, et al.
Pubblicazione: (2025)
Simulation, Modelling and Classification of Wiki Contributors: Spotting The Good, The Bad, and The Ugly
di: Méndez, Silvia García, et al.
Pubblicazione: (2024)
di: Méndez, Silvia García, et al.
Pubblicazione: (2024)
Model Merging and Safety Alignment: One Bad Model Spoils the Bunch
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2024)
di: Hammoud, Hasan Abed Al Kader, et al.
Pubblicazione: (2024)
Dr Genre: Reinforcement Learning from Decoupled LLM Feedback for Generic Text Rewriting
di: Li, Yufei, et al.
Pubblicazione: (2025)
di: Li, Yufei, et al.
Pubblicazione: (2025)
AgentCollabBench: Diagnosing When Good Agents Make Bad Collaborators
di: Mazumder, Aritra, et al.
Pubblicazione: (2026)
di: Mazumder, Aritra, et al.
Pubblicazione: (2026)
Router-R1: Teaching LLMs Multi-Round Routing and Aggregation via Reinforcement Learning
di: Zhang, Haozhen, et al.
Pubblicazione: (2025)
di: Zhang, Haozhen, et al.
Pubblicazione: (2025)
RouteLLM: Learning to Route LLMs with Preference Data
di: Ong, Isaac, et al.
Pubblicazione: (2024)
di: Ong, Isaac, et al.
Pubblicazione: (2024)
How Bad is Training on Synthetic Data? A Statistical Analysis of Language Model Collapse
di: Seddik, Mohamed El Amine, et al.
Pubblicazione: (2024)
di: Seddik, Mohamed El Amine, et al.
Pubblicazione: (2024)
SEMQA: Semi-Extractive Multi-Source Question Answering
di: Schuster, Tal, et al.
Pubblicazione: (2023)
di: Schuster, Tal, et al.
Pubblicazione: (2023)
Are Hallucinations Bad Estimations?
di: Liu, Hude, et al.
Pubblicazione: (2025)
di: Liu, Hude, et al.
Pubblicazione: (2025)
Dynamic Latent Routing
di: Yu, Fangyuan, et al.
Pubblicazione: (2026)
di: Yu, Fangyuan, et al.
Pubblicazione: (2026)
Large Language Models Badly Generalize across Option Length, Problem Types, and Irrelevant Noun Replacements
di: Zhao, Guangxiang, et al.
Pubblicazione: (2025)
di: Zhao, Guangxiang, et al.
Pubblicazione: (2025)
Comparing Bad Apples to Good Oranges: Aligning Large Language Models via Joint Preference Optimization
di: Bansal, Hritik, et al.
Pubblicazione: (2024)
di: Bansal, Hritik, et al.
Pubblicazione: (2024)
SynapseRoute: An Auto-Route Switching Framework on Dual-State Large Language Model
di: Zhang, Wencheng, et al.
Pubblicazione: (2025)
di: Zhang, Wencheng, et al.
Pubblicazione: (2025)
JuStRank: Benchmarking LLM Judges for System Ranking
di: Gera, Ariel, et al.
Pubblicazione: (2024)
di: Gera, Ariel, et al.
Pubblicazione: (2024)
Large Language Models Are Bad Dice Players: LLMs Struggle to Generate Random Numbers from Statistical Distributions
di: Zhao, Minda, et al.
Pubblicazione: (2026)
di: Zhao, Minda, et al.
Pubblicazione: (2026)
Documenti analoghi
-
FISHNET: Financial Intelligence from Sub-querying, Harmonizing, Neural-Conditioning, Expert Swarms, and Task Planning
di: Cho, Nicole, et al.
Pubblicazione: (2024) -
MultiQ&A: An Analysis in Measuring Robustness via Automated Crowdsourcing of Question Perturbations and Answers
di: Cho, Nicole, et al.
Pubblicazione: (2025) -
QueryBandits for Hallucination Mitigation: Exploiting Semantic Features for No-Regret Rewriting
di: Cho, Nicole, et al.
Pubblicazione: (2025) -
AdaptAgent: Adapting Multimodal Web Agents with Few-Shot Learning from Human Demonstrations
di: Verma, Gaurav, et al.
Pubblicazione: (2024) -
R-Bot: An LLM-based Query Rewrite System
di: Sun, Zhaoyan, et al.
Pubblicazione: (2024)