Evaluating Reliability Asymmetries in Chinese Factual Search and AI Answers
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Geng, Feng, Li, Zhu, Mengxiao, Pierri, Francesco |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Search Engines in an AI Era: The False Promise of Factual and Verifiable Source-Cited Responses
by: Venkit, Pranav Narayanan, et al.
Published: (2024)
by: Venkit, Pranav Narayanan, et al.
Published: (2024)
Evaluating Robustness of Generative Search Engine on Adversarial Factual Questions
by: Hu, Xuming, et al.
Published: (2024)
by: Hu, Xuming, et al.
Published: (2024)
PaperAsk: A Benchmark for Reliability Evaluation of LLMs in Paper Search and Reading
by: Wu, Yutao, et al.
Published: (2025)
by: Wu, Yutao, et al.
Published: (2025)
How Generative AI Disrupts Search: An Empirical Study of Google Search, Gemini, and AI Overviews
by: Grossman, Riley, et al.
Published: (2026)
by: Grossman, Riley, et al.
Published: (2026)
Retrieval Improvements Do Not Guarantee Better Answers: A Study of RAG for AI Policy QA
by: Mathur, Saahil, et al.
Published: (2026)
by: Mathur, Saahil, et al.
Published: (2026)
Stop DDoS Attacking the Research Community with AI-Generated Survey Papers
by: Lin, Jianghao, et al.
Published: (2025)
by: Lin, Jianghao, et al.
Published: (2025)
Taxonomy and Analysis of Sensitive User Queries in Generative AI Search
by: Jo, Hwiyeol, et al.
Published: (2024)
by: Jo, Hwiyeol, et al.
Published: (2024)
Search Engines Post-ChatGPT: How Generative Artificial Intelligence Could Make Search Less Reliable
by: Memon, Shahan Ali, et al.
Published: (2024)
by: Memon, Shahan Ali, et al.
Published: (2024)
Large Language Model-Based Knowledge Graph System Construction for Sustainable Development Goals: An AI-Based Speculative Design Perspective
by: Lin, Yi-De, et al.
Published: (2025)
by: Lin, Yi-De, et al.
Published: (2025)
Matching Meaning at Scale: Evaluating Semantic Search for 18th-Century Intellectual History through the Case of Locke
by: Wu, Yu, et al.
Published: (2026)
by: Wu, Yu, et al.
Published: (2026)
Document-Level Event Extraction with Definition-Driven ICL
by: Liu, Zhuoyuan, et al.
Published: (2024)
by: Liu, Zhuoyuan, et al.
Published: (2024)
Suicide Phenotyping from Clinical Notes in Safety-Net Psychiatric Hospital Using Multi-Label Classification with Pre-Trained Language Models
by: Li, Zehan, et al.
Published: (2024)
by: Li, Zehan, et al.
Published: (2024)
Model-Document Protocol for AI Search
by: Qian, Hongjin, et al.
Published: (2025)
by: Qian, Hongjin, et al.
Published: (2025)
Reliable Answers for Recurring Questions: Boosting Text-to-SQL Accuracy with Template Constrained Decoding
by: Jivani, Smit, et al.
Published: (2026)
by: Jivani, Smit, et al.
Published: (2026)
Towards AI Search Paradigm
by: Li, Yuchen, et al.
Published: (2025)
by: Li, Yuchen, et al.
Published: (2025)
TURA: Tool-Augmented Unified Retrieval Agent for AI Search
by: Zhao, Zhejun, et al.
Published: (2025)
by: Zhao, Zhejun, et al.
Published: (2025)
Reliable Evaluation Protocol for Low-Precision Retrieval
by: Yang, Kisu, et al.
Published: (2025)
by: Yang, Kisu, et al.
Published: (2025)
Global-Liar: Factuality of LLMs over Time and Geographic Regions
by: Mirza, Shujaat, et al.
Published: (2024)
by: Mirza, Shujaat, et al.
Published: (2024)
From Facts to Conclusions : Integrating Deductive Reasoning in Retrieval-Augmented LLMs
by: Mishra, Shubham, et al.
Published: (2025)
by: Mishra, Shubham, et al.
Published: (2025)
Stairway to Fairness: Connecting Group and Individual Fairness
by: Rampisela, Theresia Veronika, et al.
Published: (2025)
by: Rampisela, Theresia Veronika, et al.
Published: (2025)
Scraping the Shadows: Deep Learning Breakthroughs in Dark Web Intelligence
by: Bakermans, Ingmar, et al.
Published: (2025)
by: Bakermans, Ingmar, et al.
Published: (2025)
SynDy: Synthetic Dynamic Dataset Generation Framework for Misinformation Tasks
by: Shliselberg, Michael, et al.
Published: (2024)
by: Shliselberg, Michael, et al.
Published: (2024)
PatentEdits: Framing Patent Novelty as Textual Entailment
by: Lee, Ryan, et al.
Published: (2024)
by: Lee, Ryan, et al.
Published: (2024)
Inducing Sustained Creativity and Diversity in Large Language Models
by: Luo, Queenie, et al.
Published: (2026)
by: Luo, Queenie, et al.
Published: (2026)
Level-Navi Agent: A Framework and benchmark for Chinese Web Search Agents
by: Hu, Chuanrui, et al.
Published: (2024)
by: Hu, Chuanrui, et al.
Published: (2024)
EncouRAGe: Evaluating RAG Local, Fast, and Reliable
by: Strich, Jan, et al.
Published: (2025)
by: Strich, Jan, et al.
Published: (2025)
Beyond Factual Grounding: The Case for Opinion-Aware Retrieval-Augmented Generation
by: Agrawal, Aditya, et al.
Published: (2026)
by: Agrawal, Aditya, et al.
Published: (2026)
A Survey of LLM-based Deep Search Agents: Paradigm, Optimization, Evaluation, and Challenges
by: Xi, Yunjia, et al.
Published: (2025)
by: Xi, Yunjia, et al.
Published: (2025)
OASES: Outcome-Aligned Search-Evaluation Co-Training for Agentic Search
by: Zhang, Erhan, et al.
Published: (2026)
by: Zhang, Erhan, et al.
Published: (2026)
Superplatforms Have to Attack AI Agents
by: Lin, Jianghao, et al.
Published: (2025)
by: Lin, Jianghao, et al.
Published: (2025)
Evaluation of LLMs for Process Model Analysis and Optimization
by: Kumar, Akhil, et al.
Published: (2025)
by: Kumar, Akhil, et al.
Published: (2025)
The Quest for Reliable Metrics of Responsible AI
by: Rampisela, Theresia Veronika, et al.
Published: (2025)
by: Rampisela, Theresia Veronika, et al.
Published: (2025)
Talking the Talk Does Not Entail Walking the Walk: On the Limits of Large Language Models in Lexical Entailment Recognition
by: Greco, Candida M., et al.
Published: (2024)
by: Greco, Candida M., et al.
Published: (2024)
Search-o1: Agentic Search-Enhanced Large Reasoning Models
by: Li, Xiaoxi, et al.
Published: (2025)
by: Li, Xiaoxi, et al.
Published: (2025)
RAVine: Reality-Aligned Evaluation for Agentic Search
by: Xu, Yilong, et al.
Published: (2025)
by: Xu, Yilong, et al.
Published: (2025)
Evaluating LLMs' Assessment of Mixed-Context Hallucination Through the Lens of Summarization
by: Qi, Siya, et al.
Published: (2025)
by: Qi, Siya, et al.
Published: (2025)
Same Question, Different Source, Different Answer: Auditing Source-Dependence in Medical Multi-Source RAG
by: Li, Yubo, et al.
Published: (2026)
by: Li, Yubo, et al.
Published: (2026)
Take Care of Your Prompt Bias! Investigating and Mitigating Prompt Bias in Factual Knowledge Extraction
by: Xu, Ziyang, et al.
Published: (2024)
by: Xu, Ziyang, et al.
Published: (2024)
Beyond Catalogue Counts: the Dataset Visibility Asymmetry in Low-Resource Multilingual NLP
by: Tan, Zhiyin, et al.
Published: (2026)
by: Tan, Zhiyin, et al.
Published: (2026)
Auditing Google's AI Overviews and Featured Snippets: A Case Study on Baby Care and Pregnancy
by: Hu, Desheng, et al.
Published: (2025)
by: Hu, Desheng, et al.
Published: (2025)
Similar Items
-
Search Engines in an AI Era: The False Promise of Factual and Verifiable Source-Cited Responses
by: Venkit, Pranav Narayanan, et al.
Published: (2024) -
Evaluating Robustness of Generative Search Engine on Adversarial Factual Questions
by: Hu, Xuming, et al.
Published: (2024) -
PaperAsk: A Benchmark for Reliability Evaluation of LLMs in Paper Search and Reading
by: Wu, Yutao, et al.
Published: (2025) -
How Generative AI Disrupts Search: An Empirical Study of Google Search, Gemini, and AI Overviews
by: Grossman, Riley, et al.
Published: (2026) -
Retrieval Improvements Do Not Guarantee Better Answers: A Study of RAG for AI Policy QA
by: Mathur, Saahil, et al.
Published: (2026)