Verifying the Robustness of Automatic Credibility Assessment
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Przybyła, Piotr, Shvets, Alexander, Saggion, Horacio |
|---|---|
| Format: | Preprint |
| Publié: |
2023
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Attacking Misinformation Detection Using Adversarial Examples Generated by Language Models
par: Przybyła, Piotr, et autres
Publié: (2024)
par: Przybyła, Piotr, et autres
Publié: (2024)
PolQA: Polish Question Answering Dataset
par: Rybak, Piotr, et autres
Publié: (2022)
par: Rybak, Piotr, et autres
Publié: (2022)
Deanthropomorphising NLP: Can a Language Model Be Conscious?
par: Shardlow, Matthew, et autres
Publié: (2022)
par: Shardlow, Matthew, et autres
Publié: (2022)
CrEst: Credibility Estimation for Contexts in LLMs via Weak Supervision
par: Adila, Dyah, et autres
Publié: (2025)
par: Adila, Dyah, et autres
Publié: (2025)
Towards Trustworthy Lexical Simplification: Exploring Safety and Efficiency with Small LLMs
par: Hayakawa, Akio, et autres
Publié: (2025)
par: Hayakawa, Akio, et autres
Publié: (2025)
ERBench: An Entity-Relationship based Automatically Verifiable Hallucination Benchmark for Large Language Models
par: Oh, Jio, et autres
Publié: (2024)
par: Oh, Jio, et autres
Publié: (2024)
Continuously Learning New Words in Automatic Speech Recognition
par: Huber, Christian, et autres
Publié: (2024)
par: Huber, Christian, et autres
Publié: (2024)
Confidence-Credibility Aware Weighted Ensembles of Small LLMs Outperform Large LLMs in Emotion Detection
par: Elgabry, Menna, et autres
Publié: (2025)
par: Elgabry, Menna, et autres
Publié: (2025)
Context Biasing for Pronunciation-Orthography Mismatch in Automatic Speech Recognition
par: Huber, Christian, et autres
Publié: (2025)
par: Huber, Christian, et autres
Publié: (2025)
AutoPyVerifier: Learning Compact Executable Verifiers for Large Language Model Outputs
par: Pezeshkpour, Pouya, et autres
Publié: (2026)
par: Pezeshkpour, Pouya, et autres
Publié: (2026)
Exploiting LLMs for Automatic Hypothesis Assessment via a Logit-Based Calibrated Prior
par: Gong, Yue, et autres
Publié: (2025)
par: Gong, Yue, et autres
Publié: (2025)
Weakly Supervised Veracity Classification with LLM-Predicted Credibility Signals
par: Leite, João A., et autres
Publié: (2023)
par: Leite, João A., et autres
Publié: (2023)
VerifierQ: Enhancing LLM Test Time Compute with Q-Learning-based Verifiers
par: Qi, Jianing, et autres
Publié: (2024)
par: Qi, Jianing, et autres
Publié: (2024)
Verifying the Verifiers: Unveiling Pitfalls and Potentials in Fact Verifiers
par: Seo, Wooseok, et autres
Publié: (2025)
par: Seo, Wooseok, et autres
Publié: (2025)
Q-NL Verifier: Leveraging Synthetic Data for Robust Knowledge Graph Question Answering
par: Schwabe, Tim, et autres
Publié: (2025)
par: Schwabe, Tim, et autres
Publié: (2025)
On the Ability of Transformers to Verify Plans
par: Sarrof, Yash, et autres
Publié: (2026)
par: Sarrof, Yash, et autres
Publié: (2026)
Investigating Automatic Scoring and Feedback using Large Language Models
par: Katuka, Gloria Ashiya, et autres
Publié: (2024)
par: Katuka, Gloria Ashiya, et autres
Publié: (2024)
SCI-Verifier: Scientific Verifier with Thinking
par: Zheng, Shenghe, et autres
Publié: (2025)
par: Zheng, Shenghe, et autres
Publié: (2025)
Reinforcing General Reasoning without Verifiers
par: Zhou, Xiangxin, et autres
Publié: (2025)
par: Zhou, Xiangxin, et autres
Publié: (2025)
CSP-Atlas: Concept-Specific Neural Circuits in a Sparse Python Transformer
par: Wilam, Piotr
Publié: (2026)
par: Wilam, Piotr
Publié: (2026)
vCache: Verified Semantic Prompt Caching
par: Schroeder, Luis Gaspar, et autres
Publié: (2025)
par: Schroeder, Luis Gaspar, et autres
Publié: (2025)
On the Query Complexity of Verifier-Assisted Language Generation
par: Botta, Edoardo, et autres
Publié: (2025)
par: Botta, Edoardo, et autres
Publié: (2025)
FUSE: Ensembling Verifiers with Zero Labeled Data
par: Lee, Joonhyuk, et autres
Publié: (2026)
par: Lee, Joonhyuk, et autres
Publié: (2026)
AutoPSV: Automated Process-Supervised Verifier
par: Lu, Jianqiao, et autres
Publié: (2024)
par: Lu, Jianqiao, et autres
Publié: (2024)
From Accuracy to Robustness: A Study of Rule- and Model-based Verifiers in Mathematical Reasoning
par: Huang, Yuzhen, et autres
Publié: (2025)
par: Huang, Yuzhen, et autres
Publié: (2025)
Blockwise Advantage Estimation for Multi-Objective RL with Verifiable Rewards
par: Pavlenko, Kirill, et autres
Publié: (2026)
par: Pavlenko, Kirill, et autres
Publié: (2026)
FlagEval Findings Report: A Preliminary Evaluation of Large Reasoning Models on Automatically Verifiable Textual and Visual Questions
par: Qin, Bowen, et autres
Publié: (2025)
par: Qin, Bowen, et autres
Publié: (2025)
Verify with Caution: The Pitfalls of Relying on Imperfect Factuality Metrics
par: Godbole, Ameya, et autres
Publié: (2025)
par: Godbole, Ameya, et autres
Publié: (2025)
References Improve LLM Alignment in Non-Verifiable Domains
par: Shi, Kejian, et autres
Publié: (2026)
par: Shi, Kejian, et autres
Publié: (2026)
Towards High Data Efficiency in Reinforcement Learning with Verifiable Reward
par: Tang, Xinyu, et autres
Publié: (2025)
par: Tang, Xinyu, et autres
Publié: (2025)
Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning
par: Setlur, Amrith, et autres
Publié: (2024)
par: Setlur, Amrith, et autres
Publié: (2024)
Let it Calm: Exploratory Annealed Decoding for Verifiable Reinforcement Learning
par: Yang, Chenghao, et autres
Publié: (2025)
par: Yang, Chenghao, et autres
Publié: (2025)
A Survey on Automatic Credibility Assessment Using Textual Credibility Signals in the Era of Large Language Models
par: Srba, Ivan, et autres
Publié: (2024)
par: Srba, Ivan, et autres
Publié: (2024)
Automatic Functional Differentiation in JAX
par: Lin, Min
Publié: (2023)
par: Lin, Min
Publié: (2023)
Unmasking and Improving Data Credibility: A Study with Datasets for Training Harmless Language Models
par: Zhu, Zhaowei, et autres
Publié: (2023)
par: Zhu, Zhaowei, et autres
Publié: (2023)
Boosting Reinforcement Learning with Verifiable Rewards via Randomly Selected Few-Shot Guidance
par: Yan, Kai, et autres
Publié: (2026)
par: Yan, Kai, et autres
Publié: (2026)
Low-probability Tokens Sustain Exploration in Reinforcement Learning with Verifiable Reward
par: Huang, Guanhua, et autres
Publié: (2025)
par: Huang, Guanhua, et autres
Publié: (2025)
The Alignment Auditor: A Bayesian Framework for Verifying and Refining LLM Objectives
par: Bou, Matthieu, et autres
Publié: (2025)
par: Bou, Matthieu, et autres
Publié: (2025)
A Time-Aware Approach to Early Detection of Anorexia: UNSL at eRisk 2024
par: Thompson, Horacio, et autres
Publié: (2024)
par: Thompson, Horacio, et autres
Publié: (2024)
OPV: Outcome-based Process Verifier for Efficient Long Chain-of-Thought Verification
par: Wu, Zijian, et autres
Publié: (2025)
par: Wu, Zijian, et autres
Publié: (2025)
Documents similaires
-
Attacking Misinformation Detection Using Adversarial Examples Generated by Language Models
par: Przybyła, Piotr, et autres
Publié: (2024) -
PolQA: Polish Question Answering Dataset
par: Rybak, Piotr, et autres
Publié: (2022) -
Deanthropomorphising NLP: Can a Language Model Be Conscious?
par: Shardlow, Matthew, et autres
Publié: (2022) -
CrEst: Credibility Estimation for Contexts in LLMs via Weak Supervision
par: Adila, Dyah, et autres
Publié: (2025) -
Towards Trustworthy Lexical Simplification: Exploring Safety and Efficiency with Small LLMs
par: Hayakawa, Akio, et autres
Publié: (2025)