FaStfact: Faster, Stronger Long-Form Factuality Evaluations in LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Wan, Yingjia, Tan, Haochen, Zhu, Xiao, Zhou, Xinyu, Li, Zhiwei, Lv, Qingsong, Sun, Changxuan, Zeng, Jiaqi, Xu, Yi, Lu, Jianqiao, Liu, Yinhong, Guo, Zhijiang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Deciphering Bitcoin Blockchain Data by Cohort Analysis
by: Liu, Yulin, et al.
Published: (2021)
by: Liu, Yulin, et al.
Published: (2021)
Open-FinLLMs: Open Multimodal Large Language Models for Financial Applications
by: Huang, Jimin, et al.
Published: (2024)
by: Huang, Jimin, et al.
Published: (2024)
Faster Monotone Implied Volatility Solver
by: Floc'h, Fabien Le
Published: (2026)
by: Floc'h, Fabien Le
Published: (2026)
The Information Dynamics of Insider Intent: How Reporting Inversions (Form 144) Mask Informational Rents in Insider Sales (Form 4)
by: Neupane, Krishna
Published: (2026)
by: Neupane, Krishna
Published: (2026)
MR-Ben: A Meta-Reasoning Benchmark for Evaluating System-2 Thinking in LLMs
by: Zeng, Zhongshen, et al.
Published: (2024)
by: Zeng, Zhongshen, et al.
Published: (2024)
All That Glisters Is Not Gold: A Benchmark for Reference-Free Counterfactual Financial Misinformation Detection
by: Jiang, Yuechen, et al.
Published: (2026)
by: Jiang, Yuechen, et al.
Published: (2026)
Identifying Evidence Subgraphs for Financial Risk Detection via Graph Counterfactual and Factual Reasoning
by: Du, Huaming, et al.
Published: (2025)
by: Du, Huaming, et al.
Published: (2025)
DiFaR: Enhancing Multimodal Misinformation Detection with Diverse, Factual, and Relevant Rationales
by: Wan, Herun, et al.
Published: (2025)
by: Wan, Herun, et al.
Published: (2025)
AutoPSV: Automated Process-Supervised Verifier
by: Lu, Jianqiao, et al.
Published: (2024)
by: Lu, Jianqiao, et al.
Published: (2024)
TL;DR: Too Long, Do Re-weighting for Efficient LLM Reasoning Compression
by: Li, Zhong-Zhi, et al.
Published: (2025)
by: Li, Zhong-Zhi, et al.
Published: (2025)
Long-Range Dependence in Financial Markets: Empirical Evidence and Generative Modeling Challenges
by: He, Yifan, et al.
Published: (2025)
by: He, Yifan, et al.
Published: (2025)
FormalAlign: Automated Alignment Evaluation for Autoformalization
by: Lu, Jianqiao, et al.
Published: (2024)
by: Lu, Jianqiao, et al.
Published: (2024)
Are Large Language Models Good In-context Learners for Financial Sentiment Analysis?
by: Wei, Xinyu, et al.
Published: (2025)
by: Wei, Xinyu, et al.
Published: (2025)
Evaluating Large Language Models (LLMs) in Financial NLP: A Comparative Study on Financial Report Analysis
by: Mohsin, Md Talha
Published: (2025)
by: Mohsin, Md Talha
Published: (2025)
An Empirical Analysis on Financial Markets: Insights from the Application of Statistical Physics
by: Li, Haochen, et al.
Published: (2023)
by: Li, Haochen, et al.
Published: (2023)
A Sinusoidal Hull-White Model for Interest Rate Dynamics: Capturing Long-Term Periodicity in U.S. Treasury Yields
by: Jha, Amit Kumar
Published: (2025)
by: Jha, Amit Kumar
Published: (2025)
A Survey of Large Language Models for Financial Applications: Progress, Prospects and Challenges
by: Nie, Yuqi, et al.
Published: (2024)
by: Nie, Yuqi, et al.
Published: (2024)
Enhancing Black-Scholes Delta Hedging via Deep Learning
by: Qiao, Chunhui, et al.
Published: (2024)
by: Qiao, Chunhui, et al.
Published: (2024)
Generative Multi-Form Bayesian Optimization
by: Guo, Zhendong, et al.
Published: (2025)
by: Guo, Zhendong, et al.
Published: (2025)
QuantBench: Benchmarking AI Methods for Quantitative Investment
by: Wang, Saizhuo, et al.
Published: (2025)
by: Wang, Saizhuo, et al.
Published: (2025)
Flow of Knowledge: Federated Fine-Tuning of LLMs in Healthcare under Non-IID Conditions
by: Chen, Zeyu, et al.
Published: (2025)
by: Chen, Zeyu, et al.
Published: (2025)
Risk forecasting using Long Short-Term Memory Mixture Density Networks
by: Herrig, Nico
Published: (2025)
by: Herrig, Nico
Published: (2025)
Evaluating LLMs in Finance Requires Explicit Bias Consideration
by: Kong, Yaxuan, et al.
Published: (2026)
by: Kong, Yaxuan, et al.
Published: (2026)
BizCompass: Benchmarking the Reasoning Capabilities of LLMs in Business Knowledge and Applications
by: Hao, Jianing, et al.
Published: (2026)
by: Hao, Jianing, et al.
Published: (2026)
Modeling of Measurement Error in Financial Returns Data
by: Jasra, Ajay, et al.
Published: (2024)
by: Jasra, Ajay, et al.
Published: (2024)
EDINET-Bench: Evaluating LLMs on Complex Financial Tasks using Japanese Financial Statements
by: Sugiura, Issa, et al.
Published: (2025)
by: Sugiura, Issa, et al.
Published: (2025)
Decentralized Token Economy Theory (DeTEcT)
by: Sadykhov, Rem, et al.
Published: (2023)
by: Sadykhov, Rem, et al.
Published: (2023)
Identifying and Quantifying Financial Bubbles with the Hyped Log-Periodic Power Law Model
by: Cao, Zheng, et al.
Published: (2025)
by: Cao, Zheng, et al.
Published: (2025)
DeTEcT: Dynamic and Probabilistic Parameters Extension
by: Sadykhov, Rem, et al.
Published: (2024)
by: Sadykhov, Rem, et al.
Published: (2024)
Impacts of Economic Policies on Wealth Distribution in Token Economies
by: Sadykhov, Rem, et al.
Published: (2026)
by: Sadykhov, Rem, et al.
Published: (2026)
Look-Ahead-Bench: a Standardized Benchmark of Look-ahead Bias in Point-in-Time LLMs for Finance
by: Benhenda, Mostapha
Published: (2026)
by: Benhenda, Mostapha
Published: (2026)
Operator-based machine learning framework for generalizable prediction of unsteady treatment dynamics in stormwater infrastructure
by: Shatarah, Mohamed, et al.
Published: (2025)
by: Shatarah, Mohamed, et al.
Published: (2025)
StableAML: Machine Learning for Behavioral Wallet Detection in Stablecoin Anti-Money Laundering on Ethereum
by: Juvinski, Luciano, et al.
Published: (2026)
by: Juvinski, Luciano, et al.
Published: (2026)
UCFE: A User-Centric Financial Expertise Benchmark for Large Language Models
by: Yang, Yuzhe, et al.
Published: (2024)
by: Yang, Yuzhe, et al.
Published: (2024)
MoA is All You Need: Building LLM Research Team using Mixture of Agents
by: Chen, Sandy, et al.
Published: (2024)
by: Chen, Sandy, et al.
Published: (2024)
Leveraging Multi-modal Representations to Predict Protein Melting Temperatures
by: Zhang, Daiheng, et al.
Published: (2024)
by: Zhang, Daiheng, et al.
Published: (2024)
A force-based beam element model based on the modified higher-order shear deformation theory for accurate analysis of FG beams
by: Li, Wenxiong, et al.
Published: (2023)
by: Li, Wenxiong, et al.
Published: (2023)
"Global is Good, Local is Bad?": Understanding Brand Bias in LLMs
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
by: Kamruzzaman, Mahammed, et al.
Published: (2024)
Can LLMs Learn Macroeconomic Narratives from Social Media?
by: Gueta, Almog, et al.
Published: (2024)
by: Gueta, Almog, et al.
Published: (2024)
The Statistical Significance of the Inclusion of Graph Neural Networks in the Financial Time Series Forecasting Problem
by: Gregnanin, Marco, et al.
Published: (2026)
by: Gregnanin, Marco, et al.
Published: (2026)
Similar Items
-
Deciphering Bitcoin Blockchain Data by Cohort Analysis
by: Liu, Yulin, et al.
Published: (2021) -
Open-FinLLMs: Open Multimodal Large Language Models for Financial Applications
by: Huang, Jimin, et al.
Published: (2024) -
Faster Monotone Implied Volatility Solver
by: Floc'h, Fabien Le
Published: (2026) -
The Information Dynamics of Insider Intent: How Reporting Inversions (Form 144) Mask Informational Rents in Insider Sales (Form 4)
by: Neupane, Krishna
Published: (2026) -
MR-Ben: A Meta-Reasoning Benchmark for Evaluating System-2 Thinking in LLMs
by: Zeng, Zhongshen, et al.
Published: (2024)