Salvato in:
| Autori principali: | Zhou, Yuhang, Ni, Yuchen, Gan, Yunhui, Yin, Zhangyue, Liu, Xiang, Zhang, Jian, Liu, Sen, Qiu, Xipeng, Ye, Guangnan, Chai, Hongfeng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2402.12713 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
$R^3$-NL2GQL: A Model Coordination and Knowledge Graph Alignment Approach for NL2GQL
di: Zhou, Yuhang, et al.
Pubblicazione: (2023)
di: Zhou, Yuhang, et al.
Pubblicazione: (2023)
SilverSight: A Multi-Task Chinese Financial Large Language Model Based on Adaptive Semantic Space Learning
di: Zhou, Yuhang, et al.
Pubblicazione: (2024)
di: Zhou, Yuhang, et al.
Pubblicazione: (2024)
Teach2Eval: An Indirect Evaluation Method for LLM by Judging How It Teaches
di: Zhou, Yuhang, et al.
Pubblicazione: (2025)
di: Zhou, Yuhang, et al.
Pubblicazione: (2025)
RAGFormer: Learning Semantic Attributes and Topological Structure for Fraud Detection
di: Li, Haolin, et al.
Pubblicazione: (2024)
di: Li, Haolin, et al.
Pubblicazione: (2024)
FedCAda: Adaptive Client-Side Optimization for Accelerated and Stable Federated Learning
di: Zhou, Liuzhi, et al.
Pubblicazione: (2024)
di: Zhou, Liuzhi, et al.
Pubblicazione: (2024)
DogLayout: Denoising Diffusion GAN for Discrete and Continuous Layout Generation
di: Gan, Zhaoxing, et al.
Pubblicazione: (2024)
di: Gan, Zhaoxing, et al.
Pubblicazione: (2024)
Corex: Pushing the Boundaries of Complex Reasoning through Multi-Model Collaboration
di: Sun, Qiushi, et al.
Pubblicazione: (2023)
di: Sun, Qiushi, et al.
Pubblicazione: (2023)
RLoop: An Self-Improving Framework for Reinforcement Learning with Iterative Policy Initialization
di: Zhiyuan, Zeng, et al.
Pubblicazione: (2025)
di: Zhiyuan, Zeng, et al.
Pubblicazione: (2025)
Revisiting the Test-Time Scaling of o1-like Models: Do they Truly Possess Test-Time Scaling Capabilities?
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2025)
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2025)
Error Classification of Large Language Models on Math Word Problems: A Dynamically Adaptive Framework
di: Sun, Yuhong, et al.
Pubblicazione: (2025)
di: Sun, Yuhong, et al.
Pubblicazione: (2025)
GAM-RAG: Gain-Adaptive Memory for Evolving Retrieval in Retrieval-Augmented Generation
di: Wang, Yifan, et al.
Pubblicazione: (2026)
di: Wang, Yifan, et al.
Pubblicazione: (2026)
R3-RAG: Learning Step-by-Step Reasoning and Retrieval for LLMs via Reinforcement Learning
di: Li, Yuan, et al.
Pubblicazione: (2025)
di: Li, Yuan, et al.
Pubblicazione: (2025)
Limited or Biased: Modeling Sub-Rational Human Investors in Financial Markets
di: Liu, Penghang, et al.
Pubblicazione: (2022)
di: Liu, Penghang, et al.
Pubblicazione: (2022)
Balanced Aggregation: Understanding and Fixing Aggregation Bias in GRPO
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2026)
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2026)
Dynamic and Generalizable Process Reward Modeling
di: Yin, Zhangyue, et al.
Pubblicazione: (2025)
di: Yin, Zhangyue, et al.
Pubblicazione: (2025)
LLatrieval: LLM-Verified Retrieval for Verifiable Generation
di: Li, Xiaonan, et al.
Pubblicazione: (2023)
di: Li, Xiaonan, et al.
Pubblicazione: (2023)
Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem
di: Sun, Yuhong, et al.
Pubblicazione: (2024)
di: Sun, Yuhong, et al.
Pubblicazione: (2024)
Behavioral Consistency Validation for LLM Agents: An Analysis of Trading-Style Switching through Stock-Market Simulation
di: Li, Zeping, et al.
Pubblicazione: (2026)
di: Li, Zeping, et al.
Pubblicazione: (2026)
Rethinking the Role of Entropy in Optimizing Tool-Use Behaviors for Large Language Model Agents
di: Li, Zeping, et al.
Pubblicazione: (2026)
di: Li, Zeping, et al.
Pubblicazione: (2026)
TransXion: A High-Fidelity Graph Benchmark for Realistic Anti-Money Laundering
di: Chen, Keyang, et al.
Pubblicazione: (2026)
di: Chen, Keyang, et al.
Pubblicazione: (2026)
BandPO: Bridging Trust Regions and Ratio Clipping via Probability-Aware Bounds for LLM Reinforcement Learning
di: Li, Yuan, et al.
Pubblicazione: (2026)
di: Li, Yuan, et al.
Pubblicazione: (2026)
ARISE: An Adaptive Resolution-Aware Metric for Test-Time Scaling Evaluation in Large Reasoning Models
di: Yin, Zhangyue, et al.
Pubblicazione: (2025)
di: Yin, Zhangyue, et al.
Pubblicazione: (2025)
H2ST: Hierarchical Two-Sample Tests for Continual Out-of-Distribution Detection
di: Liu, Yuhang, et al.
Pubblicazione: (2025)
di: Liu, Yuhang, et al.
Pubblicazione: (2025)
LongLLaDA: Unlocking Long Context Capabilities in Diffusion LLMs
di: Liu, Xiaoran, et al.
Pubblicazione: (2025)
di: Liu, Xiaoran, et al.
Pubblicazione: (2025)
Understanding the Role of LLMs in Multimodal Evaluation Benchmarks
di: Jiang, Botian, et al.
Pubblicazione: (2024)
di: Jiang, Botian, et al.
Pubblicazione: (2024)
Sparse-dLLM: Accelerating Diffusion LLMs with Dynamic Cache Eviction
di: Song, Yuerong, et al.
Pubblicazione: (2025)
di: Song, Yuerong, et al.
Pubblicazione: (2025)
Can AI Assistants Know What They Don't Know?
di: Cheng, Qinyuan, et al.
Pubblicazione: (2024)
di: Cheng, Qinyuan, et al.
Pubblicazione: (2024)
BiasAlert: A Plug-and-play Tool for Social Bias Detection in LLMs
di: Fan, Zhiting, et al.
Pubblicazione: (2024)
di: Fan, Zhiting, et al.
Pubblicazione: (2024)
Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective
di: Chandna, Bhavik, et al.
Pubblicazione: (2025)
di: Chandna, Bhavik, et al.
Pubblicazione: (2025)
Perceived Political Bias in LLMs Reduces Persuasive Abilities
di: DiGiuseppe, Matthew, et al.
Pubblicazione: (2026)
di: DiGiuseppe, Matthew, et al.
Pubblicazione: (2026)
Beyond Real: Imaginary Extension of Rotary Position Embeddings for Long-Context LLMs
di: Liu, Xiaoran, et al.
Pubblicazione: (2025)
di: Liu, Xiaoran, et al.
Pubblicazione: (2025)
Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2024)
di: Zeng, Zhiyuan, et al.
Pubblicazione: (2024)
MetaRank: Task-Aware Metric Selection for Model Transferability Estimation
di: Liu, Yuhang, et al.
Pubblicazione: (2025)
di: Liu, Yuhang, et al.
Pubblicazione: (2025)
How AI Agents Follow the Herd of AI? Network Effects, History, and Machine Optimism
di: Liu, Yu, et al.
Pubblicazione: (2025)
di: Liu, Yu, et al.
Pubblicazione: (2025)
When Machines Meet Each Other: Network Effects and the Strategic Role of History in Multi-Agent AI
di: Liu, Yu, et al.
Pubblicazione: (2025)
di: Liu, Yu, et al.
Pubblicazione: (2025)
Limited Reference, Reliable Generation: A Two-Component Framework for Tabular Data Generation in Low-Data Regimes
di: Jiang, Mingxuan, et al.
Pubblicazione: (2025)
di: Jiang, Mingxuan, et al.
Pubblicazione: (2025)
Does Financial Statement Comparability Reduce Differences in Sentiment‐induced Investor Trading Behaviour?
di: Eun Hye Jo, et al.
Pubblicazione: (2025)
di: Eun Hye Jo, et al.
Pubblicazione: (2025)
Elucidating Mechanisms of Demographic Bias in LLMs for Healthcare
di: Ahsan, Hiba, et al.
Pubblicazione: (2025)
di: Ahsan, Hiba, et al.
Pubblicazione: (2025)
Reducing False Positives in Static Bug Detection with LLMs: An Empirical Study in Industry
di: Du, Xueying, et al.
Pubblicazione: (2026)
di: Du, Xueying, et al.
Pubblicazione: (2026)
Hint-before-Solving Prompting: Guiding LLMs to Effectively Utilize Encoded Knowledge
di: Fu, Jinlan, et al.
Pubblicazione: (2024)
di: Fu, Jinlan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
$R^3$-NL2GQL: A Model Coordination and Knowledge Graph Alignment Approach for NL2GQL
di: Zhou, Yuhang, et al.
Pubblicazione: (2023) -
SilverSight: A Multi-Task Chinese Financial Large Language Model Based on Adaptive Semantic Space Learning
di: Zhou, Yuhang, et al.
Pubblicazione: (2024) -
Teach2Eval: An Indirect Evaluation Method for LLM by Judging How It Teaches
di: Zhou, Yuhang, et al.
Pubblicazione: (2025) -
RAGFormer: Learning Semantic Attributes and Topological Structure for Fraud Detection
di: Li, Haolin, et al.
Pubblicazione: (2024) -
FedCAda: Adaptive Client-Side Optimization for Accelerated and Stable Federated Learning
di: Zhou, Liuzhi, et al.
Pubblicazione: (2024)