Gespeichert in:
| Hauptverfasser: | Chen, Yanxu, Yao, Zijun, Liu, Yantao, Xin, Amy, Ye, Jin, Yu, Jianing, Hou, Lei, Li, Juanzi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2510.02209 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RM-Bench: Benchmarking Reward Models of Language Models with Subtlety and Style
von: Liu, Yantao, et al.
Veröffentlicht: (2024)
von: Liu, Yantao, et al.
Veröffentlicht: (2024)
Are Reasoning Models More Prone to Hallucination?
von: Yao, Zijun, et al.
Veröffentlicht: (2025)
von: Yao, Zijun, et al.
Veröffentlicht: (2025)
Evaluating Generative Language Models in Information Extraction as Subjective Question Correction
von: Fan, Yuchen, et al.
Veröffentlicht: (2024)
von: Fan, Yuchen, et al.
Veröffentlicht: (2024)
PairJudge RM: Perform Best-of-N Sampling with Knockout Tournament
von: Liu, Yantao, et al.
Veröffentlicht: (2025)
von: Liu, Yantao, et al.
Veröffentlicht: (2025)
Aligning Teacher with Student Preferences for Tailored Training Data Generation
von: Liu, Yantao, et al.
Veröffentlicht: (2024)
von: Liu, Yantao, et al.
Veröffentlicht: (2024)
Untangle the KNOT: Interweaving Conflicting Knowledge and Reasoning Skills in Large Language Models
von: Liu, Yantao, et al.
Veröffentlicht: (2024)
von: Liu, Yantao, et al.
Veröffentlicht: (2024)
AtomR: Atomic Operator-Empowered Large Language Models for Heterogeneous Knowledge Reasoning
von: Xin, Amy, et al.
Veröffentlicht: (2024)
von: Xin, Amy, et al.
Veröffentlicht: (2024)
DICE: Detecting In-distribution Contamination in LLM's Fine-tuning Phase for Math Reasoning
von: Tu, Shangqing, et al.
Veröffentlicht: (2024)
von: Tu, Shangqing, et al.
Veröffentlicht: (2024)
Auxiliary Metrics Help Decoding Skill Neurons in the Wild
von: Zhao, Yixiu, et al.
Veröffentlicht: (2025)
von: Zhao, Yixiu, et al.
Veröffentlicht: (2025)
LLMAEL: Large Language Models are Good Context Augmenters for Entity Linking
von: Xin, Amy, et al.
Veröffentlicht: (2024)
von: Xin, Amy, et al.
Veröffentlicht: (2024)
Guiding LLM Post-training Data Engineering with Model Internals from Sparse Autoencoders
von: Jing, Yi, et al.
Veröffentlicht: (2026)
von: Jing, Yi, et al.
Veröffentlicht: (2026)
Pre-training Distillation for Large Language Models: A Design Space Exploration
von: Peng, Hao, et al.
Veröffentlicht: (2024)
von: Peng, Hao, et al.
Veröffentlicht: (2024)
WildReward: Learning Reward Models from In-the-Wild Human Interactions
von: Peng, Hao, et al.
Veröffentlicht: (2026)
von: Peng, Hao, et al.
Veröffentlicht: (2026)
Can Artificial Intelligence Trade the Stock Market?
von: Maskiewicz, Jędrzej, et al.
Veröffentlicht: (2025)
von: Maskiewicz, Jędrzej, et al.
Veröffentlicht: (2025)
Towards Understanding Safety Alignment: A Mechanistic Perspective from Safety Neurons
von: Chen, Jianhui, et al.
Veröffentlicht: (2024)
von: Chen, Jianhui, et al.
Veröffentlicht: (2024)
When AI Meets Finance (StockAgent): Large Language Model-based Stock Trading in Simulated Real-world Environments
von: Zhang, Chong, et al.
Veröffentlicht: (2024)
von: Zhang, Chong, et al.
Veröffentlicht: (2024)
A Cause-Effect Look at Alleviating Hallucination of Knowledge-grounded Dialogue Generation
von: Yu, Jifan, et al.
Veröffentlicht: (2024)
von: Yu, Jifan, et al.
Veröffentlicht: (2024)
LinguaLens: Towards Interpreting Linguistic Mechanisms of Large Language Models via Sparse Auto-Encoder
von: Jing, Yi, et al.
Veröffentlicht: (2025)
von: Jing, Yi, et al.
Veröffentlicht: (2025)
Chaining the Evidence: Robust Reinforcement Learning for Deep Search Agents with Citation-Aware Rubric Rewards
von: Zhang, Jiajie, et al.
Veröffentlicht: (2026)
von: Zhang, Jiajie, et al.
Veröffentlicht: (2026)
Prospects of Imitating Trading Agents in the Stock Market
von: Wilinski, Mateusz, et al.
Veröffentlicht: (2025)
von: Wilinski, Mateusz, et al.
Veröffentlicht: (2025)
WaterBench: Towards Holistic Evaluation of Watermarks for Large Language Models
von: Tu, Shangqing, et al.
Veröffentlicht: (2023)
von: Tu, Shangqing, et al.
Veröffentlicht: (2023)
Agentic Reward Modeling: Integrating Human Preferences with Verifiable Correctness Signals for Reliable Reward Systems
von: Peng, Hao, et al.
Veröffentlicht: (2025)
von: Peng, Hao, et al.
Veröffentlicht: (2025)
LLMFactor: Extracting Profitable Factors through Prompts for Explainable Stock Movement Prediction
von: Wang, Meiyun, et al.
Veröffentlicht: (2024)
von: Wang, Meiyun, et al.
Veröffentlicht: (2024)
VitaBench: Benchmarking LLM Agents with Versatile Interactive Tasks in Real-world Applications
von: He, Wei, et al.
Veröffentlicht: (2025)
von: He, Wei, et al.
Veröffentlicht: (2025)
Establishing Trustworthy LLM Evaluation via Shortcut Neuron Analysis
von: Zhu, Kejian, et al.
Veröffentlicht: (2025)
von: Zhu, Kejian, et al.
Veröffentlicht: (2025)
When Agents Trade: Live Multi-Market Trading Benchmark for LLM Agents
von: Qian, Lingfei, et al.
Veröffentlicht: (2025)
von: Qian, Lingfei, et al.
Veröffentlicht: (2025)
SeaKR: Self-aware Knowledge Retrieval for Adaptive Retrieval Augmented Generation
von: Yao, Zijun, et al.
Veröffentlicht: (2024)
von: Yao, Zijun, et al.
Veröffentlicht: (2024)
Behavioral Consistency Validation for LLM Agents: An Analysis of Trading-Style Switching through Stock-Market Simulation
von: Li, Zeping, et al.
Veröffentlicht: (2026)
von: Li, Zeping, et al.
Veröffentlicht: (2026)
AdaptThink: Reasoning Models Can Learn When to Think
von: Zhang, Jiajie, et al.
Veröffentlicht: (2025)
von: Zhang, Jiajie, et al.
Veröffentlicht: (2025)
Generalized Stock Price Prediction for Multiple Stocks Combined with News Fusion
von: Liao, Pei-Jun, et al.
Veröffentlicht: (2026)
von: Liao, Pei-Jun, et al.
Veröffentlicht: (2026)
MarketSenseAI 2.0: Enhancing Stock Analysis through LLM Agents
von: Fatouros, George, et al.
Veröffentlicht: (2025)
von: Fatouros, George, et al.
Veröffentlicht: (2025)
Analyst Reports and Stock Performance: Evidence from the Chinese Market
von: Liu, Rui, et al.
Veröffentlicht: (2024)
von: Liu, Rui, et al.
Veröffentlicht: (2024)
Reverse That Number! Decoding Order Matters in Arithmetic Learning
von: Zhang-Li, Daniel, et al.
Veröffentlicht: (2024)
von: Zhang-Li, Daniel, et al.
Veröffentlicht: (2024)
T1: Advancing Language Model Reasoning through Reinforcement Learning and Inference Scaling
von: Hou, Zhenyu, et al.
Veröffentlicht: (2025)
von: Hou, Zhenyu, et al.
Veröffentlicht: (2025)
Words that Matter: The Impact of Negative Words on News Sentiment and Stock Market Index
von: Kim, Wonseong
Veröffentlicht: (2023)
von: Kim, Wonseong
Veröffentlicht: (2023)
BERTopic-Driven Stock Market Predictions: Unraveling Sentiment Insights
von: Zhu, Enmin, et al.
Veröffentlicht: (2024)
von: Zhu, Enmin, et al.
Veröffentlicht: (2024)
ECom-Bench: Can LLM Agent Resolve Real-World E-commerce Customer Support Issues?
von: Wang, Haoxin, et al.
Veröffentlicht: (2025)
von: Wang, Haoxin, et al.
Veröffentlicht: (2025)
On the Paradoxical Interference between Instruction-Following and Task Solving
von: Qi, Yunjia, et al.
Veröffentlicht: (2026)
von: Qi, Yunjia, et al.
Veröffentlicht: (2026)
Transferable and Efficient Non-Factual Content Detection via Probe Training with Offline Consistency Checking
von: Zhang, Xiaokang, et al.
Veröffentlicht: (2024)
von: Zhang, Xiaokang, et al.
Veröffentlicht: (2024)
WebSeer: Training Deeper Search Agents through Reinforcement Learning with Self-Reflection
von: He, Guanzhong, et al.
Veröffentlicht: (2025)
von: He, Guanzhong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
RM-Bench: Benchmarking Reward Models of Language Models with Subtlety and Style
von: Liu, Yantao, et al.
Veröffentlicht: (2024) -
Are Reasoning Models More Prone to Hallucination?
von: Yao, Zijun, et al.
Veröffentlicht: (2025) -
Evaluating Generative Language Models in Information Extraction as Subjective Question Correction
von: Fan, Yuchen, et al.
Veröffentlicht: (2024) -
PairJudge RM: Perform Best-of-N Sampling with Knockout Tournament
von: Liu, Yantao, et al.
Veröffentlicht: (2025) -
Aligning Teacher with Student Preferences for Tailored Training Data Generation
von: Liu, Yantao, et al.
Veröffentlicht: (2024)