Saved in:
| Main Authors: | Yang, Chang, Wang, Ruiyu, Jiang, Junzhe, Jiang, Qi, Zhang, Qinggang, Deng, Yanchen, Li, Shuxin, Hu, Shuyue, Li, Bo, Pokorny, Florian T., Huang, Xiao, Wang, Xinrun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2504.11239 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FinMaster: A Holistic Benchmark for Mastering Full-Pipeline Financial Workflows with LLMs
by: Jiang, Junzhe, et al.
Published: (2025)
by: Jiang, Junzhe, et al.
Published: (2025)
LLM-Based World Models Can Make Decisions Solely, But Rigorous Evaluations are Needed
by: Yang, Chang, et al.
Published: (2024)
by: Yang, Chang, et al.
Published: (2024)
Why Regression? Binary Encoding Classification Brings Confidence to Stock Market Index Price Prediction
by: Jiang, Junzhe, et al.
Published: (2025)
by: Jiang, Junzhe, et al.
Published: (2025)
GDBA Revisited: Unleashing the Power of Guided Local Search for Distributed Constraint Optimization
by: Deng, Yanchen, et al.
Published: (2025)
by: Deng, Yanchen, et al.
Published: (2025)
Configurable Mirror Descent: Towards a Unification of Decision Making
by: Li, Pengdeng, et al.
Published: (2024)
by: Li, Pengdeng, et al.
Published: (2024)
Resolving Latency and Inventory Risk in Market Making with Reinforcement Learning
by: Jiang, Junzhe, et al.
Published: (2025)
by: Jiang, Junzhe, et al.
Published: (2025)
RealCraft: Attention Control as A Tool for Zero-Shot Consistent Video Editing
by: Jin, Shutong, et al.
Published: (2023)
by: Jin, Shutong, et al.
Published: (2023)
Algorithmic Analysis of Termination Problems for Nondeterministic Quantum Programs
by: Fu, Jianling, et al.
Published: (2024)
by: Fu, Jianling, et al.
Published: (2024)
PACA: Perspective-Aware Cross-Attention Representation for Zero-Shot Scene Rearrangement
by: Jin, Shutong, et al.
Published: (2024)
by: Jin, Shutong, et al.
Published: (2024)
How Physics and Background Attributes Impact Video Transformers in Robotic Manipulation: A Case Study on Planar Pushing
by: Jin, Shutong, et al.
Published: (2023)
by: Jin, Shutong, et al.
Published: (2023)
PALM: Enhanced Generalizability for Local Visuomotor Policies via Perception Alignment
by: Wang, Ruiyu, et al.
Published: (2026)
by: Wang, Ruiyu, et al.
Published: (2026)
Reinforcement Nash Equilibrium Solver
by: Wang, Xinrun, et al.
Published: (2024)
by: Wang, Xinrun, et al.
Published: (2024)
Self-adaptive PSRO: Towards an Automatic Population-based Game Solver
by: Li, Pengdeng, et al.
Published: (2024)
by: Li, Pengdeng, et al.
Published: (2024)
Grasper: A Generalist Pursuer for Pursuit-Evasion Problems
by: Li, Pengdeng, et al.
Published: (2024)
by: Li, Pengdeng, et al.
Published: (2024)
R900: Understanding the Cost-Effectiveness of Random Exploration from 900 Hours of Robotic Data Collection
by: Jin, Shutong, et al.
Published: (2025)
by: Jin, Shutong, et al.
Published: (2025)
Simulating Polynomial-Time Nondeterministic Turing Machines via Nondeterministic Turing Machines
by: Lin, Tianrong
Published: (2024)
by: Lin, Tianrong
Published: (2024)
STAR: A First-Ever Dataset and A Large-Scale Benchmark for Scene Graph Generation in Large-Size Satellite Imagery
by: Li, Yansheng, et al.
Published: (2024)
by: Li, Yansheng, et al.
Published: (2024)
StarDojo: Benchmarking Open-Ended Behaviors of Agentic Multimodal LLMs in Production-Living Simulations with Stardew Valley
by: Tan, Weihao, et al.
Published: (2025)
by: Tan, Weihao, et al.
Published: (2025)
In-Context Exploiter for Extensive-Form Games
by: Li, Shuxin, et al.
Published: (2024)
by: Li, Shuxin, et al.
Published: (2024)
Solving Urban Network Security Games: Learning Platform, Benchmark, and Challenge for AI Research
by: Zhuang, Shuxin, et al.
Published: (2025)
by: Zhuang, Shuxin, et al.
Published: (2025)
Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning
by: Wang, Chaojie, et al.
Published: (2024)
by: Wang, Chaojie, et al.
Published: (2024)
The Agent Use of Agent Beings: Agent Cybernetics Is the Missing Science of Foundation Agents
by: Wang, Xinrun, et al.
Published: (2026)
by: Wang, Xinrun, et al.
Published: (2026)
Feature Extractor or Decision Maker: Rethinking the Role of Visual Encoders in Visuomotor Policies
by: Wang, Ruiyu, et al.
Published: (2024)
by: Wang, Ruiyu, et al.
Published: (2024)
Double Oracle Neural Architecture Search for Game Theoretic Deep Learning Models
by: Aung, Aye Phyu Phyu, et al.
Published: (2024)
by: Aung, Aye Phyu Phyu, et al.
Published: (2024)
Federated Learning for Large-Scale Cloud Robotic Manipulation: Opportunities and Challenges
by: Zaland, Obaidullah, et al.
Published: (2025)
by: Zaland, Obaidullah, et al.
Published: (2025)
FaithfulRAG: Fact-Level Conflict Modeling for Context-Faithful Retrieval-Augmented Generation
by: Zhang, Qinggang, et al.
Published: (2025)
by: Zhang, Qinggang, et al.
Published: (2025)
Xiezhi: An Ever-Updating Benchmark for Holistic Domain Knowledge Evaluation
by: Gu, Zhouhong, et al.
Published: (2023)
by: Gu, Zhouhong, et al.
Published: (2023)
DeepPHY: Benchmarking Agentic VLMs on Physical Reasoning
by: Xu, Xinrun, et al.
Published: (2025)
by: Xu, Xinrun, et al.
Published: (2025)
On Approximation Algorithms for Commutative Quaternion Polynomial Optimization
by: He, Chang, et al.
Published: (2025)
by: He, Chang, et al.
Published: (2025)
A Soft, Centimeter‐Scaled, Thin‐Cable‐Crawling Robot for Narrow Space Inspection
by: Wentao Ma, et al.
Published: (2024)
by: Wentao Ma, et al.
Published: (2024)
Reliable Reasoning Path: Distilling Effective Guidance for LLM Reasoning with Knowledge Graphs
by: Xiao, Yilin, et al.
Published: (2025)
by: Xiao, Yilin, et al.
Published: (2025)
Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
by: Fu, Chaoyou, et al.
Published: (2024)
by: Fu, Chaoyou, et al.
Published: (2024)
LLMs for Relational Reasoning: How Far are We?
by: Li, Zhiming, et al.
Published: (2024)
by: Li, Zhiming, et al.
Published: (2024)
Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs
by: Lyu, Zhiyi, et al.
Published: (2025)
by: Lyu, Zhiyi, et al.
Published: (2025)
All Changes May Have Invariant Principles: Improving Ever-Shifting Harmful Meme Detection via Design Concept Reproduction
by: Jiang, Ziyou, et al.
Published: (2026)
by: Jiang, Ziyou, et al.
Published: (2026)
Polynomial Complementation of Nondeterministic 2-Way Finite Automata by 1-Limited Automata
by: Guillon, Bruno, et al.
Published: (2025)
by: Guillon, Bruno, et al.
Published: (2025)
Diagonalization of Polynomial-Time Deterministic Turing Machines via Nondeterministic Turing Machines
by: Lin, Tianrong
Published: (2021)
by: Lin, Tianrong
Published: (2021)
Benchmarking Spatiotemporal Reasoning in LLMs and Reasoning Models: Capabilities and Challenges
by: Quan, Pengrui, et al.
Published: (2025)
by: Quan, Pengrui, et al.
Published: (2025)
INTEGRALBENCH: Benchmarking LLMs with Definite Integral Problems
by: Tang, Bintao, et al.
Published: (2025)
by: Tang, Bintao, et al.
Published: (2025)
Fat Replacers in Frozen Desserts: Functions, Challenges, and Strategies
by: Zhaoyi Tang, et al.
Published: (2025)
by: Zhaoyi Tang, et al.
Published: (2025)
Similar Items
-
FinMaster: A Holistic Benchmark for Mastering Full-Pipeline Financial Workflows with LLMs
by: Jiang, Junzhe, et al.
Published: (2025) -
LLM-Based World Models Can Make Decisions Solely, But Rigorous Evaluations are Needed
by: Yang, Chang, et al.
Published: (2024) -
Why Regression? Binary Encoding Classification Brings Confidence to Stock Market Index Price Prediction
by: Jiang, Junzhe, et al.
Published: (2025) -
GDBA Revisited: Unleashing the Power of Guided Local Search for Distributed Constraint Optimization
by: Deng, Yanchen, et al.
Published: (2025) -
Configurable Mirror Descent: Towards a Unification of Decision Making
by: Li, Pengdeng, et al.
Published: (2024)