FrontierCS: Evolving Challenges for Evolving Intelligence
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mang, Qiuyang, Chai, Wenhao, Li, Zhifei, Mao, Huanzhi, Zhou, Shang, Du, Alexander, Li, Hanchen, Liu, Shu, Chen, Edwin, Wang, Yichuan, Chu, Xieting, Cheng, Zerui, Xu, Yuan, Xia, Tian, Wang, Zirui, Shi, Tianneng, Yao, Jianzhu, Zhao, Yilong, Zhang, Qizheng, Ruan, Charlie, Shen, Zeyu, Liu, Kaiyuan, He, Runyuan, Xing, Dong, Li, Zerui, Zeng, Zirong, Jiang, Yige, Cheng, Lufeng, Zhao, Ziyi, Sun, Youran, Zheng, Wesley, Zhang, Meiyuwang, Ji, Ruyi, Tu, Xuechang, Zheng, Zihan, Chen, Zexing, Zhou, Kangyang, Wang, Zhaozi, Chen, Jingbang, Korolova, Aleksandra, Henderson, Peter, Viswanath, Pramod, Ganesh, Vijay, Xie, Saining, Liu, Zhuang, Song, Dawn, Min, Sewon, Stoica, Ion, Gonzalez, Joseph E., Shang, Jingbo, Cheung, Alvin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AutoCode: LLMs as Problem Setters for Competitive Programming
von: Zhou, Shang, et al.
Veröffentlicht: (2025)
von: Zhou, Shang, et al.
Veröffentlicht: (2025)
FrontierSmith: Synthesizing Open-Ended Coding Problems at Scale
von: He, Runyuan, et al.
Veröffentlicht: (2026)
von: He, Runyuan, et al.
Veröffentlicht: (2026)
STAA: Spatio-Temporal Attention Attribution for Real-Time Interpreting Transformer-based Video Models
von: Wang, Zerui, et al.
Veröffentlicht: (2024)
von: Wang, Zerui, et al.
Veröffentlicht: (2024)
Cloud-based XAI Services for Assessing Open Repository Models Under Adversarial Attacks
von: Wang, Zerui, et al.
Veröffentlicht: (2024)
von: Wang, Zerui, et al.
Veröffentlicht: (2024)
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
von: Zheng, Zihan, et al.
Veröffentlicht: (2025)
von: Zheng, Zihan, et al.
Veröffentlicht: (2025)
TabularMath: Evaluating Computational Extrapolation in Tabular Learning via Program-Verified Synthesis
von: Cheng, Zerui, et al.
Veröffentlicht: (2026)
von: Cheng, Zerui, et al.
Veröffentlicht: (2026)
rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
von: Guan, Xinyu, et al.
Veröffentlicht: (2025)
von: Guan, Xinyu, et al.
Veröffentlicht: (2025)
VeRA: Verified Reasoning Data Augmentation at Scale
von: Cheng, Zerui, et al.
Veröffentlicht: (2026)
von: Cheng, Zerui, et al.
Veröffentlicht: (2026)
An Open API Architecture to Discover the Trustworthy Explanation of Cloud AI Services
von: Wang, Zerui, et al.
Veröffentlicht: (2024)
von: Wang, Zerui, et al.
Veröffentlicht: (2024)
TAO: Tolerance-Aware Optimistic Verification for Floating-Point Neural Networks
von: Yao, Jianzhu, et al.
Veröffentlicht: (2025)
von: Yao, Jianzhu, et al.
Veröffentlicht: (2025)
MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents
von: Zhang, Haozhen, et al.
Veröffentlicht: (2026)
von: Zhang, Haozhen, et al.
Veröffentlicht: (2026)
Tools as Continuous Flow for Evolving Agentic Reasoning
von: Huang, Tairan, et al.
Veröffentlicht: (2026)
von: Huang, Tairan, et al.
Veröffentlicht: (2026)
SPIN-Bench: How Well Do LLMs Plan Strategically and Reason Socially?
von: Yao, Jianzhu, et al.
Veröffentlicht: (2025)
von: Yao, Jianzhu, et al.
Veröffentlicht: (2025)
Investigating Public Fine-Tuning Datasets: A Complex Review of Current Practices from a Construction Perspective
von: Ma, Runyuan, et al.
Veröffentlicht: (2024)
von: Ma, Runyuan, et al.
Veröffentlicht: (2024)
Detection of HI filament: Pair Stacking vs. Filament Stacking
von: Meng, Yuxi, et al.
Veröffentlicht: (2025)
von: Meng, Yuxi, et al.
Veröffentlicht: (2025)
An Energy‐Based Deep Learning Method for Shakedown Analysis of Elastoplastic Structures
von: Yuzhou Lin, et al.
Veröffentlicht: (2026)
von: Yuzhou Lin, et al.
Veröffentlicht: (2026)
Learning to Self-Evolve
von: Chen, Xiaoyin, et al.
Veröffentlicht: (2026)
von: Chen, Xiaoyin, et al.
Veröffentlicht: (2026)
Enhancing Generalization in Medical Visual Question Answering Tasks via Gradient-Guided Model Perturbation
von: Liu, Gang, et al.
Veröffentlicht: (2024)
von: Liu, Gang, et al.
Veröffentlicht: (2024)
Learning to Evolve for Optimization via Stability-Inducing Neural Unrolling
von: Gao, Jiaxin, et al.
Veröffentlicht: (2025)
von: Gao, Jiaxin, et al.
Veröffentlicht: (2025)
When Hallucination Costs Millions: Benchmarking AI Agents in High-Stakes Adversarial Financial Markets
von: Dai, Zeshi, et al.
Veröffentlicht: (2025)
von: Dai, Zeshi, et al.
Veröffentlicht: (2025)
Learning Explicit Contact for Implicit Reconstruction of Hand-held Objects from Monocular Images
von: Hu, Junxing, et al.
Veröffentlicht: (2023)
von: Hu, Junxing, et al.
Veröffentlicht: (2023)
Google is all you need: Semi-Supervised Transfer Learning Strategy For Light Multimodal Multi-Task Classification Model
von: Liu, Haixu, et al.
Veröffentlicht: (2025)
von: Liu, Haixu, et al.
Veröffentlicht: (2025)
COMAP: Co-Evolving World Models and Agent Policies for LLM Agents
von: Liu, Youwei, et al.
Veröffentlicht: (2026)
von: Liu, Youwei, et al.
Veröffentlicht: (2026)
Shangri-La, paraíso terrenal / Liu Huanzhi y Zhang Tao
von: Liu Huanzhi
Veröffentlicht: (2008)
von: Liu Huanzhi
Veröffentlicht: (2008)
Campesinos buscan una vida modestamente acomodada / Liu Huanzhi
von: Liu Huanzhi
von: Liu Huanzhi
DisenTS: Disentangled Channel Evolving Pattern Modeling for Multivariate Time Series Forecasting
von: Liu, Zhiding, et al.
Veröffentlicht: (2024)
von: Liu, Zhiding, et al.
Veröffentlicht: (2024)
The Evolving Duet of Two Modalities: A Survey on Integrating Text and Visualization for Data Communication
von: Lan, Xingyu, et al.
Veröffentlicht: (2026)
von: Lan, Xingyu, et al.
Veröffentlicht: (2026)
Diving into Self-Evolving Training for Multimodal Reasoning
von: Liu, Wei, et al.
Veröffentlicht: (2024)
von: Liu, Wei, et al.
Veröffentlicht: (2024)
Chain-of-Table: Evolving Tables in the Reasoning Chain for Table Understanding
von: Wang, Zilong, et al.
Veröffentlicht: (2024)
von: Wang, Zilong, et al.
Veröffentlicht: (2024)
InferenceEvolve: Towards Automated Causal Effect Estimators through Self-Evolving AI
von: Wang, Can, et al.
Veröffentlicht: (2026)
von: Wang, Can, et al.
Veröffentlicht: (2026)
Data-Efficient Training by Evolved Sampling
von: Cheng, Ziheng, et al.
Veröffentlicht: (2025)
von: Cheng, Ziheng, et al.
Veröffentlicht: (2025)
Mem$^2$Evolve: Towards Self-Evolving Agents via Co-Evolutionary Capability Expansion and Experience Distillation
von: Cheng, Zihao, et al.
Veröffentlicht: (2026)
von: Cheng, Zihao, et al.
Veröffentlicht: (2026)
Self-Evolving Spatial Reasoning in Vision Language Models via Geometric Logic Consistency
von: Liu, Junming, et al.
Veröffentlicht: (2026)
von: Liu, Junming, et al.
Veröffentlicht: (2026)
OpenDeepThink: Parallel Reasoning via Bradley-Terry Aggregation
von: Zhou, Shang, et al.
Veröffentlicht: (2026)
von: Zhou, Shang, et al.
Veröffentlicht: (2026)
Fast Multichannel Topology Discovery in Cognitive Radio Networks
von: Wang, Yung-Li, et al.
Veröffentlicht: (2025)
von: Wang, Yung-Li, et al.
Veröffentlicht: (2025)
EvolveReason: Self-Evolving Reasoning Paradigm for Explainable Deepfake Facial Image Identification
von: Zhou, Binjia, et al.
Veröffentlicht: (2026)
von: Zhou, Binjia, et al.
Veröffentlicht: (2026)
Evolve the Method, Not the Prompts: Evolutionary Synthesis of Jailbreak Attacks on LLMs
von: Chen, Yunhao, et al.
Veröffentlicht: (2025)
von: Chen, Yunhao, et al.
Veröffentlicht: (2025)
Nature‐Inspired Structure Design for Robust and Efficient Flexible Perovskite Solar Cells
von: Zerui Du, et al.
Veröffentlicht: (2025)
von: Zerui Du, et al.
Veröffentlicht: (2025)
PolyJailbreak: Cross-Modal Jailbreaking Attacks on Black-Box Multimodal LLMs
von: Wang, Xinkai, et al.
Veröffentlicht: (2025)
von: Wang, Xinkai, et al.
Veröffentlicht: (2025)
Best Transition Matrix Esitimation or Best Label Noise Robustness Classifier? Two Possible Methods to Enhance the Performance of T-revision
von: Liu, Haixu, et al.
Veröffentlicht: (2025)
von: Liu, Haixu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
AutoCode: LLMs as Problem Setters for Competitive Programming
von: Zhou, Shang, et al.
Veröffentlicht: (2025) -
FrontierSmith: Synthesizing Open-Ended Coding Problems at Scale
von: He, Runyuan, et al.
Veröffentlicht: (2026) -
STAA: Spatio-Temporal Attention Attribution for Real-Time Interpreting Transformer-based Video Models
von: Wang, Zerui, et al.
Veröffentlicht: (2024) -
Cloud-based XAI Services for Assessing Open Repository Models Under Adversarial Attacks
von: Wang, Zerui, et al.
Veröffentlicht: (2024) -
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
von: Zheng, Zihan, et al.
Veröffentlicht: (2025)