Salvato in:
| Autori principali: | Wang, Xinglin, Liu, Zishen, Feng, Shaoxiong, Yuan, Peiwen, Li, Yiwei, Shi, Jiayi, Zhang, Yueqi, Tan, Chuyi, Zhang, Ji, Pan, Boyuan, Hu, Yao, Li, Kan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2605.06110 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Every Rollout Counts: Optimal Resource Allocation for Efficient Test-Time Scaling
di: Wang, Xinglin, et al.
Pubblicazione: (2025)
di: Wang, Xinglin, et al.
Pubblicazione: (2025)
Make Every Penny Count: Difficulty-Adaptive Self-Consistency for Cost-Efficient Reasoning
di: Wang, Xinglin, et al.
Pubblicazione: (2024)
di: Wang, Xinglin, et al.
Pubblicazione: (2024)
UniCBE: An Uniformity-driven Comparing Based Evaluation Framework with Unified Multi-Objective Optimization
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
Silencer: From Discovery to Mitigation of Self-Bias in LLM-as-Benchmark-Generator
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
Beyond One-Size-Fits-All: Tailored Benchmarks for Efficient Evaluation
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
Mind the Quote: Enabling Quotation-Aware Dialogue in LLMs via Plug-and-Play Modules
di: Zhang, Yueqi, et al.
Pubblicazione: (2025)
di: Zhang, Yueqi, et al.
Pubblicazione: (2025)
From Sub-Ability Diagnosis to Human-Aligned Generation: Bridging the Gap for Text Length Control via MARKERGEN
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
LLM-Powered Benchmark Factory: Reliable, Generic, and Efficient
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)
Revisiting Self-Consistency from Dynamic Distributional Alignment Perspective on Answer Aggregation
di: Li, Yiwei, et al.
Pubblicazione: (2025)
di: Li, Yiwei, et al.
Pubblicazione: (2025)
Diagnosing and Mitigating System Bias in Self-Rewarding RL
di: Tan, Chuyi, et al.
Pubblicazione: (2025)
di: Tan, Chuyi, et al.
Pubblicazione: (2025)
PatternKV: Flattening KV Representation Expands Quantization Headroom
di: Zhang, Ji, et al.
Pubblicazione: (2025)
di: Zhang, Ji, et al.
Pubblicazione: (2025)
Do Not Waste Your Rollouts: Recycling Search Experience for Efficient Test-Time Scaling
di: Wang, Xinglin, et al.
Pubblicazione: (2026)
di: Wang, Xinglin, et al.
Pubblicazione: (2026)
Speculative Decoding for Multi-Sample Inference
di: Li, Yiwei, et al.
Pubblicazione: (2025)
di: Li, Yiwei, et al.
Pubblicazione: (2025)
InsBank: Evolving Instruction Subset for Ongoing Alignment
di: Shi, Jiayi, et al.
Pubblicazione: (2025)
di: Shi, Jiayi, et al.
Pubblicazione: (2025)
Focused Large Language Models are Stable Many-Shot Learners
di: Yuan, Peiwen, et al.
Pubblicazione: (2024)
di: Yuan, Peiwen, et al.
Pubblicazione: (2024)
Learning More from Less: Unlocking Internal Representations for Benchmark Compression
di: Zhang, Yueqi, et al.
Pubblicazione: (2026)
di: Zhang, Yueqi, et al.
Pubblicazione: (2026)
Instruction Embedding: Latent Representations of Instructions Towards Task Identification
di: Li, Yiwei, et al.
Pubblicazione: (2024)
di: Li, Yiwei, et al.
Pubblicazione: (2024)
BatchEval: Towards Human-like Text Evaluation
di: Yuan, Peiwen, et al.
Pubblicazione: (2023)
di: Yuan, Peiwen, et al.
Pubblicazione: (2023)
CogLM: Tracking Cognitive Development of Large Language Models
di: Wang, Xinglin, et al.
Pubblicazione: (2024)
di: Wang, Xinglin, et al.
Pubblicazione: (2024)
Poor-Supervised Evaluation for SuperLLM via Mutual Consistency
di: Yuan, Peiwen, et al.
Pubblicazione: (2024)
di: Yuan, Peiwen, et al.
Pubblicazione: (2024)
Integrate the Essence and Eliminate the Dross: Fine-Grained Self-Consistency for Free-Form Language Generation
di: Wang, Xinglin, et al.
Pubblicazione: (2024)
di: Wang, Xinglin, et al.
Pubblicazione: (2024)
Escape Sky-high Cost: Early-stopping Self-Consistency for Multi-step Reasoning
di: Li, Yiwei, et al.
Pubblicazione: (2024)
di: Li, Yiwei, et al.
Pubblicazione: (2024)
Generative Dense Retrieval: Memory Can Be a Burden
di: Yuan, Peiwen, et al.
Pubblicazione: (2024)
di: Yuan, Peiwen, et al.
Pubblicazione: (2024)
Online Resource Allocation with Average Budget Constraints
di: Ao, Ruicheng, et al.
Pubblicazione: (2024)
di: Ao, Ruicheng, et al.
Pubblicazione: (2024)
Balancing Fairness and Efficiency in Energy Resource Allocations
di: Li, Jiayi, et al.
Pubblicazione: (2024)
di: Li, Jiayi, et al.
Pubblicazione: (2024)
Online Resource Allocation With General Constraints
di: Chiefari, Eleonora Fidelia, et al.
Pubblicazione: (2026)
di: Chiefari, Eleonora Fidelia, et al.
Pubblicazione: (2026)
A Budget-Adaptive Allocation Rule for Optimal Computing Budget Allocation
di: Cao, Zirui, et al.
Pubblicazione: (2023)
di: Cao, Zirui, et al.
Pubblicazione: (2023)
Resource Allocation in Electricity Markets with Budget Constrained Customers
di: Perkins, Lila, et al.
Pubblicazione: (2026)
di: Perkins, Lila, et al.
Pubblicazione: (2026)
Stitch and Tell: A Structured Multimodal Data Augmentation Method for Spatial Understanding
di: Yin, Hang, et al.
Pubblicazione: (2025)
di: Yin, Hang, et al.
Pubblicazione: (2025)
RobustFlow: Towards Robust Agentic Workflow Generation
di: Xu, Shengxiang, et al.
Pubblicazione: (2025)
di: Xu, Shengxiang, et al.
Pubblicazione: (2025)
Adaptive Resource Allocation for Workflow Containerization on Kubernetes
di: Shan, Chenggang, et al.
Pubblicazione: (2023)
di: Shan, Chenggang, et al.
Pubblicazione: (2023)
Online Allocation with Replenishable Budgets: Worst Case and Beyond
di: Yang, Jianyi, et al.
Pubblicazione: (2024)
di: Yang, Jianyi, et al.
Pubblicazione: (2024)
Reliable Online Resource Allocation for Multi-User Semantic Communications: A Constraint Bayesian Optimization Approach
di: Hou, Huawei, et al.
Pubblicazione: (2026)
di: Hou, Huawei, et al.
Pubblicazione: (2026)
Dynamic Stochastic Decoding Strategy for Open-Domain Dialogue Generation
di: Li, Yiwei, et al.
Pubblicazione: (2024)
di: Li, Yiwei, et al.
Pubblicazione: (2024)
Rethinking Multilingual Continual Pretraining: Data Mixing for Adapting LLMs Across Languages and Resources
di: Li, Zihao, et al.
Pubblicazione: (2025)
di: Li, Zihao, et al.
Pubblicazione: (2025)
Learning with a Budget: Identifying the Best Arm with Resource Constraints
di: Li, Zitian, et al.
Pubblicazione: (2026)
di: Li, Zitian, et al.
Pubblicazione: (2026)
Shape Memory‐Driven Intelligent Composite Film for Infrared Stealth and Adjustable EMI Shielding
di: Yang Bai, et al.
Pubblicazione: (2025)
di: Yang Bai, et al.
Pubblicazione: (2025)
Online Optimization for Randomized Network Resource Allocation with Long-Term Constraints
di: Sid-Ali, Ahmed, et al.
Pubblicazione: (2023)
di: Sid-Ali, Ahmed, et al.
Pubblicazione: (2023)
Data-Driven Online Resource Allocation for User Experience Improvement in Mobile Edge Clouds
di: Fu, Liqun, et al.
Pubblicazione: (2024)
di: Fu, Liqun, et al.
Pubblicazione: (2024)
Flow: Modularized Agentic Workflow Automation
di: Niu, Boye, et al.
Pubblicazione: (2025)
di: Niu, Boye, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Every Rollout Counts: Optimal Resource Allocation for Efficient Test-Time Scaling
di: Wang, Xinglin, et al.
Pubblicazione: (2025) -
Make Every Penny Count: Difficulty-Adaptive Self-Consistency for Cost-Efficient Reasoning
di: Wang, Xinglin, et al.
Pubblicazione: (2024) -
UniCBE: An Uniformity-driven Comparing Based Evaluation Framework with Unified Multi-Objective Optimization
di: Yuan, Peiwen, et al.
Pubblicazione: (2025) -
Silencer: From Discovery to Mitigation of Self-Bias in LLM-as-Benchmark-Generator
di: Yuan, Peiwen, et al.
Pubblicazione: (2025) -
Beyond One-Size-Fits-All: Tailored Benchmarks for Efficient Evaluation
di: Yuan, Peiwen, et al.
Pubblicazione: (2025)