STAR : Bridging Statistical and Agentic Reasoning for Large Model Performance Prediction
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Xiaoxiao, Li, Chunxiao, Wang, Junying, Guo, Yijin, Chen, Zijian, Li, Chunyi, Liu, Xiaohong, Zhang, Zicheng, Zhai, Guangtao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MODP: Multi Objective Directional Prompting
di: Nema, Aashutosh, et al.
Pubblicazione: (2025)
di: Nema, Aashutosh, et al.
Pubblicazione: (2025)
LLM Reasoning for Cold-Start Item Recommendation
di: Li, Shijun, et al.
Pubblicazione: (2025)
di: Li, Shijun, et al.
Pubblicazione: (2025)
What Matters in LLM-Based Feature Extractor for Recommender? A Systematic Analysis of Prompts, Models, and Adaptation
di: Shi, Kainan, et al.
Pubblicazione: (2025)
di: Shi, Kainan, et al.
Pubblicazione: (2025)
LENS: A Staged Design for Interaction Granularityin Sequential CTR Prediction
di: Wang, Yuan, et al.
Pubblicazione: (2026)
di: Wang, Yuan, et al.
Pubblicazione: (2026)
LLM-as-a-Judge: Rapid Evaluation of Legal Document Recommendation for Retrieval-Augmented Generation
di: Pradhan, Anu, et al.
Pubblicazione: (2025)
di: Pradhan, Anu, et al.
Pubblicazione: (2025)
CourseTimeQA: A Lecture-Video Benchmark and a Latency-Constrained Cross-Modal Fusion Method for Timestamped QA
di: Kovalev, Vsevolod, et al.
Pubblicazione: (2025)
di: Kovalev, Vsevolod, et al.
Pubblicazione: (2025)
Retrieval Is Not Enough: Why Organizational AI Needs Epistemic Infrastructure
di: Bottino, Federico, et al.
Pubblicazione: (2026)
di: Bottino, Federico, et al.
Pubblicazione: (2026)
When to Forget: A Memory Governance Primitive
di: Simsek, Baris
Pubblicazione: (2026)
di: Simsek, Baris
Pubblicazione: (2026)
Adaptive$^2$: Adaptive Domain Mining for Fine-grained Domain Adaptation Modeling
di: Sun, Wenxuan, et al.
Pubblicazione: (2024)
di: Sun, Wenxuan, et al.
Pubblicazione: (2024)
EQUATOR: A Deterministic Framework for Evaluating LLM Reasoning with Open-Ended Questions. # v1.0.0-beta
di: Bernard, Raymond, et al.
Pubblicazione: (2024)
di: Bernard, Raymond, et al.
Pubblicazione: (2024)
Mubeen AI: A Specialized Arabic Language Model for Heritage Preservation and User Intent Understanding
di: Aljafari, Mohammed, et al.
Pubblicazione: (2025)
di: Aljafari, Mohammed, et al.
Pubblicazione: (2025)
Tulip Agent -- Enabling LLM-Based Agents to Solve Tasks Using Large Tool Libraries
di: Ocker, Felix, et al.
Pubblicazione: (2024)
di: Ocker, Felix, et al.
Pubblicazione: (2024)
Myriad People Open Source Software for New Media Arts
di: Baudry, Benoit, et al.
Pubblicazione: (2025)
di: Baudry, Benoit, et al.
Pubblicazione: (2025)
PRECEPT: Planning Resilience via Experience, Context Engineering & Probing Trajectories A Unified Framework for Test-Time Adaptation with Compositional Rule Learning and Pareto-Guided Prompt Evolution
di: Shahmansoori, Arash
Pubblicazione: (2026)
di: Shahmansoori, Arash
Pubblicazione: (2026)
Bayesian Coreset Optimization for Personalized Federated Learning
di: Chanda, Prateek, et al.
Pubblicazione: (2025)
di: Chanda, Prateek, et al.
Pubblicazione: (2025)
Beyond the Flat Sequence: Hierarchical and Preference-Aware Generative Recommendations
di: Chen, Zerui, et al.
Pubblicazione: (2026)
di: Chen, Zerui, et al.
Pubblicazione: (2026)
Drawing on Memory: Dual-Trace Encoding Improves Cross-Session Recall in LLM Agents
di: Stern, Benjamin, et al.
Pubblicazione: (2026)
di: Stern, Benjamin, et al.
Pubblicazione: (2026)
Optimizing Retrieval-Augmented Generation (RAG) for Colloquial Cantonese: A LoRA-Based Systematic Review
di: Calonge, David Santandreu, et al.
Pubblicazione: (2025)
di: Calonge, David Santandreu, et al.
Pubblicazione: (2025)
Evolve: A Persistent Knowledge Lifecycle for Small Language Models
di: Hovagimian, Dikran
Pubblicazione: (2026)
di: Hovagimian, Dikran
Pubblicazione: (2026)
Qtok: A Comprehensive Framework for Evaluating Multilingual Tokenizer Quality in Large Language Models
di: Chelombitko, Iaroslav, et al.
Pubblicazione: (2024)
di: Chelombitko, Iaroslav, et al.
Pubblicazione: (2024)
Task Memory Engine: Spatial Memory for Robust Multi-Step LLM Agents
di: Ye, Ye
Pubblicazione: (2025)
di: Ye, Ye
Pubblicazione: (2025)
Adaptive Data Flywheel: Applying MAPE Control Loops to AI Agent Improvement
di: Shukla, Aaditya, et al.
Pubblicazione: (2025)
di: Shukla, Aaditya, et al.
Pubblicazione: (2025)
MoVoC: Morphology-Aware Subword Construction for Geez Script Languages
di: Teklehaymanot, Hailay Kidu, et al.
Pubblicazione: (2025)
di: Teklehaymanot, Hailay Kidu, et al.
Pubblicazione: (2025)
BudgetMem: Learning Selective Memory Policies for Cost-Efficient Long-Context Processing in Language Models
di: Alla, Chandra Vamsi Krishna, et al.
Pubblicazione: (2025)
di: Alla, Chandra Vamsi Krishna, et al.
Pubblicazione: (2025)
EdgeJury: Cross-Reviewed Small-Model Ensembles for Truthful Question Answering on Serverless Edge Inference
di: Kumar, Aayush
Pubblicazione: (2025)
di: Kumar, Aayush
Pubblicazione: (2025)
Improving the Performance of Sequential Recommendation Systems with an Extended Large Language Model
di: Choi, Sinnyum, et al.
Pubblicazione: (2025)
di: Choi, Sinnyum, et al.
Pubblicazione: (2025)
LaPro-DTA: Latent Dual-View Drug Representations and Salient Protein Feature Extraction for Generalizable Drug--Target Affinity Prediction
di: Dun, Zihan, et al.
Pubblicazione: (2026)
di: Dun, Zihan, et al.
Pubblicazione: (2026)
Knowledge-Aware Iterative Retrieval for Multi-Agent Systems
di: Song, Seyoung
Pubblicazione: (2025)
di: Song, Seyoung
Pubblicazione: (2025)
Experimentation Accelerator: Interpretable Insights and Creative Recommendations for A/B Testing with Content-Aware ranking
di: Hu, Zhengmian, et al.
Pubblicazione: (2026)
di: Hu, Zhengmian, et al.
Pubblicazione: (2026)
An Audio-centric Multi-task Learning Framework for Streaming Ads Targeting on Spotify
di: Verma, Shivam, et al.
Pubblicazione: (2025)
di: Verma, Shivam, et al.
Pubblicazione: (2025)
CogRec: A Cognitive Recommender Agent Fusing Large Language Models and Soar for Explainable Recommendation
di: Hu, Jiaxin, et al.
Pubblicazione: (2025)
di: Hu, Jiaxin, et al.
Pubblicazione: (2025)
Conversational No-code, Multi-agentic Disease Module Identification and Drug Repurposing Prediction with ChatDRex
di: Süwer, Simon, et al.
Pubblicazione: (2025)
di: Süwer, Simon, et al.
Pubblicazione: (2025)
Evaluating Few-Shot Temporal Reasoning of LLMs for Human Activity Prediction in Smart Environments
di: Doctorarastoo, Maral, et al.
Pubblicazione: (2026)
di: Doctorarastoo, Maral, et al.
Pubblicazione: (2026)
Accelerating Transfer Function Update for Distance Map based Volume Rendering
di: Rauter, Michael, et al.
Pubblicazione: (2024)
di: Rauter, Michael, et al.
Pubblicazione: (2024)
Meta Knowledge for Retrieval Augmented Large Language Models
di: Mombaerts, Laurent, et al.
Pubblicazione: (2024)
di: Mombaerts, Laurent, et al.
Pubblicazione: (2024)
ComplianceNLP: Knowledge-Graph-Augmented RAG for Multi-Framework Regulatory Gap Detection
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
UniRec: A Dual Enhancement of Uniformity and Frequency in Sequential Recommendations
di: Liu, Yang, et al.
Pubblicazione: (2024)
di: Liu, Yang, et al.
Pubblicazione: (2024)
Covariance Structure and Coordinate Heterogeneity Govern Binary Quantization of Contrastive Embeddings
di: Xiao, Wenxuan
Pubblicazione: (2026)
di: Xiao, Wenxuan
Pubblicazione: (2026)
SERP Interference Network and Its Applications in Search Advertising
di: Jain, Purak, et al.
Pubblicazione: (2025)
di: Jain, Purak, et al.
Pubblicazione: (2025)
IMDMR: An Intelligent Multi-Dimensional Memory Retrieval System for Enhanced Conversational AI
di: Pawar, Tejas, et al.
Pubblicazione: (2025)
di: Pawar, Tejas, et al.
Pubblicazione: (2025)
Documenti analoghi
-
MODP: Multi Objective Directional Prompting
di: Nema, Aashutosh, et al.
Pubblicazione: (2025) -
LLM Reasoning for Cold-Start Item Recommendation
di: Li, Shijun, et al.
Pubblicazione: (2025) -
What Matters in LLM-Based Feature Extractor for Recommender? A Systematic Analysis of Prompts, Models, and Adaptation
di: Shi, Kainan, et al.
Pubblicazione: (2025) -
LENS: A Staged Design for Interaction Granularityin Sequential CTR Prediction
di: Wang, Yuan, et al.
Pubblicazione: (2026) -
LLM-as-a-Judge: Rapid Evaluation of Legal Document Recommendation for Retrieval-Augmented Generation
di: Pradhan, Anu, et al.
Pubblicazione: (2025)