RTTC: Reward-Guided Collaborative Test-Time Compute
Fuente:
arXiv
Saved in:
| Main Authors: | Muñoz, J. Pablo, Yuan, Jinjie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CoMaPOI: A Collaborative Multi-Agent Framework for Next POI Prediction Bridging the Gap Between Trajectory and Language
by: Zhong, Lin, et al.
Published: (2025)
by: Zhong, Lin, et al.
Published: (2025)
Mamba-Shedder: Post-Transformer Compression for Efficient Selective Structured State Space Models
by: Muñoz, J. Pablo, et al.
Published: (2025)
by: Muñoz, J. Pablo, et al.
Published: (2025)
MATH-PT: A Math Reasoning Benchmark for European and Brazilian Portuguese
by: Teixeira, Tiago, et al.
Published: (2026)
by: Teixeira, Tiago, et al.
Published: (2026)
Deploying Large Language Models With Retrieval Augmented Generation
by: Prabhune, Sonal, et al.
Published: (2024)
by: Prabhune, Sonal, et al.
Published: (2024)
Rewarding Creativity: A Human-Aligned Generative Reward Model for Reinforcement Learning in Storytelling
by: Li, Zhaoyan, et al.
Published: (2026)
by: Li, Zhaoyan, et al.
Published: (2026)
Equip Pre-ranking with Target Attention by Residual Quantization
by: Li, Yutong, et al.
Published: (2025)
by: Li, Yutong, et al.
Published: (2025)
MultiPruner: Balanced Structure Removal in Foundation Models
by: Muñoz, J. Pablo, et al.
Published: (2025)
by: Muñoz, J. Pablo, et al.
Published: (2025)
CAG: Chunked Augmented Generation for Google Chrome's Built-in Gemini Nano
by: Surulimuthu, Vivek Vellaiyappan, et al.
Published: (2024)
by: Surulimuthu, Vivek Vellaiyappan, et al.
Published: (2024)
FRAGATA: Semantic Retrieval of HPC Support Tickets via Hybrid RAG over 20 Years of Request Tracker History
by: Paramés-Estévez, Santiago, et al.
Published: (2026)
by: Paramés-Estévez, Santiago, et al.
Published: (2026)
A Novel Approach to Scalable and Automatic Topic-Controlled Question Generation in Education
by: Li, Ziqing, et al.
Published: (2025)
by: Li, Ziqing, et al.
Published: (2025)
Unlocking the Potential of Metaverse in Innovative and Immersive Digital Health
by: Ebrahimzadeh, Fatemeh, et al.
Published: (2024)
by: Ebrahimzadeh, Fatemeh, et al.
Published: (2024)
Scene-wise Adaptive Network for Dynamic Cold-start Scenes Optimization in CTR Prediction
by: Li, Wenhao, et al.
Published: (2024)
by: Li, Wenhao, et al.
Published: (2024)
KnowThyself: An Agentic Assistant for LLM Interpretability
by: Prasai, Suraj, et al.
Published: (2025)
by: Prasai, Suraj, et al.
Published: (2025)
Knowledge-Aware Iterative Retrieval for Multi-Agent Systems
by: Song, Seyoung
Published: (2025)
by: Song, Seyoung
Published: (2025)
Fanar: An Arabic-Centric Multimodal Generative AI Platform
by: Fanar Team, et al.
Published: (2025)
by: Fanar Team, et al.
Published: (2025)
Reducing Selection Bias in Large Language Models
by: Eicher, J. E., et al.
Published: (2024)
by: Eicher, J. E., et al.
Published: (2024)
Beyond Direct Generation: A Decomposed Approach to Well-Crafted Screenwriting with LLMs
by: Lei, Hang, et al.
Published: (2025)
by: Lei, Hang, et al.
Published: (2025)
ConCISE: A Reference-Free Conciseness Evaluation Metric for LLM-Generated Answers
by: Ghafari, Seyed Mohssen, et al.
Published: (2025)
by: Ghafari, Seyed Mohssen, et al.
Published: (2025)
Box Maze: A Process-Control Architecture for Reliable LLM Reasoning
by: Qiang, Zou
Published: (2026)
by: Qiang, Zou
Published: (2026)
Cognitively Inspired Components for Social Conversational Agents
by: Clay, Alex, et al.
Published: (2023)
by: Clay, Alex, et al.
Published: (2023)
CuentosIE: can a chatbot about "tales with a message" help to teach emotional intelligence?
by: Ferrández, Antonio, et al.
Published: (2024)
by: Ferrández, Antonio, et al.
Published: (2024)
In-Context Learning May Not Elicit Trustworthy Reasoning: A-Not-B Errors in Pretrained Language Models
by: Han, Pengrui, et al.
Published: (2024)
by: Han, Pengrui, et al.
Published: (2024)
Modeling Emotions and Ethics with Large Language Models
by: Chang, Edward Y.
Published: (2024)
by: Chang, Edward Y.
Published: (2024)
ConciseRL: Conciseness-Guided Reinforcement Learning for Efficient Reasoning Models
by: Dumitru, Razvan-Gabriel, et al.
Published: (2025)
by: Dumitru, Razvan-Gabriel, et al.
Published: (2025)
QuickSilver -- Speeding up LLM Inference through Dynamic Token Halting, KV Skipping, Contextual Token Fusion, and Adaptive Matryoshka Quantization
by: Khanna, Danush, et al.
Published: (2025)
by: Khanna, Danush, et al.
Published: (2025)
Quo Vadis ChatGPT? From Large Language Models to Large Knowledge Models
by: Venkatasubramanian, Venkat, et al.
Published: (2024)
by: Venkatasubramanian, Venkat, et al.
Published: (2024)
Training Language Models to Win Debates with Self-Play Improves Judge Accuracy
by: Arnesen, Samuel, et al.
Published: (2024)
by: Arnesen, Samuel, et al.
Published: (2024)
OG-RAG: Ontology-Grounded Retrieval-Augmented Generation For Large Language Models
by: Sharma, Kartik, et al.
Published: (2024)
by: Sharma, Kartik, et al.
Published: (2024)
OpenAI Cribbed Our Tax Example, But Can GPT-4 Really Do Tax?
by: Blair-Stanek, Andrew, et al.
Published: (2023)
by: Blair-Stanek, Andrew, et al.
Published: (2023)
Revisiting Parameter-Based Knowledge Editing in Large Language Models: Theoretical Limits and Empirical Evidence
by: Ren, Wanying, et al.
Published: (2026)
by: Ren, Wanying, et al.
Published: (2026)
Agent-Based Detection and Resolution of Incompleteness and Ambiguity in Interactions with Large Language Models
by: Naik, Riya, et al.
Published: (2025)
by: Naik, Riya, et al.
Published: (2025)
Multi-Perspective Attention Mechanism for Bias-Aware Sequential Recommendation
by: Fu, Mingjian, et al.
Published: (2025)
by: Fu, Mingjian, et al.
Published: (2025)
NRR-Core: Non-Resolution Reasoning as a Computational Framework for Contextual Identity and Ambiguity Preservation
by: Saito, Kei
Published: (2025)
by: Saito, Kei
Published: (2025)
SEDA: A Self-Adapted Entity-Centric Data Augmentation for Boosting Gird-based Discontinuous NER Models
by: Su, Wen-Fang, et al.
Published: (2025)
by: Su, Wen-Fang, et al.
Published: (2025)
Accelerating Complex Disease Treatment through Network Medicine and GenAI: A Case Study on Drug Repurposing for Breast Cancer
by: Hamed, Ahmed Abdeen, et al.
Published: (2024)
by: Hamed, Ahmed Abdeen, et al.
Published: (2024)
GACL: Graph Attention Collaborative Learning for Temporal QoS Prediction
by: Hu, Shengxiang, et al.
Published: (2024)
by: Hu, Shengxiang, et al.
Published: (2024)
The Drill-Down and Fabricate Test (DDFT): A Protocol for Measuring Epistemic Robustness in Language Models
by: Baxi, Rahul
Published: (2025)
by: Baxi, Rahul
Published: (2025)
Future of AI Models: A Computational perspective on Model collapse
by: Satharasi, Trivikram, et al.
Published: (2025)
by: Satharasi, Trivikram, et al.
Published: (2025)
EnterpriseEM: Fine-tuned Embeddings for Enterprise Semantic Search
by: Rathinasamy, Kamalkumar, et al.
Published: (2024)
by: Rathinasamy, Kamalkumar, et al.
Published: (2024)
Meta Knowledge for Retrieval Augmented Large Language Models
by: Mombaerts, Laurent, et al.
Published: (2024)
by: Mombaerts, Laurent, et al.
Published: (2024)
Similar Items
-
CoMaPOI: A Collaborative Multi-Agent Framework for Next POI Prediction Bridging the Gap Between Trajectory and Language
by: Zhong, Lin, et al.
Published: (2025) -
Mamba-Shedder: Post-Transformer Compression for Efficient Selective Structured State Space Models
by: Muñoz, J. Pablo, et al.
Published: (2025) -
MATH-PT: A Math Reasoning Benchmark for European and Brazilian Portuguese
by: Teixeira, Tiago, et al.
Published: (2026) -
Deploying Large Language Models With Retrieval Augmented Generation
by: Prabhune, Sonal, et al.
Published: (2024) -
Rewarding Creativity: A Human-Aligned Generative Reward Model for Reinforcement Learning in Storytelling
by: Li, Zhaoyan, et al.
Published: (2026)