Optimas: Optimizing Compound AI Systems with Globally Aligned Local Rewards
Fuente:
arXiv
Salvato in:
| Autori principali: | Wu, Shirley, Sarthi, Parth, Zhao, Shiyu, Lee, Aaron, Shandilya, Herumb, Grobelnik, Adrian Mladenic, Choudhary, Nurendra, Huang, Eddie, Subbian, Karthik, Zhang, Linjun, Yang, Diyi, Zou, James, Leskovec, Jure |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Commonsense-Infused Language-Agnostic Learning Framework for Enhancing Prediction of Political Polarity in Multilingual News Headlines
di: Swati, Swati, et al.
Pubblicazione: (2022)
di: Swati, Swati, et al.
Pubblicazione: (2022)
An Interpretable Ensemble of Graph and Language Models for Improving Search Relevance in E-Commerce
di: Choudhary, Nurendra, et al.
Pubblicazione: (2024)
di: Choudhary, Nurendra, et al.
Pubblicazione: (2024)
AvaTaR: Optimizing LLM Agents for Tool Usage via Contrastive Reasoning
di: Wu, Shirley, et al.
Pubblicazione: (2024)
di: Wu, Shirley, et al.
Pubblicazione: (2024)
STaRK: Benchmarking LLM Retrieval on Textual and Relational Knowledge Bases
di: Wu, Shirley, et al.
Pubblicazione: (2024)
di: Wu, Shirley, et al.
Pubblicazione: (2024)
GraphMETRO: Mitigating Complex Graph Distribution Shifts via Mixture of Aligned Experts
di: Wu, Shirley, et al.
Pubblicazione: (2023)
di: Wu, Shirley, et al.
Pubblicazione: (2023)
The 2021 Tokyo Olympics Multilingual News Article Dataset
di: Novak, Erik, et al.
Pubblicazione: (2025)
di: Novak, Erik, et al.
Pubblicazione: (2025)
BAR-Analytics: A Web-based Platform for Analyzing Information Spreading Barriers in News: Comparative Analysis Across Multiple Barriers and Events
di: Sittar, Abdul, et al.
Pubblicazione: (2025)
di: Sittar, Abdul, et al.
Pubblicazione: (2025)
Found in Conversation: LLMs Teach Themselves to Close the Multi-Turn Gap
di: Chen, Tianlang, et al.
Pubblicazione: (2026)
di: Chen, Tianlang, et al.
Pubblicazione: (2026)
AgentDR: Dynamic Recommendation with Implicit Item-Item Relations via LLM-based Agents
di: Yang, Mingdai, et al.
Pubblicazione: (2025)
di: Yang, Mingdai, et al.
Pubblicazione: (2025)
Uncalibrated Reasoning: GRPO Induces Overconfidence for Stochastic Outcomes
di: Bereket, Michael, et al.
Pubblicazione: (2025)
di: Bereket, Michael, et al.
Pubblicazione: (2025)
RFG: Test-Time Scaling for Diffusion Large Language Model Reasoning with Reward-Free Guidance
di: Chen, Tianlang, et al.
Pubblicazione: (2025)
di: Chen, Tianlang, et al.
Pubblicazione: (2025)
Complex Logical Reasoning over Knowledge Graphs using Large Language Models
di: Choudhary, Nurendra, et al.
Pubblicazione: (2023)
di: Choudhary, Nurendra, et al.
Pubblicazione: (2023)
All Against Some: Efficient Integration of Large Language Models for Message Passing in Graph Neural Networks
di: Jaiswal, Ajay, et al.
Pubblicazione: (2024)
di: Jaiswal, Ajay, et al.
Pubblicazione: (2024)
Automatic Normalization of Word Variations in Code-Mixed Social Media Text
di: Singh, Rajat, et al.
Pubblicazione: (2018)
di: Singh, Rajat, et al.
Pubblicazione: (2018)
HumanLM: Simulating Users with State Alignment Beats Response Imitation
di: Wu, Shirley, et al.
Pubblicazione: (2026)
di: Wu, Shirley, et al.
Pubblicazione: (2026)
RelGNN: Composite Message Passing for Relational Deep Learning
di: Chen, Tianlang, et al.
Pubblicazione: (2025)
di: Chen, Tianlang, et al.
Pubblicazione: (2025)
Reverse Image Retrieval Cues Parametric Memory in Multimodal LLMs
di: Xu, Jialiang, et al.
Pubblicazione: (2024)
di: Xu, Jialiang, et al.
Pubblicazione: (2024)
Uncertainty Quantification for Forward and Inverse Problems of PDEs via Latent Global Evolution
di: Wu, Tailin, et al.
Pubblicazione: (2024)
di: Wu, Tailin, et al.
Pubblicazione: (2024)
Large Language Models are Good Relational Learners
di: Wu, Fang, et al.
Pubblicazione: (2025)
di: Wu, Fang, et al.
Pubblicazione: (2025)
Sentiment Analysis of Code-Mixed Languages leveraging Resource Rich Languages
di: Choudhary, Nurendra, et al.
Pubblicazione: (2018)
di: Choudhary, Nurendra, et al.
Pubblicazione: (2018)
Emotions are Universal: Learning Sentiment Based Representations of Resource-Poor Languages using Siamese Networks
di: Choudhary, Nurendra, et al.
Pubblicazione: (2018)
di: Choudhary, Nurendra, et al.
Pubblicazione: (2018)
Neural Network Architecture for Credibility Assessment of Textual Claims
di: Choudhary, Nurendra, et al.
Pubblicazione: (2018)
di: Choudhary, Nurendra, et al.
Pubblicazione: (2018)
Contrastive Learning of Emoji-based Representations for Resource-Poor Languages
di: Choudhary, Nurendra, et al.
Pubblicazione: (2018)
di: Choudhary, Nurendra, et al.
Pubblicazione: (2018)
GT2Vec: Large Language Models as Multi-Modal Encoders for Text and Graph-Structured Data
di: Lin, Jiacheng, et al.
Pubblicazione: (2024)
di: Lin, Jiacheng, et al.
Pubblicazione: (2024)
MLAgentBench: Evaluating Language Agents on Machine Learning Experimentation
di: Huang, Qian, et al.
Pubblicazione: (2023)
di: Huang, Qian, et al.
Pubblicazione: (2023)
Escaping Local Optima in Global Placement
di: Xue, Ke, et al.
Pubblicazione: (2024)
di: Xue, Ke, et al.
Pubblicazione: (2024)
Relational Deep Learning: Challenges, Foundations and Next-Generation Architectures
di: Dwivedi, Vijay Prakash, et al.
Pubblicazione: (2025)
di: Dwivedi, Vijay Prakash, et al.
Pubblicazione: (2025)
Aligning Target-Aware Molecule Diffusion Models with Exact Energy Optimization
di: Gu, Siyi, et al.
Pubblicazione: (2024)
di: Gu, Siyi, et al.
Pubblicazione: (2024)
The Collection Developer's Link to Global Education.
di: Aaron, Shirley L.
Pubblicazione: (1990)
di: Aaron, Shirley L.
Pubblicazione: (1990)
RAPTOR: Recursive Abstractive Processing for Tree-Organized Retrieval
di: Sarthi, Parth, et al.
Pubblicazione: (2024)
di: Sarthi, Parth, et al.
Pubblicazione: (2024)
CollabLLM: From Passive Responders to Active Collaborators
di: Wu, Shirley, et al.
Pubblicazione: (2025)
di: Wu, Shirley, et al.
Pubblicazione: (2025)
Local and Global Optima: Sheaves and Polynomial Approximations
di: Tohmé, Fernando
Pubblicazione: (2025)
di: Tohmé, Fernando
Pubblicazione: (2025)
TimeGraphs: Graph-based Temporal Reasoning
di: Maheshwari, Paridhi, et al.
Pubblicazione: (2024)
di: Maheshwari, Paridhi, et al.
Pubblicazione: (2024)
Inferring Dynamic Networks from Marginals with Iterative Proportional Fitting
di: Chang, Serina, et al.
Pubblicazione: (2024)
di: Chang, Serina, et al.
Pubblicazione: (2024)
Learning over Positive and Negative Edges with Contrastive Message Passing
di: Pao-Huang, Peter, et al.
Pubblicazione: (2026)
di: Pao-Huang, Peter, et al.
Pubblicazione: (2026)
Computational Assessment of Clinical Drugs against SARS‐CoV‐2: Foreseeing Molecular Mechanisms and Potent Mpro Inhibitors
di: Saroj Kumar Panda, et al.
Pubblicazione: (2024)
di: Saroj Kumar Panda, et al.
Pubblicazione: (2024)
Uncomputability of Global Optima for Nonconvex Functions in the Oracle Model
di: Lakshmanan, K
Pubblicazione: (2023)
di: Lakshmanan, K
Pubblicazione: (2023)
Learning Efficient Positional Encodings with Graph Neural Networks
di: Kanatsoulis, Charilaos I., et al.
Pubblicazione: (2025)
di: Kanatsoulis, Charilaos I., et al.
Pubblicazione: (2025)
An Efficient Plugin Method for Metric Optimization of Black-Box Models
di: Devic, Siddartha, et al.
Pubblicazione: (2025)
di: Devic, Siddartha, et al.
Pubblicazione: (2025)
LLMs generate structurally realistic social networks but overestimate political homophily
di: Chang, Serina, et al.
Pubblicazione: (2024)
di: Chang, Serina, et al.
Pubblicazione: (2024)
Documenti analoghi
-
A Commonsense-Infused Language-Agnostic Learning Framework for Enhancing Prediction of Political Polarity in Multilingual News Headlines
di: Swati, Swati, et al.
Pubblicazione: (2022) -
An Interpretable Ensemble of Graph and Language Models for Improving Search Relevance in E-Commerce
di: Choudhary, Nurendra, et al.
Pubblicazione: (2024) -
AvaTaR: Optimizing LLM Agents for Tool Usage via Contrastive Reasoning
di: Wu, Shirley, et al.
Pubblicazione: (2024) -
STaRK: Benchmarking LLM Retrieval on Textual and Relational Knowledge Bases
di: Wu, Shirley, et al.
Pubblicazione: (2024) -
GraphMETRO: Mitigating Complex Graph Distribution Shifts via Mixture of Aligned Experts
di: Wu, Shirley, et al.
Pubblicazione: (2023)