Saved in:
| Main Authors: | Carrión, Héctor, Bai, Yutong, Castro, Víctor A. Hernández, Panaganti, Kishan, Zenith, Ayush, Trang, Matthew, Zhang, Tony, Perona, Pietro, Malik, Jitendra |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2506.11302 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SDQM: Synthetic Data Quality Metric for Object Detection Dataset Evaluation
by: Zenith, Ayush, et al.
Published: (2025)
by: Zenith, Ayush, et al.
Published: (2025)
Tractable Equilibrium Computation in Markov Games through Risk Aversion
by: Mazumdar, Eric, et al.
Published: (2024)
by: Mazumdar, Eric, et al.
Published: (2024)
Model-Free Robust $ϕ$-Divergence Reinforcement Learning Using Both Offline and Online Data
by: Panaganti, Kishan, et al.
Published: (2024)
by: Panaganti, Kishan, et al.
Published: (2024)
TARDIS: Mitigating Temporal Misalignment via Representation Steering
by: Shin, Changho, et al.
Published: (2025)
by: Shin, Changho, et al.
Published: (2025)
Efficient Policy Optimization in Robust Constrained MDPs with Iteration Complexity Guarantees
by: Ganguly, Sourav, et al.
Published: (2025)
by: Ganguly, Sourav, et al.
Published: (2025)
The Single-Multi Evolution Loop for Self-Improving Model Collaboration Systems
by: Feng, Shangbin, et al.
Published: (2026)
by: Feng, Shangbin, et al.
Published: (2026)
Hybrid Transfer Reinforcement Learning: Provable Sample Efficiency from Shifted-Dynamics Data
by: Qu, Chengrui, et al.
Published: (2024)
by: Qu, Chengrui, et al.
Published: (2024)
KL-regularization Itself is Differentially Private in Bandits and RLHF
by: Zhang, Yizhou, et al.
Published: (2025)
by: Zhang, Yizhou, et al.
Published: (2025)
Group Distributionally Robust Optimization-Driven Reinforcement Learning for LLM Reasoning
by: Panaganti, Kishan, et al.
Published: (2026)
by: Panaganti, Kishan, et al.
Published: (2026)
Distributionally Robust Cooperative Multi-Agent Reinforcement Learning via Robust Value Factorization
by: Qu, Chengrui, et al.
Published: (2026)
by: Qu, Chengrui, et al.
Published: (2026)
A Number Sense as an Emergent Property of the Manipulating Brain
by: Kondapaneni, Neehar, et al.
Published: (2020)
by: Kondapaneni, Neehar, et al.
Published: (2020)
A ATUAÇÃO DO ENFERMEIRO NAS PRÁTICAS DE GERENCIAR EM UNIDADES HOSPITALARES: UMA REVISÃO INTEGRATIVA
by: Zenith Rosa Silvino
Published: (2011)
by: Zenith Rosa Silvino
Published: (2011)
AS ESTRATÉGIAS DEFENSIVAS UTILIZADAS PELO TRABALHADOR DE ENFERMAGEM: UMA REVISÃO INTEGRAL DA LITERATURA
by: Zenith Rosa Silvino
Published: (2010)
by: Zenith Rosa Silvino
Published: (2010)
Risk-Averse Total-Reward Reinforcement Learning
by: Su, Xihong, et al.
Published: (2025)
by: Su, Xihong, et al.
Published: (2025)
Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis
by: Shi, Yucheng, et al.
Published: (2026)
by: Shi, Yucheng, et al.
Published: (2026)
Robust LLM Alignment via Distributionally Robust Direct Preference Optimization
by: Xu, Zaiyan, et al.
Published: (2025)
by: Xu, Zaiyan, et al.
Published: (2025)
Every Question Has Its Own Value: Reinforcement Learning with Explicit Human Values
by: Yu, Dian, et al.
Published: (2025)
by: Yu, Dian, et al.
Published: (2025)
Distributionally Robust Constrained Reinforcement Learning under Strong Duality
by: Zhang, Zhengfei, et al.
Published: (2024)
by: Zhang, Zhengfei, et al.
Published: (2024)
Off-Policy Evaluation Using Information Borrowing and Context-Based Switching
by: Dasgupta, Sutanoy, et al.
Published: (2021)
by: Dasgupta, Sutanoy, et al.
Published: (2021)
Incidencia de las tecnologías de información y comunicación en la formación de los estudiantes de la Universidad de la Guajira
by: Zenith Palencia del Valle
Published: (2016)
by: Zenith Palencia del Valle
Published: (2016)
Abundance Stratification in Type Iax SN 2020rea with TARDIS
by: Kayal, Sohini, et al.
Published: (2025)
by: Kayal, Sohini, et al.
Published: (2025)
Whole-Body Conditioned Egocentric Video Prediction
by: Bai, Yutong, et al.
Published: (2025)
by: Bai, Yutong, et al.
Published: (2025)
Single View Seafloor Recovery from Imaging Sonar via Differentiable Rendering
by: Brodjian, Sevan, et al.
Published: (2026)
by: Brodjian, Sevan, et al.
Published: (2026)
Confidence Intervals for Error Rates in 1:1 Matching Tasks: Critical Statistical Analysis and Recommendations
by: Fogliato, Riccardo, et al.
Published: (2023)
by: Fogliato, Riccardo, et al.
Published: (2023)
FlowNet: Modeling Dynamic Spatio-Temporal Systems via Flow Propagation
by: Feng, Yutong, et al.
Published: (2025)
by: Feng, Yutong, et al.
Published: (2025)
Spatio-Temporal Field Neural Networks for Air Quality Inference
by: Feng, Yutong, et al.
Published: (2024)
by: Feng, Yutong, et al.
Published: (2024)
Can LLMs Guide Their Own Exploration? Gradient-Guided Reinforcement Learning for LLM Reasoning
by: Liang, Zhenwen, et al.
Published: (2025)
by: Liang, Zhenwen, et al.
Published: (2025)
Guided Self-Evolving LLMs with Minimal Human Supervision
by: Yu, Wenhao, et al.
Published: (2025)
by: Yu, Wenhao, et al.
Published: (2025)
Representational Similarity via Interpretable Visual Concepts
by: Kondapaneni, Neehar, et al.
Published: (2025)
by: Kondapaneni, Neehar, et al.
Published: (2025)
Representational Difference Explanations
by: Kondapaneni, Neehar, et al.
Published: (2025)
by: Kondapaneni, Neehar, et al.
Published: (2025)
Counting Fish with Temporal Representations of Sonar Video
by: Van Brunt, Kai, et al.
Published: (2025)
by: Van Brunt, Kai, et al.
Published: (2025)
AURA: An Agent Autonomy Risk Assessment Framework
by: Chiris, Lorenzo Satta, et al.
Published: (2025)
by: Chiris, Lorenzo Satta, et al.
Published: (2025)
A Dynamical Model for Spatio-Temporal Processes Motivated by Second-Order Partial Differential Equations
by: Zhang, Yutong, et al.
Published: (2025)
by: Zhang, Yutong, et al.
Published: (2025)
Exploring the origins of high-velocity features in SNe Ia with the spectral synthesis code TARDIS
by: Harvey, Luke, et al.
Published: (2025)
by: Harvey, Luke, et al.
Published: (2025)
Towards Unbiased and Robust Spatio-Temporal Scene Graph Generation and Anticipation
by: Peddi, Rohith, et al.
Published: (2024)
by: Peddi, Rohith, et al.
Published: (2024)
SAVeD: Learning to Denoise Low-SNR Video for Improved Downstream Performance
by: Stathatos, Suzanne, et al.
Published: (2025)
by: Stathatos, Suzanne, et al.
Published: (2025)
Linear Mechanisms for Spatiotemporal Reasoning in Vision Language Models
by: Kang, Raphi, et al.
Published: (2026)
by: Kang, Raphi, et al.
Published: (2026)
Diffusion-Based Action Recognition Generalizes to Untrained Domains
by: Guimaraes, Rogerio, et al.
Published: (2025)
by: Guimaraes, Rogerio, et al.
Published: (2025)
Is CLIP ideal? No. Can we fix it? Yes!
by: Kang, Raphi, et al.
Published: (2025)
by: Kang, Raphi, et al.
Published: (2025)
A Rapid Test for Accuracy and Bias of Face Recognition Technology
by: Knott, Manuel, et al.
Published: (2025)
by: Knott, Manuel, et al.
Published: (2025)
Similar Items
-
SDQM: Synthetic Data Quality Metric for Object Detection Dataset Evaluation
by: Zenith, Ayush, et al.
Published: (2025) -
Tractable Equilibrium Computation in Markov Games through Risk Aversion
by: Mazumdar, Eric, et al.
Published: (2024) -
Model-Free Robust $ϕ$-Divergence Reinforcement Learning Using Both Offline and Online Data
by: Panaganti, Kishan, et al.
Published: (2024) -
TARDIS: Mitigating Temporal Misalignment via Representation Steering
by: Shin, Changho, et al.
Published: (2025) -
Efficient Policy Optimization in Robust Constrained MDPs with Iteration Complexity Guarantees
by: Ganguly, Sourav, et al.
Published: (2025)