Guardado en:
| Autores principales: | Ge, Luise, Zhang, Yongyan, Vorobeychik, Yevgeniy |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.15173 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Optimized Distortion in Linear Social Choice
por: Ge, Luise, et al.
Publicado: (2025)
por: Ge, Luise, et al.
Publicado: (2025)
Linear Social Choice with Few Queries: A Moment-Based Approach
por: Ge, Luise, et al.
Publicado: (2026)
por: Ge, Luise, et al.
Publicado: (2026)
Learning Linear Utility Functions From Pairwise Comparison Queries
por: Ge, Luise, et al.
Publicado: (2024)
por: Ge, Luise, et al.
Publicado: (2024)
Learning Policy Committees for Effective Personalization in MDPs with Diverse Tasks
por: Ge, Luise, et al.
Publicado: (2025)
por: Ge, Luise, et al.
Publicado: (2025)
CoFineLLM: Conformal Finetuning of LLMs for Language-Instructed Robot Planning
por: Wang, Jun, et al.
Publicado: (2025)
por: Wang, Jun, et al.
Publicado: (2025)
Verified Safe Reinforcement Learning for Neural Network Dynamic Models
por: Wu, Junlin, et al.
Publicado: (2024)
por: Wu, Junlin, et al.
Publicado: (2024)
Adversarial Reinforcement Learning for Detecting False Data Injection Attacks in Vehicular Routing
por: Eghtesad, Taha, et al.
Publicado: (2026)
por: Eghtesad, Taha, et al.
Publicado: (2026)
Online Feedback Efficient Active Target Discovery in Partially Observable Environments
por: Sarkar, Anindya, et al.
Publicado: (2025)
por: Sarkar, Anindya, et al.
Publicado: (2025)
Protecting Language Models Against Unauthorized Distillation through Trace Rewriting
por: Ma, Xinhang, et al.
Publicado: (2026)
por: Ma, Xinhang, et al.
Publicado: (2026)
Axioms for AI Alignment from Human Feedback
por: Ge, Luise, et al.
Publicado: (2024)
por: Ge, Luise, et al.
Publicado: (2024)
Conformal Reachability for Safe Control in Unknown Environments
por: Ma, Xinhang, et al.
Publicado: (2026)
por: Ma, Xinhang, et al.
Publicado: (2026)
Multi-Agent Reinforcement Learning for Assessing False-Data Injection Attacks on Transportation Networks
por: Eghtesad, Taha, et al.
Publicado: (2023)
por: Eghtesad, Taha, et al.
Publicado: (2023)
Learning Interpretable Policies in Hindsight-Observable POMDPs through Partially Supervised Reinforcement Learning
por: Lanier, Michael, et al.
Publicado: (2024)
por: Lanier, Michael, et al.
Publicado: (2024)
Learning Vision-Based Neural Network Controllers with Semi-Probabilistic Safety Guarantees
por: Ma, Xinhang, et al.
Publicado: (2025)
por: Ma, Xinhang, et al.
Publicado: (2025)
Conformal Temporal Logic Planning using Large Language Models
por: Wang, Jun, et al.
Publicado: (2023)
por: Wang, Jun, et al.
Publicado: (2023)
DiffVAS: Diffusion-Guided Visual Active Search in Partially Observable Environments
por: Sarkar, Anindya, et al.
Publicado: (2026)
por: Sarkar, Anindya, et al.
Publicado: (2026)
Preference Poisoning Attacks on Reward Model Learning
por: Wu, Junlin, et al.
Publicado: (2024)
por: Wu, Junlin, et al.
Publicado: (2024)
GOMAA-Geo: GOal Modality Agnostic Active Geo-localization
por: Sarkar, Anindya, et al.
Publicado: (2024)
por: Sarkar, Anindya, et al.
Publicado: (2024)
Learned Neighbor Trust for Collaborative Deployment in Model-Agnostic Decentralized Learning
por: Lanier, Michael, et al.
Publicado: (2026)
por: Lanier, Michael, et al.
Publicado: (2026)
RLHFPoison: Reward Poisoning Attack for Reinforcement Learning with Human Feedback in Large Language Models
por: Wang, Jiongxiao, et al.
Publicado: (2023)
por: Wang, Jiongxiao, et al.
Publicado: (2023)
Mind the Gap! Choice Independence in Using Multilingual LLMs for Persuasive Co-Writing Tasks in Different Languages
por: Biswas, Shreyan, et al.
Publicado: (2025)
por: Biswas, Shreyan, et al.
Publicado: (2025)
AutoDAN-Turbo: A Lifelong Agent for Strategy Self-Exploration to Jailbreak LLMs
por: Liu, Xiaogeng, et al.
Publicado: (2024)
por: Liu, Xiaogeng, et al.
Publicado: (2024)
COSMOS: Model-Agnostic Personalized Federated Learning with Clustered Server Models and Pseudo-Label-Only Communication
por: Rachmut, Ben, et al.
Publicado: (2026)
por: Rachmut, Ben, et al.
Publicado: (2026)
Mind the Gap: The Divergence Between Human and LLM-Generated Tasks
por: Lu, Yi-Long, et al.
Publicado: (2025)
por: Lu, Yi-Long, et al.
Publicado: (2025)
Mind the Gap Between Spatial Reasoning and Acting! Step-by-Step Evaluation of Agents With Spatial-Gym
por: Kaesberg, Lars Benedikt, et al.
Publicado: (2026)
por: Kaesberg, Lars Benedikt, et al.
Publicado: (2026)
Active Geospatial Search for Efficient Tenant Eviction Outreach
por: Sarkar, Anindya, et al.
Publicado: (2024)
por: Sarkar, Anindya, et al.
Publicado: (2024)
Mind the (Belief) Gap: Group Identity in the World of LLMs
por: Borah, Angana, et al.
Publicado: (2025)
por: Borah, Angana, et al.
Publicado: (2025)
Mind The Gap: Quantifying Mechanistic Gaps in Algorithmic Reasoning via Neural Compilation
por: Saldyt, Lucas, et al.
Publicado: (2025)
por: Saldyt, Lucas, et al.
Publicado: (2025)
Language Models Trained to do Arithmetic Predict Human Risky and Intertemporal Choice
por: Zhu, Jian-Qiao, et al.
Publicado: (2024)
por: Zhu, Jian-Qiao, et al.
Publicado: (2024)
MindGap: A Conversational AI Framework for Upstream Neuroplastic Intervention in Post-Traumatic Stress Disorder
por: Bandara, Eranga, et al.
Publicado: (2026)
por: Bandara, Eranga, et al.
Publicado: (2026)
Mind the Data Gap: Bridging LLMs to Enterprise Data Integration
por: Kayali, Moe, et al.
Publicado: (2024)
por: Kayali, Moe, et al.
Publicado: (2024)
Sliced Rényi Pufferfish Privacy: Directional Additive Noise Mechanism and Private Learning with Gradient Clipping
por: Zhang, Tao, et al.
Publicado: (2025)
por: Zhang, Tao, et al.
Publicado: (2025)
Residual-PAC Privacy: Automatic Privacy Control Beyond the Gaussian Barrier
por: Zhang, Tao, et al.
Publicado: (2025)
por: Zhang, Tao, et al.
Publicado: (2025)
Mind the Gap: Bridging the Divide Between AI Aspirations and the Reality of Autonomous Characterization
por: Guinan, Grace, et al.
Publicado: (2025)
por: Guinan, Grace, et al.
Publicado: (2025)
A Scalable Approach to Solving Simulation-Based Network Security Games
por: Lanier, Michael, et al.
Publicado: (2026)
por: Lanier, Michael, et al.
Publicado: (2026)
CyGym: A Simulation-Based Game-Theoretic Analysis Framework for Cybersecurity
por: Lanier, Michael, et al.
Publicado: (2025)
por: Lanier, Michael, et al.
Publicado: (2025)
Positive and Risky Message Assessment for Music Products
por: Zhang, Yigeng, et al.
Publicado: (2023)
por: Zhang, Yigeng, et al.
Publicado: (2023)
InMind: Evaluating LLMs in Capturing and Applying Individual Human Reasoning Styles
por: Li, Zizhen, et al.
Publicado: (2025)
por: Li, Zizhen, et al.
Publicado: (2025)
Mind the Value-Action Gap: Do LLMs Act in Alignment with Their Values?
por: Shen, Hua, et al.
Publicado: (2025)
por: Shen, Hua, et al.
Publicado: (2025)
MultiMind: Enhancing Werewolf Agents with Multimodal Reasoning and Theory of Mind
por: Zhang, Zheng, et al.
Publicado: (2025)
por: Zhang, Zheng, et al.
Publicado: (2025)
Ejemplares similares
-
Optimized Distortion in Linear Social Choice
por: Ge, Luise, et al.
Publicado: (2025) -
Linear Social Choice with Few Queries: A Moment-Based Approach
por: Ge, Luise, et al.
Publicado: (2026) -
Learning Linear Utility Functions From Pairwise Comparison Queries
por: Ge, Luise, et al.
Publicado: (2024) -
Learning Policy Committees for Effective Personalization in MDPs with Diverse Tasks
por: Ge, Luise, et al.
Publicado: (2025) -
CoFineLLM: Conformal Finetuning of LLMs for Language-Instructed Robot Planning
por: Wang, Jun, et al.
Publicado: (2025)