Auditing and Generating Synthetic Data with Controllable Trust Trade-offs
Fuente:
arXiv
Guardado en:
| Autores principales: | Belgodere, Brian, Dognin, Pierre, Ivankay, Adam, Melnyk, Igor, Mroueh, Youssef, Mojsilovic, Aleksandra, Navratil, Jiri, Nitsure, Apoorva, Padhi, Inkit, Rigotti, Mattia, Ross, Jerret, Schiff, Yair, Vedpathak, Radhika, Young, Richard A. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Risk Aware Benchmarking of Large Language Models
por: Nitsure, Apoorva, et al.
Publicado: (2023)
por: Nitsure, Apoorva, et al.
Publicado: (2023)
Distributional Preference Alignment of LLMs via Optimal Transport
por: Melnyk, Igor, et al.
Publicado: (2024)
por: Melnyk, Igor, et al.
Publicado: (2024)
Image Captioning as an Assistive Technology: Lessons Learned from VizWiz 2020 Challenge
por: Dognin, Pierre, et al.
Publicado: (2020)
por: Dognin, Pierre, et al.
Publicado: (2020)
Revisiting Group Relative Policy Optimization: Insights into On-Policy and Off-Policy Training
por: Mroueh, Youssef, et al.
Publicado: (2025)
por: Mroueh, Youssef, et al.
Publicado: (2025)
Multivariate Stochastic Dominance via Optimal Transport and Applications to Models Benchmarking
por: Rioux, Gabriel, et al.
Publicado: (2024)
por: Rioux, Gabriel, et al.
Publicado: (2024)
GP-MoLFormer: A Foundation Model For Molecular Generation
por: Ross, Jerret, et al.
Publicado: (2024)
por: Ross, Jerret, et al.
Publicado: (2024)
GP-MoLFormer-Sim: Test Time Molecular Optimization through Contextual Similarity Guidance
por: Navratil, Jiri, et al.
Publicado: (2025)
por: Navratil, Jiri, et al.
Publicado: (2025)
CliffSearch: Structured Agentic Co-Evolution over Theory and Code for Scientific Algorithm Discovery
por: Mroueh, Youssef, et al.
Publicado: (2026)
por: Mroueh, Youssef, et al.
Publicado: (2026)
Value Alignment from Unstructured Text
por: Padhi, Inkit, et al.
Publicado: (2024)
por: Padhi, Inkit, et al.
Publicado: (2024)
Reinforcement Learning with Verifiable Rewards: GRPO's Effective Loss, Dynamics, and Success Amplification
por: Mroueh, Youssef
Publicado: (2025)
por: Mroueh, Youssef
Publicado: (2025)
Information Theoretic Guarantees For Policy Alignment In Large Language Models
por: Mroueh, Youssef
Publicado: (2024)
por: Mroueh, Youssef
Publicado: (2024)
Programming Refusal with Conditional Activation Steering
por: Lee, Bruce W., et al.
Publicado: (2024)
por: Lee, Bruce W., et al.
Publicado: (2024)
When in Doubt, Cascade: Towards Building Efficient and Capable Guardrails
por: Nagireddy, Manish, et al.
Publicado: (2024)
por: Nagireddy, Manish, et al.
Publicado: (2024)
Answering the Wrong Question: Reasoning Trace Inversion for Abstention in LLMs
por: Gourabathina, Abinitha, et al.
Publicado: (2026)
por: Gourabathina, Abinitha, et al.
Publicado: (2026)
Contextual Moral Value Alignment Through Context-Based Aggregation
por: Dognin, Pierre, et al.
Publicado: (2024)
por: Dognin, Pierre, et al.
Publicado: (2024)
Alignment Studio: Aligning Large Language Models to Particular Contextual Regulations
por: Achintalwar, Swapnaja, et al.
Publicado: (2024)
por: Achintalwar, Swapnaja, et al.
Publicado: (2024)
Split, Unlearn, Merge: Leveraging Data Attributes for More Effective Unlearning in LLMs
por: Kadhe, Swanand Ravindra, et al.
Publicado: (2024)
por: Kadhe, Swanand Ravindra, et al.
Publicado: (2024)
The Narasimhan-Seshadri Theorem revisited
por: Nitsure, Nitin
Publicado: (2025)
por: Nitsure, Nitin
Publicado: (2025)
Trade, migration and welfare : the impact of social capital / Maurice Schiff
por: Schiff, Maurice
Publicado: (1999)
por: Schiff, Maurice
Publicado: (1999)
Guided Speculative Inference for Efficient Test-Time Alignment of LLMs
por: Geuter, Jonathan, et al.
Publicado: (2025)
por: Geuter, Jonathan, et al.
Publicado: (2025)
Hétérocères nouveaux de l'Amérique du Sud
por: Dognin, Paul
Publicado: (1901)
por: Dognin, Paul
Publicado: (1901)
Heterocores nouveaux de l'Amerique du Sud
por: Dognin, Paul
Publicado: (1913)
por: Dognin, Paul
Publicado: (1913)
Regional integration and technology diffusion : the case of the North America Free Trade Agreement / Maurice Schiff, Yanling Wang
por: Schiff, Maurice
Publicado: (2003)
por: Schiff, Maurice
Publicado: (2003)
Final-Model-Only Data Attribution with a Unifying View of Gradient-Based Methods
por: Wei, Dennis, et al.
Publicado: (2024)
por: Wei, Dennis, et al.
Publicado: (2024)
GIST: Gauge-Invariant Spectral Transformers for Scalable Graph Neural Operators
por: Rigotti, Mattia, et al.
Publicado: (2026)
por: Rigotti, Mattia, et al.
Publicado: (2026)
Eliciting Reasoning in Language Models with Cognitive Tools
por: Ebouky, Brown, et al.
Publicado: (2025)
por: Ebouky, Brown, et al.
Publicado: (2025)
Verify when Uncertain: Beyond Self-Consistency in Black Box Hallucination Detection
por: Xue, Yihao, et al.
Publicado: (2025)
por: Xue, Yihao, et al.
Publicado: (2025)
Evaluation of medication adherence among Lebanese diabetic patients
por: Lara Mroueh
Publicado: (2018)
por: Lara Mroueh
Publicado: (2018)
KL-Regularized RLHF with Multiple Reference Models: Exact Solutions and Sample Complexity
por: Aminian, Gholamali, et al.
Publicado: (2025)
por: Aminian, Gholamali, et al.
Publicado: (2025)
Optimization and Mechanistic Insights of Zinc Ascorbate Catalyst for Ring‐Opening Polymerization of Caprolactone Using RSM Methodology and DFT Calculations
por: Sonali S. Naik, et al.
Publicado: (2025)
por: Sonali S. Naik, et al.
Publicado: (2025)
VHDL-Eval: A Framework for Evaluating Large Language Models in VHDL Code Generation
por: Vijayaraghavan, Prashanth, et al.
Publicado: (2024)
por: Vijayaraghavan, Prashanth, et al.
Publicado: (2024)
Outline-Guided Object Inpainting with Diffusion Models
por: Pobitzer, Markus, et al.
Publicado: (2024)
por: Pobitzer, Markus, et al.
Publicado: (2024)
Nerve function impairment and quality of life in patients with leprosy: a prospective, observational study
por: Apoorva Sharma, et al.
Publicado: (2024)
por: Apoorva Sharma, et al.
Publicado: (2024)
Rh Potenziale Moiré come Operatore di Scattering e il confinamento topologico sulla varietà di Klein
por: Rigotti, Alex
Publicado: (2026)
por: Rigotti, Alex
Publicado: (2026)
El Conflicto del Campo. Matrices culturales e identificaciones políticas
por: Sebastián Rigotti
Publicado: (2014)
por: Sebastián Rigotti
Publicado: (2014)
Explain First, Trust Later: LLM-Augmented Explanations for Graph-Based Crypto Anomaly Detection
por: Watson, Adriana, et al.
Publicado: (2025)
por: Watson, Adriana, et al.
Publicado: (2025)
WikiContradict: A Benchmark for Evaluating LLMs on Real-World Knowledge Conflicts from Wikipedia
por: Hou, Yufang, et al.
Publicado: (2024)
por: Hou, Yufang, et al.
Publicado: (2024)
FESTA: Functionally Equivalent Sampling for Trust Assessment of Multimodal LLMs
por: Bhattacharya, Debarpan, et al.
Publicado: (2025)
por: Bhattacharya, Debarpan, et al.
Publicado: (2025)
Gradient Flows and Riemannian Structure in the Gromov-Wasserstein Geometry
por: Zhang, Zhengxin, et al.
Publicado: (2024)
por: Zhang, Zhengxin, et al.
Publicado: (2024)
Best-of-N through the Smoothing Lens: KL Divergence and Regret Analysis
por: Aminian, Gholamali, et al.
Publicado: (2025)
por: Aminian, Gholamali, et al.
Publicado: (2025)
Ejemplares similares
-
Risk Aware Benchmarking of Large Language Models
por: Nitsure, Apoorva, et al.
Publicado: (2023) -
Distributional Preference Alignment of LLMs via Optimal Transport
por: Melnyk, Igor, et al.
Publicado: (2024) -
Image Captioning as an Assistive Technology: Lessons Learned from VizWiz 2020 Challenge
por: Dognin, Pierre, et al.
Publicado: (2020) -
Revisiting Group Relative Policy Optimization: Insights into On-Policy and Off-Policy Training
por: Mroueh, Youssef, et al.
Publicado: (2025) -
Multivariate Stochastic Dominance via Optimal Transport and Applications to Models Benchmarking
por: Rioux, Gabriel, et al.
Publicado: (2024)