Towards Operationalizing Right to Data Protection
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Java, Abhinav, Shahid, Simra, Agarwal, Chirag |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Towards Understanding the Robustness of Sparse Autoencoders
par: Saiyed, Ahson, et autres
Publié: (2026)
par: Saiyed, Ahson, et autres
Publié: (2026)
Toward Understanding Unlearning Difficulty: A Mechanistic Perspective and Circuit-Guided Difficulty Metric
par: Cheng, Jiali, et autres
Publié: (2026)
par: Cheng, Jiali, et autres
Publié: (2026)
In-Context Explainers: Harnessing LLMs for Explaining Black Box Models
par: Kroeger, Nicholas, et autres
Publié: (2023)
par: Kroeger, Nicholas, et autres
Publié: (2023)
Do Students Debias Like Teachers? On the Distillability of Bias Mitigation Methods
par: Cheng, Jiali, et autres
Publié: (2025)
par: Cheng, Jiali, et autres
Publié: (2025)
Towards Quantifying Commonsense Reasoning with Mechanistic Insights
par: Joshi, Abhinav, et autres
Publié: (2025)
par: Joshi, Abhinav, et autres
Publié: (2025)
Towards Robust Evaluation of Unlearning in LLMs via Data Transformations
par: Joshi, Abhinav, et autres
Publié: (2024)
par: Joshi, Abhinav, et autres
Publié: (2024)
Agnostic Language Identification and Generation
par: Høgsgaard, Mikael Møller, et autres
Publié: (2026)
par: Høgsgaard, Mikael Møller, et autres
Publié: (2026)
DiscoveryBench: Towards Data-Driven Discovery with Large Language Models
par: Majumder, Bodhisattwa Prasad, et autres
Publié: (2024)
par: Majumder, Bodhisattwa Prasad, et autres
Publié: (2024)
Neural Networks for Learnable and Scalable Influence Estimation of Instruction Fine-Tuning Data
par: Agarwal, Ishika, et autres
Publié: (2025)
par: Agarwal, Ishika, et autres
Publié: (2025)
Meursault as a Data Point
par: Pratap, Abhinav
Publié: (2025)
par: Pratap, Abhinav
Publié: (2025)
How Reliable are Causal Probing Interventions?
par: Canby, Marc, et autres
Publié: (2024)
par: Canby, Marc, et autres
Publié: (2024)
Certifying LLM Safety against Adversarial Prompting
par: Kumar, Aounon, et autres
Publié: (2023)
par: Kumar, Aounon, et autres
Publié: (2023)
Rethinking Explainability in the Era of Multimodal AI
par: Agarwal, Chirag
Publié: (2025)
par: Agarwal, Chirag
Publié: (2025)
The Right Time Matters: Data Arrangement Affects Zero-Shot Generalization in Instruction Tuning
par: He, Bingxiang, et autres
Publié: (2024)
par: He, Bingxiang, et autres
Publié: (2024)
Feedback-Aware Monte Carlo Tree Search for Efficient Information Seeking in Goal-Oriented Conversations
par: Chopra, Harshita, et autres
Publié: (2025)
par: Chopra, Harshita, et autres
Publié: (2025)
Operationalizing the Blueprint for an AI Bill of Rights: Recommendations for Practitioners, Researchers, and Policy Makers
par: Oesterling, Alex, et autres
Publié: (2024)
par: Oesterling, Alex, et autres
Publié: (2024)
Data-driven Discovery with Large Generative Models
par: Majumder, Bodhisattwa Prasad, et autres
Publié: (2024)
par: Majumder, Bodhisattwa Prasad, et autres
Publié: (2024)
COLD: Causal reasOning in cLosed Daily activities
par: Joshi, Abhinav, et autres
Publié: (2024)
par: Joshi, Abhinav, et autres
Publié: (2024)
Beyond Components: Singular Vector-Based Interpretability of Transformer Circuits
par: Ahmad, Areeb, et autres
Publié: (2025)
par: Ahmad, Areeb, et autres
Publié: (2025)
Geometry of Decision Making in Language Models
par: Joshi, Abhinav, et autres
Publié: (2025)
par: Joshi, Abhinav, et autres
Publié: (2025)
Calibration Across Layers: Understanding Calibration Evolution in LLMs
par: Joshi, Abhinav, et autres
Publié: (2025)
par: Joshi, Abhinav, et autres
Publié: (2025)
AutoEval Done Right: Using Synthetic Data for Model Evaluation
par: Boyeau, Pierre, et autres
Publié: (2024)
par: Boyeau, Pierre, et autres
Publié: (2024)
Languages are Modalities: Cross-Lingual Alignment via Encoder Injection
par: Agarwal, Rajan, et autres
Publié: (2025)
par: Agarwal, Rajan, et autres
Publié: (2025)
Calibrating LLMs for Text-to-SQL Parsing by Leveraging Sub-clause Frequencies
par: Liu, Terrance, et autres
Publié: (2025)
par: Liu, Terrance, et autres
Publié: (2025)
TrICy: Trigger-guided Data-to-text Generation with Intent aware Attention-Copy
par: Agarwal, Vibhav, et autres
Publié: (2024)
par: Agarwal, Vibhav, et autres
Publié: (2024)
Towards Compute-Optimal Many-Shot In-Context Learning
par: Golchin, Shahriar, et autres
Publié: (2025)
par: Golchin, Shahriar, et autres
Publié: (2025)
AcquisitionSynthesis: Targeted Data Generation using Acquisition Functions
par: Agarwal, Ishika, et autres
Publié: (2026)
par: Agarwal, Ishika, et autres
Publié: (2026)
Representation Learning of Structured Data for Medical Foundation Models
par: Dwivedi, Vijay Prakash, et autres
Publié: (2024)
par: Dwivedi, Vijay Prakash, et autres
Publié: (2024)
Exploring Facets of Language Generation in the Limit
par: Charikar, Moses, et autres
Publié: (2024)
par: Charikar, Moses, et autres
Publié: (2024)
Pareto-optimal Non-uniform Language Generation
par: Charikar, Moses, et autres
Publié: (2025)
par: Charikar, Moses, et autres
Publié: (2025)
On the Emergence of Thinking in LLMs I: Searching for the Right Intuition
par: Ye, Guanghao, et autres
Publié: (2025)
par: Ye, Guanghao, et autres
Publié: (2025)
SAEs Are Good for Steering -- If You Select the Right Features
par: Arad, Dana, et autres
Publié: (2025)
par: Arad, Dana, et autres
Publié: (2025)
Perplexity Cannot Always Tell Right from Wrong
par: Veličković, Petar, et autres
Publié: (2026)
par: Veličković, Petar, et autres
Publié: (2026)
Towards a Zero-Data, Controllable, Adaptive Dialog System
par: Väth, Dirk, et autres
Publié: (2024)
par: Väth, Dirk, et autres
Publié: (2024)
Think Inside the JSON: Reinforcement Strategy for Strict LLM Schema Adherence
par: Agarwal, Bhavik, et autres
Publié: (2025)
par: Agarwal, Bhavik, et autres
Publié: (2025)
G-Loss: Graph-Guided Fine-Tuning of Language Models
par: Sharma, Aditya, et autres
Publié: (2026)
par: Sharma, Aditya, et autres
Publié: (2026)
Operationalizing AI: Empirical Evidence on MLOps Practices, User Satisfaction, and Organizational Context
par: Pasch, Stefan
Publié: (2025)
par: Pasch, Stefan
Publié: (2025)
Thinking Fair and Slow: On the Efficacy of Structured Prompts for Debiasing Language Models
par: Furniturewala, Shaz, et autres
Publié: (2024)
par: Furniturewala, Shaz, et autres
Publié: (2024)
Towards Data-Centric RLHF: Simple Metrics for Preference Dataset Comparison
par: Shen, Judy Hanwen, et autres
Publié: (2024)
par: Shen, Judy Hanwen, et autres
Publié: (2024)
RAG-Modulo: Solving Sequential Tasks using Experience, Critics, and Language Models
par: Jain, Abhinav, et autres
Publié: (2024)
par: Jain, Abhinav, et autres
Publié: (2024)
Documents similaires
-
Towards Understanding the Robustness of Sparse Autoencoders
par: Saiyed, Ahson, et autres
Publié: (2026) -
Toward Understanding Unlearning Difficulty: A Mechanistic Perspective and Circuit-Guided Difficulty Metric
par: Cheng, Jiali, et autres
Publié: (2026) -
In-Context Explainers: Harnessing LLMs for Explaining Black Box Models
par: Kroeger, Nicholas, et autres
Publié: (2023) -
Do Students Debias Like Teachers? On the Distillability of Bias Mitigation Methods
par: Cheng, Jiali, et autres
Publié: (2025) -
Towards Quantifying Commonsense Reasoning with Mechanistic Insights
par: Joshi, Abhinav, et autres
Publié: (2025)