Aligning Language Models with Observational Data: Opportunities and Risks from a Causal Perspective
Fuente:
arXiv
Salvato in:
| Autore principale: | Loghmani, Erfan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Big Data Approach to Understand Sub-national Determinants of FDI in Africa
di: Colladon, A. Fronzetti, et al.
Pubblicazione: (2024)
di: Colladon, A. Fronzetti, et al.
Pubblicazione: (2024)
Repetition Without Exclusivity: Scale Sensitivity of Referential Mechanisms in Child-Scale Language Models
di: Cacioli, Jon-Paul
Pubblicazione: (2026)
di: Cacioli, Jon-Paul
Pubblicazione: (2026)
Stroke Lesions as a Rosetta Stone for Language Model Interpretability
di: Fridriksson, Julius, et al.
Pubblicazione: (2026)
di: Fridriksson, Julius, et al.
Pubblicazione: (2026)
Leveraging Natural Language Processing and Machine Learning for Evidence-Based Food Security Policy Decision-Making in Data-Scarce Making
di: Singh, Karan Kumar, et al.
Pubblicazione: (2026)
di: Singh, Karan Kumar, et al.
Pubblicazione: (2026)
Measure what Matters: Psychometric Evaluation of AI with Situational Judgment Tests
di: Yost, Alexandra, et al.
Pubblicazione: (2025)
di: Yost, Alexandra, et al.
Pubblicazione: (2025)
On the Limits of LLM Adaptability: Impact of Model-Internalized Priors on Annotation Task Performance
di: Casanova, Etienne, et al.
Pubblicazione: (2026)
di: Casanova, Etienne, et al.
Pubblicazione: (2026)
FlightSense: An End-to-End MLOps Platform for Real-Time Flight Delay Prediction via Rotation-Chain Propagation Features and Agentic Conversational AI
di: Shelke, Aditi J., et al.
Pubblicazione: (2026)
di: Shelke, Aditi J., et al.
Pubblicazione: (2026)
Super-additive Cooperation in Language Model Agents
di: Tonini, Filippo, et al.
Pubblicazione: (2025)
di: Tonini, Filippo, et al.
Pubblicazione: (2025)
A Domain-Agnostic Neurosymbolic Approach for Big Social Data Analysis: Evaluating Mental Health Sentiment on Social Media during COVID-19
di: Khandelwal, Vedant, et al.
Pubblicazione: (2024)
di: Khandelwal, Vedant, et al.
Pubblicazione: (2024)
Language Model Guided Reinforcement Learning in Quantitative Trading
di: Darmanin, Adam, et al.
Pubblicazione: (2025)
di: Darmanin, Adam, et al.
Pubblicazione: (2025)
Robustness of Spatio-temporal Graph Neural Networks for Fault Location in Partially Observable Distribution Grids
di: Karabulut, Burak, et al.
Pubblicazione: (2026)
di: Karabulut, Burak, et al.
Pubblicazione: (2026)
Evo-DKD: Dual-Knowledge Decoding for Autonomous Ontology Evolution in Large Language Models
di: Raman, Vishal, et al.
Pubblicazione: (2025)
di: Raman, Vishal, et al.
Pubblicazione: (2025)
Context-dependent Causality (the Non-Nonotonic Case)
di: Billfeld, Nir, et al.
Pubblicazione: (2024)
di: Billfeld, Nir, et al.
Pubblicazione: (2024)
Automated Circuit Interpretation via Probe Prompting
di: Birardi, Giuseppe
Pubblicazione: (2025)
di: Birardi, Giuseppe
Pubblicazione: (2025)
Knowing Isn't Understanding: Re-grounding Generative Proactivity with Epistemic and Behavioral Insight
di: Kaur, Kirandeep, et al.
Pubblicazione: (2026)
di: Kaur, Kirandeep, et al.
Pubblicazione: (2026)
Reward Model Interpretability via Optimal and Pessimal Tokens
di: Christian, Brian, et al.
Pubblicazione: (2025)
di: Christian, Brian, et al.
Pubblicazione: (2025)
Categorical Perception in Large Language Model Hidden States: Structural Warping at Digit-Count Boundaries
di: Cacioli, Jon-Paul
Pubblicazione: (2026)
di: Cacioli, Jon-Paul
Pubblicazione: (2026)
Evaluation of Hate Speech Detection Using Large Language Models and Geographical Contextualization
di: Zahid, Anwar Hossain, et al.
Pubblicazione: (2025)
di: Zahid, Anwar Hossain, et al.
Pubblicazione: (2025)
From Syntax to Semantics: Unveiling the Emergence of Chirality in SMILES Translation Models
di: Li, Zehao, et al.
Pubblicazione: (2026)
di: Li, Zehao, et al.
Pubblicazione: (2026)
Auditing Marketing Budget Allocation with Hindsight Regret
di: Pathak, Nilavra, et al.
Pubblicazione: (2026)
di: Pathak, Nilavra, et al.
Pubblicazione: (2026)
Analysis of LLM as a grammatical feature tagger for African American English
di: Porwal, Rahul, et al.
Pubblicazione: (2025)
di: Porwal, Rahul, et al.
Pubblicazione: (2025)
RouteNLP: Closed-Loop LLM Routing with Conformal Cascading and Distillation Co-Optimization
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
Computers as Bad Social Actors: Dark Patterns and Anti-Patterns in Interfaces that Act Socially
di: Alberts, Lize, et al.
Pubblicazione: (2023)
di: Alberts, Lize, et al.
Pubblicazione: (2023)
A Benchmark of Classical and Deep Learning Models for Agricultural Commodity Price Forecasting on A Novel Bangladeshi Market Price Dataset
di: Muhammad, Tashreef, et al.
Pubblicazione: (2026)
di: Muhammad, Tashreef, et al.
Pubblicazione: (2026)
OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models
di: Cotti, Luca, et al.
Pubblicazione: (2025)
di: Cotti, Luca, et al.
Pubblicazione: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
di: Fadli, Samih
Pubblicazione: (2025)
di: Fadli, Samih
Pubblicazione: (2025)
ComplianceNLP: Knowledge-Graph-Augmented RAG for Multi-Framework Regulatory Gap Detection
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
Interpretable Clustering: A Survey
di: Hu, Lianyu, et al.
Pubblicazione: (2024)
di: Hu, Lianyu, et al.
Pubblicazione: (2024)
Merge-Bench: Resolve Merge Conflicts with Large Language Models
di: Schesch, Benedikt, et al.
Pubblicazione: (2026)
di: Schesch, Benedikt, et al.
Pubblicazione: (2026)
Learning What Matters: Probabilistic Task Selection via Mutual Information for Model Finetuning
di: Chanda, Prateek, et al.
Pubblicazione: (2025)
di: Chanda, Prateek, et al.
Pubblicazione: (2025)
Learning, Fast and Slow: Towards LLMs That Adapt Continually
di: Tiwari, Rishabh, et al.
Pubblicazione: (2026)
di: Tiwari, Rishabh, et al.
Pubblicazione: (2026)
Better Schedules for Low Precision Training of Deep Neural Networks
di: Wolfe, Cameron R., et al.
Pubblicazione: (2024)
di: Wolfe, Cameron R., et al.
Pubblicazione: (2024)
Explainable Classifier for Malignant Lymphoma Subtyping via Cell Graph and Image Fusion
di: Nishiyama, Daiki, et al.
Pubblicazione: (2025)
di: Nishiyama, Daiki, et al.
Pubblicazione: (2025)
Extreme Self-Preference in Language Models
di: Lehr, Steven A., et al.
Pubblicazione: (2025)
di: Lehr, Steven A., et al.
Pubblicazione: (2025)
propella-1: Multi-Property Document Annotation for LLM Data Curation at Scale
di: Idahl, Maximilian, et al.
Pubblicazione: (2026)
di: Idahl, Maximilian, et al.
Pubblicazione: (2026)
Survey Transfer Learning: Recycling Data with Silicon Responses
di: Amini, Ali
Pubblicazione: (2025)
di: Amini, Ali
Pubblicazione: (2025)
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring
di: Heyman, Alex, et al.
Pubblicazione: (2025)
di: Heyman, Alex, et al.
Pubblicazione: (2025)
Unveiling the Secret Recipe: A Guide For Supervised Fine-Tuning Small LLMs
di: Pareja, Aldo, et al.
Pubblicazione: (2024)
di: Pareja, Aldo, et al.
Pubblicazione: (2024)
HIP Network: Historical Information Passing Network for Extrapolation Reasoning on Temporal Knowledge Graph
di: He, Yongquan, et al.
Pubblicazione: (2024)
di: He, Yongquan, et al.
Pubblicazione: (2024)
SAGE: A Strategy-Aware Graph-Enhanced Generation Framework For Online Counseling
di: Aharon, Eliya Naomi, et al.
Pubblicazione: (2026)
di: Aharon, Eliya Naomi, et al.
Pubblicazione: (2026)
Documenti analoghi
-
A Big Data Approach to Understand Sub-national Determinants of FDI in Africa
di: Colladon, A. Fronzetti, et al.
Pubblicazione: (2024) -
Repetition Without Exclusivity: Scale Sensitivity of Referential Mechanisms in Child-Scale Language Models
di: Cacioli, Jon-Paul
Pubblicazione: (2026) -
Stroke Lesions as a Rosetta Stone for Language Model Interpretability
di: Fridriksson, Julius, et al.
Pubblicazione: (2026) -
Leveraging Natural Language Processing and Machine Learning for Evidence-Based Food Security Policy Decision-Making in Data-Scarce Making
di: Singh, Karan Kumar, et al.
Pubblicazione: (2026) -
Measure what Matters: Psychometric Evaluation of AI with Situational Judgment Tests
di: Yost, Alexandra, et al.
Pubblicazione: (2025)