Experiments with truth using Machine Learning: Spectral analysis and explainable classification of synthetic, false, and genuine information
Fuente:
arXiv
Saved in:
| Main Authors: | Pendyala, Vishnu S., Dutta, Madhulika |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Human and the Mechanical: logos, truthfulness, and ChatGPT
by: Giannakidou, Anastasia, et al.
Published: (2024)
by: Giannakidou, Anastasia, et al.
Published: (2024)
Exploring the generalization of LLM truth directions on conversational formats
by: Ichmoukhamedov, Timour, et al.
Published: (2025)
by: Ichmoukhamedov, Timour, et al.
Published: (2025)
Evaluating and explaining training strategies for zero-shot cross-lingual news sentiment analysis
by: Andrenšek, Luka, et al.
Published: (2024)
by: Andrenšek, Luka, et al.
Published: (2024)
An overview of model uncertainty and variability in LLM-based sentiment analysis. Challenges, mitigation strategies and the role of explainability
by: Herrera-Poyatos, David, et al.
Published: (2025)
by: Herrera-Poyatos, David, et al.
Published: (2025)
Illuminate: A novel approach for depression detection with explainable analysis and proactive therapy using prompt engineering
by: Agrawal, Aryan
Published: (2024)
by: Agrawal, Aryan
Published: (2024)
Tell me the truth: A system to measure the trustworthiness of Large Language Models
by: Lipizzi, Carlo
Published: (2024)
by: Lipizzi, Carlo
Published: (2024)
HITgram: A Platform for Experimenting with n-gram Language Models
by: Dasgupta, Shibaranjani, et al.
Published: (2024)
by: Dasgupta, Shibaranjani, et al.
Published: (2024)
AugSumm: towards generalizable speech summarization using synthetic labels from large language model
by: Jung, Jee-weon, et al.
Published: (2024)
by: Jung, Jee-weon, et al.
Published: (2024)
Controllable and explainable personality sliders for LLMs at inference time
by: Hoppe, Florian, et al.
Published: (2026)
by: Hoppe, Florian, et al.
Published: (2026)
Towards Understanding and Improving Refusal in Compressed Models via Mechanistic Interpretability
by: Chhabra, Vishnu Kabir, et al.
Published: (2025)
by: Chhabra, Vishnu Kabir, et al.
Published: (2025)
GRAFT: A Graph-based Flow-aware Agentic Framework for Document-level Machine Translation
by: Dutta, Himanshu, et al.
Published: (2025)
by: Dutta, Himanshu, et al.
Published: (2025)
CARE: A QLoRA-Fine Tuned Multi-Domain Chatbot With Fast Learning On Minimal Hardware
by: Dutta, Ankit, et al.
Published: (2025)
by: Dutta, Ankit, et al.
Published: (2025)
Identification of Potentially Misclassified Crash Narratives using Machine Learning (ML) and Deep Learning (DL)
by: Bhagat, Sudesh, et al.
Published: (2025)
by: Bhagat, Sudesh, et al.
Published: (2025)
Verbalizing LLMs' assumptions to explain and control sycophancy
by: Cheng, Myra, et al.
Published: (2026)
by: Cheng, Myra, et al.
Published: (2026)
Auditing medical multi-agent AI reveals risks of false consensus
by: Zhu, Yinghao, et al.
Published: (2025)
by: Zhu, Yinghao, et al.
Published: (2025)
Exploring Persona-dependent LLM Alignment for the Moral Machine Experiment
by: Kim, Jiseon, et al.
Published: (2025)
by: Kim, Jiseon, et al.
Published: (2025)
MuCRASP: Multimodal Chain-of-thought Reasoning aware Structured Pruning
by: Dutta, Aritra, et al.
Published: (2026)
by: Dutta, Aritra, et al.
Published: (2026)
An explainable transformer circuit for compositional generalization
by: Tang, Cheng, et al.
Published: (2025)
by: Tang, Cheng, et al.
Published: (2025)
PANORAMA: A synthetic PII-laced dataset for studying sensitive data memorization in LLMs
by: Selvam, Sriram, et al.
Published: (2025)
by: Selvam, Sriram, et al.
Published: (2025)
Uniform Discretized Integrated Gradients: An effective attribution based method for explaining large language models
by: Roy, Swarnava Sinha, et al.
Published: (2024)
by: Roy, Swarnava Sinha, et al.
Published: (2024)
A novel hallucination classification framework
by: Zavhorodnii, Maksym, et al.
Published: (2025)
by: Zavhorodnii, Maksym, et al.
Published: (2025)
Illusion or Algorithm? Investigating Memorization, Emergence, and Symbolic Processing in In-Context Learning
by: Niu, Jingcheng, et al.
Published: (2025)
by: Niu, Jingcheng, et al.
Published: (2025)
A Path Towards Legal Autonomy: An interoperable and explainable approach to extracting, transforming, loading and computing legal information using large language models, expert systems and Bayesian networks
by: Constant, Axel, et al.
Published: (2024)
by: Constant, Axel, et al.
Published: (2024)
CoT-Self-Instruct: Building high-quality synthetic prompts for reasoning and non-reasoning tasks
by: Yu, Ping, et al.
Published: (2025)
by: Yu, Ping, et al.
Published: (2025)
Sentiment Analysis and Emotion Classification using Machine Learning Techniques for Nagamese Language -- A Low-resource Language
by: Morang, Ekha, et al.
Published: (2025)
by: Morang, Ekha, et al.
Published: (2025)
Towards Autonomous Agents: Adaptive-planning, Reasoning, and Acting in Language Models
by: Dutta, Abhishek, et al.
Published: (2024)
by: Dutta, Abhishek, et al.
Published: (2024)
Unveiling Reasoning Thresholds in Language Models: Scaling, Fine-Tuning, and Interpretability through Attention Maps
by: Hsiao, Yen-Che, et al.
Published: (2025)
by: Hsiao, Yen-Che, et al.
Published: (2025)
Shayona@SMM4H23: COVID-19 Self diagnosis classification using BERT and LightGBM models
by: Chavda, Rushi, et al.
Published: (2024)
by: Chavda, Rushi, et al.
Published: (2024)
LLMCARE: early detection of cognitive impairment via transformer models enhanced by LLM-generated synthetic data
by: Zolnour, Ali, et al.
Published: (2025)
by: Zolnour, Ali, et al.
Published: (2025)
On the Shortcut Learning in Multilingual Neural Machine Translation
by: Wang, Wenxuan, et al.
Published: (2024)
by: Wang, Wenxuan, et al.
Published: (2024)
Curation of a Palaeohispanic Dataset for Machine Learning
by: Martínez-Fernández, Gonzalo, et al.
Published: (2026)
by: Martínez-Fernández, Gonzalo, et al.
Published: (2026)
Larger models yield better results? Streamlined severity classification of ADHD-related concerns using BERT-based knowledge distillation
by: Karim, Ahmed Akib Jawad, et al.
Published: (2024)
by: Karim, Ahmed Akib Jawad, et al.
Published: (2024)
Decoding Multilingual Moral Preferences: Unveiling LLM's Biases Through the Moral Machine Experiment
by: Vida, Karina, et al.
Published: (2024)
by: Vida, Karina, et al.
Published: (2024)
Intelligent Scientific Literature Explorer using Machine Learning (ISLE)
by: Jani, Sina, et al.
Published: (2025)
by: Jani, Sina, et al.
Published: (2025)
OSCAR: Orchestrated Self-verification and Cross-path Refinement
by: Shah, Yash, et al.
Published: (2026)
by: Shah, Yash, et al.
Published: (2026)
AI Hallucinations: A Misnomer Worth Clarifying
by: Maleki, Negar, et al.
Published: (2024)
by: Maleki, Negar, et al.
Published: (2024)
$\texttt{LM}^\texttt{2}$: A Simple Society of Language Models Solves Complex Reasoning
by: Juneja, Gurusha, et al.
Published: (2024)
by: Juneja, Gurusha, et al.
Published: (2024)
Mechanistic Behavior Editing of Language Models
by: Singh, Joykirat, et al.
Published: (2024)
by: Singh, Joykirat, et al.
Published: (2024)
Recon, Answer, Verify: Agents in Search of Truth
by: Shukla, Satyam, et al.
Published: (2025)
by: Shukla, Satyam, et al.
Published: (2025)
Revealing the impact of synthetic native samples and multi-tasking strategies in Hindi-English code-mixed humour and sarcasm detection
by: Mazumder, Debajyoti, et al.
Published: (2024)
by: Mazumder, Debajyoti, et al.
Published: (2024)
Similar Items
-
The Human and the Mechanical: logos, truthfulness, and ChatGPT
by: Giannakidou, Anastasia, et al.
Published: (2024) -
Exploring the generalization of LLM truth directions on conversational formats
by: Ichmoukhamedov, Timour, et al.
Published: (2025) -
Evaluating and explaining training strategies for zero-shot cross-lingual news sentiment analysis
by: Andrenšek, Luka, et al.
Published: (2024) -
An overview of model uncertainty and variability in LLM-based sentiment analysis. Challenges, mitigation strategies and the role of explainability
by: Herrera-Poyatos, David, et al.
Published: (2025) -
Illuminate: A novel approach for depression detection with explainable analysis and proactive therapy using prompt engineering
by: Agrawal, Aryan
Published: (2024)