What Happens to a Dataset Transformed by a Projection-based Concept Removal Method?
Fuente:
arXiv
Saved in:
| Main Author: | Johansson, Richard |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Where does In-context Translation Happen in Large Language Models
by: Sia, Suzanna, et al.
Published: (2024)
by: Sia, Suzanna, et al.
Published: (2024)
What Happened in LLMs Layers when Trained for Fast vs. Slow Thinking: A Gradient Perspective
by: Li, Ming, et al.
Published: (2024)
by: Li, Ming, et al.
Published: (2024)
When and Where Did it Happen? An Encoder-Decoder Model to Identify Scenario Context
by: Noriega-Atala, Enrique, et al.
Published: (2024)
by: Noriega-Atala, Enrique, et al.
Published: (2024)
Latent Concept Disentanglement in Transformer-based Language Models
by: Hong, Guan Zhe, et al.
Published: (2025)
by: Hong, Guan Zhe, et al.
Published: (2025)
Multimodal Sentiment Analysis on CMU-MOSEI Dataset using Transformer-based Models
by: Gajjar, Jugal, et al.
Published: (2025)
by: Gajjar, Jugal, et al.
Published: (2025)
FADE: Why Bad Descriptions Happen to Good Features
by: Puri, Bruno, et al.
Published: (2025)
by: Puri, Bruno, et al.
Published: (2025)
Generating Concept Lexicalizations via Dictionary-Based Cross-Lingual Sense Projection
by: Basil, David, et al.
Published: (2026)
by: Basil, David, et al.
Published: (2026)
Uncovering Implicit Bias in Large Language Models with Concept Learning Dataset
by: Wang, Leroy Z.
Published: (2025)
by: Wang, Leroy Z.
Published: (2025)
A Novel Prompt-tuning Method: Incorporating Scenario-specific Concepts into a Verbalizer
by: Ma, Yong, et al.
Published: (2024)
by: Ma, Yong, et al.
Published: (2024)
Meta-Judging with Large Language Models: Concepts, Methods, and Challenges
by: Silva, Hugo, et al.
Published: (2026)
by: Silva, Hugo, et al.
Published: (2026)
Removing RLHF Protections in GPT-4 via Fine-Tuning
by: Zhan, Qiusi, et al.
Published: (2023)
by: Zhan, Qiusi, et al.
Published: (2023)
AutoAugment Is What You Need: Enhancing Rule-based Augmentation Methods in Low-resource Regimes
by: Choi, Juhwan, et al.
Published: (2024)
by: Choi, Juhwan, et al.
Published: (2024)
From Concepts to Components: Concept-Agnostic Attention Module Discovery in Transformers
by: Su, Jingtong, et al.
Published: (2025)
by: Su, Jingtong, et al.
Published: (2025)
Chinese Cyberbullying Detection: Dataset, Method, and Validation
by: Zhu, Yi, et al.
Published: (2025)
by: Zhu, Yi, et al.
Published: (2025)
Is It a Free Lunch for Removing Outliers during Pretraining?
by: Liao, Baohao, et al.
Published: (2024)
by: Liao, Baohao, et al.
Published: (2024)
Estimating Text Similarity based on Semantic Concept Embeddings
by: der Brück, Tim vor, et al.
Published: (2024)
by: der Brück, Tim vor, et al.
Published: (2024)
Projected Compression: Trainable Projection for Efficient Transformer Compression
by: Stefaniak, Maciej, et al.
Published: (2025)
by: Stefaniak, Maciej, et al.
Published: (2025)
Safe-CLIP: Removing NSFW Concepts from Vision-and-Language Models
by: Poppi, Samuele, et al.
Published: (2023)
by: Poppi, Samuele, et al.
Published: (2023)
Separating Tongue from Thought: Activation Patching Reveals Language-Agnostic Concept Representations in Transformers
by: Dumas, Clément, et al.
Published: (2024)
by: Dumas, Clément, et al.
Published: (2024)
What do language models model? Transformers, automata, and the format of thought
by: Klein, Colin
Published: (2025)
by: Klein, Colin
Published: (2025)
On the Military Applications of Large Language Models
by: Johansson, Satu, et al.
Published: (2025)
by: Johansson, Satu, et al.
Published: (2025)
Machine Learning-based NLP for Emotion Classification on a Cholera X Dataset
by: Jideani, Paul, et al.
Published: (2024)
by: Jideani, Paul, et al.
Published: (2024)
Interdisciplinary Fairness in Imbalanced Research Proposal Topic Inference: A Hierarchical Transformer-based Method with Selective Interpolation
by: Xiao, Meng, et al.
Published: (2023)
by: Xiao, Meng, et al.
Published: (2023)
Emotion Concepts and their Function in a Large Language Model
by: Sofroniew, Nicholas, et al.
Published: (2026)
by: Sofroniew, Nicholas, et al.
Published: (2026)
Language Model Re-rankers are Fooled by Lexical Similarities
by: Hagström, Lovisa, et al.
Published: (2025)
by: Hagström, Lovisa, et al.
Published: (2025)
A Method for Constructing a Digital Transformation Driving Mechanism Based on Semantic Understanding of Large Models
by: Liu, Huayi
Published: (2026)
by: Liu, Huayi
Published: (2026)
Automated Concept Discovery for LLM-as-a-Judge Preference Analysis
by: Wedgwood, James, et al.
Published: (2026)
by: Wedgwood, James, et al.
Published: (2026)
What are the Essential Factors in Crafting Effective Long Context Multi-Hop Instruction Datasets? Insights and Best Practices
by: Chen, Zhi, et al.
Published: (2024)
by: Chen, Zhi, et al.
Published: (2024)
Removal of Hallucination on Hallucination: Debate-Augmented RAG
by: Hu, Wentao, et al.
Published: (2025)
by: Hu, Wentao, et al.
Published: (2025)
BERT-based model for Vietnamese Fact Verification Dataset
by: Tran, Bao, et al.
Published: (2025)
by: Tran, Bao, et al.
Published: (2025)
Knowing What LLMs DO NOT Know: A Simple Yet Effective Self-Detection Method
by: Zhao, Yukun, et al.
Published: (2023)
by: Zhao, Yukun, et al.
Published: (2023)
What Matters in Transformers? Not All Attention is Needed
by: He, Shwai, et al.
Published: (2024)
by: He, Shwai, et al.
Published: (2024)
CUICurate: A GraphRAG-based Framework for Automated Clinical Concept Curation for NLP applications
by: Blake, Victoria, et al.
Published: (2026)
by: Blake, Victoria, et al.
Published: (2026)
ConceptViz: A Visual Analytics Approach for Exploring Concepts in Large Language Models
by: Li, Haoxuan, et al.
Published: (2025)
by: Li, Haoxuan, et al.
Published: (2025)
A Large-Scale Dataset for Molecular Structure-Language Description via a Rule-Regularized Method
by: Cai, Feiyang, et al.
Published: (2026)
by: Cai, Feiyang, et al.
Published: (2026)
CF-RAG: A Dataset and Method for Carbon Footprint QA Using Retrieval-Augmented Generation
by: Zhao, Kaiwen, et al.
Published: (2025)
by: Zhao, Kaiwen, et al.
Published: (2025)
UNLEARN Efficient Removal of Knowledge in Large Language Models
by: Lizzo, Tyler, et al.
Published: (2024)
by: Lizzo, Tyler, et al.
Published: (2024)
Pre-training a Transformer-Based Generative Model Using a Small Sepedi Dataset
by: Ramalepe, Simon P., et al.
Published: (2025)
by: Ramalepe, Simon P., et al.
Published: (2025)
DREsS: Dataset for Rubric-based Essay Scoring on EFL Writing
by: Yoo, Haneul, et al.
Published: (2024)
by: Yoo, Haneul, et al.
Published: (2024)
Concept Attractors in LLMs and their Applications
by: Chytas, Sotirios Panagiotis, et al.
Published: (2025)
by: Chytas, Sotirios Panagiotis, et al.
Published: (2025)
Similar Items
-
Where does In-context Translation Happen in Large Language Models
by: Sia, Suzanna, et al.
Published: (2024) -
What Happened in LLMs Layers when Trained for Fast vs. Slow Thinking: A Gradient Perspective
by: Li, Ming, et al.
Published: (2024) -
When and Where Did it Happen? An Encoder-Decoder Model to Identify Scenario Context
by: Noriega-Atala, Enrique, et al.
Published: (2024) -
Latent Concept Disentanglement in Transformer-based Language Models
by: Hong, Guan Zhe, et al.
Published: (2025) -
Multimodal Sentiment Analysis on CMU-MOSEI Dataset using Transformer-based Models
by: Gajjar, Jugal, et al.
Published: (2025)