MedSteer: Counterfactual Endoscopic Synthesis via Training-Free Activation Steering
Fuente:
arXiv
Salvato in:
| Autori principali: | Pham, Trong-Thang, Nguyen, Loc, Nguyen, Anh, Nguyen, Hien, Le, Ngan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GazeQwen: Lightweight Gaze-Conditioned LLM Modulation for Streaming Video Understanding
di: Pham, Trong Thang, et al.
Pubblicazione: (2026)
di: Pham, Trong Thang, et al.
Pubblicazione: (2026)
Interpreting Radiologist's Intention from Eye Movements in Chest X-ray Diagnosis
di: Pham, Trong-Thang, et al.
Pubblicazione: (2025)
di: Pham, Trong-Thang, et al.
Pubblicazione: (2025)
GazeSearch: Radiology Findings Search Benchmark
di: Pham, Trong Thang, et al.
Pubblicazione: (2024)
di: Pham, Trong Thang, et al.
Pubblicazione: (2024)
DuFal: Dual-Frequency-Aware Learning for High-Fidelity Extremely Sparse-view CBCT Reconstruction
di: Van, Cuong Tran, et al.
Pubblicazione: (2026)
di: Van, Cuong Tran, et al.
Pubblicazione: (2026)
InverFill: One-Step Inversion for Enhanced Few-Step Diffusion Inpainting
di: Vu, Duc, et al.
Pubblicazione: (2026)
di: Vu, Duc, et al.
Pubblicazione: (2026)
FG-CXR: A Radiologist-Aligned Gaze Dataset for Enhancing Interpretability in Chest X-Ray Report Generation
di: Pham, Trong Thang, et al.
Pubblicazione: (2024)
di: Pham, Trong Thang, et al.
Pubblicazione: (2024)
WAVER: Writing-style Agnostic Text-Video Retrieval via Distilling Vision-Language Models Through Open-Vocabulary Knowledge
di: Le, Huy, et al.
Pubblicazione: (2023)
di: Le, Huy, et al.
Pubblicazione: (2023)
HDC: Hierarchical Distillation for Multi-level Noisy Consistency in Semi-Supervised Fetal Ultrasound Segmentation
di: Le, Tran Quoc Khanh, et al.
Pubblicazione: (2025)
di: Le, Tran Quoc Khanh, et al.
Pubblicazione: (2025)
BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance
di: Le, Huy, et al.
Pubblicazione: (2025)
di: Le, Huy, et al.
Pubblicazione: (2025)
FurniMAS: Language-Guided Furniture Decoration using Multi-Agent System
di: Nguyen, Toan, et al.
Pubblicazione: (2025)
di: Nguyen, Toan, et al.
Pubblicazione: (2025)
Bridging the Training-Deployment Gap: Gated Encoding and Multi-Scale Refinement for Efficient Quantization-Aware Image Enhancement
di: To-Thanh, Dat, et al.
Pubblicazione: (2026)
di: To-Thanh, Dat, et al.
Pubblicazione: (2026)
PEEB: Part-based Image Classifiers with an Explainable and Editable Language Bottleneck
di: Pham, Thang M., et al.
Pubblicazione: (2024)
di: Pham, Thang M., et al.
Pubblicazione: (2024)
Causally Steered Diffusion for Automated Video Counterfactual Generation
di: Spyrou, Nikos, et al.
Pubblicazione: (2025)
di: Spyrou, Nikos, et al.
Pubblicazione: (2025)
Steering to Say No: Configurable Refusal via Activation Steering in Vision Language Models
di: Yang, Jiaxi, et al.
Pubblicazione: (2026)
di: Yang, Jiaxi, et al.
Pubblicazione: (2026)
Aleatoric Uncertainty Medical Image Segmentation Estimation via Flow Matching
di: Van Nguyen, Phi, et al.
Pubblicazione: (2025)
di: Van Nguyen, Phi, et al.
Pubblicazione: (2025)
GrAInS: Gradient-based Attribution for Inference-Time Steering of LLMs and VLMs
di: Nguyen, Duy, et al.
Pubblicazione: (2025)
di: Nguyen, Duy, et al.
Pubblicazione: (2025)
SwiftBrush v2: Make Your One-step Diffusion Model Better Than Its Teacher
di: Dao, Trung, et al.
Pubblicazione: (2024)
di: Dao, Trung, et al.
Pubblicazione: (2024)
Virtual Fusion with Contrastive Learning for Single Sensor-based Activity Recognition
di: Nguyen, Duc-Anh, et al.
Pubblicazione: (2023)
di: Nguyen, Duc-Anh, et al.
Pubblicazione: (2023)
OE3DIS: Open-Ended 3D Point Cloud Instance Segmentation
di: Nguyen, Phuc D. A., et al.
Pubblicazione: (2024)
di: Nguyen, Phuc D. A., et al.
Pubblicazione: (2024)
Conditioned Activation Transport for T2I Safety Steering
di: Chrabąszcz, Maciej, et al.
Pubblicazione: (2026)
di: Chrabąszcz, Maciej, et al.
Pubblicazione: (2026)
SwiftEdit: Lightning Fast Text-Guided Image Editing via One-Step Diffusion
di: Nguyen, Trong-Tung, et al.
Pubblicazione: (2024)
di: Nguyen, Trong-Tung, et al.
Pubblicazione: (2024)
Enhancing Multimodal Entity Linking with Jaccard Distance-based Conditional Contrastive Learning and Contextual Visual Augmentation
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2025)
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2025)
CSD-VAR: Content-Style Decomposition in Visual Autoregressive Models
di: Nguyen, Quang-Binh, et al.
Pubblicazione: (2025)
di: Nguyen, Quang-Binh, et al.
Pubblicazione: (2025)
PAT: Pixel-wise Adaptive Training for Long-tailed Segmentation
di: Do, Khoi, et al.
Pubblicazione: (2024)
di: Do, Khoi, et al.
Pubblicazione: (2024)
FlexEdit: Flexible and Controllable Diffusion-based Object-centric Image Editing
di: Nguyen, Trong-Tung, et al.
Pubblicazione: (2024)
di: Nguyen, Trong-Tung, et al.
Pubblicazione: (2024)
XEdgeAI: A Human-centered Industrial Inspection Framework with Data-centric Explainable Edge AI Approach
di: Nguyen, Truong Thanh Hung, et al.
Pubblicazione: (2024)
di: Nguyen, Truong Thanh Hung, et al.
Pubblicazione: (2024)
SimGraph: A Unified Framework for Scene Graph-Based Image Generation and Editing
di: Vo, Thanh-Nhan, et al.
Pubblicazione: (2026)
di: Vo, Thanh-Nhan, et al.
Pubblicazione: (2026)
A2VIS: Amodal-Aware Approach to Video Instance Segmentation
di: Tran, Minh, et al.
Pubblicazione: (2024)
di: Tran, Minh, et al.
Pubblicazione: (2024)
Multimodal Contextualized Support for Enhancing Video Retrieval System
di: Nguyen-Le, Quoc-Bao, et al.
Pubblicazione: (2024)
di: Nguyen-Le, Quoc-Bao, et al.
Pubblicazione: (2024)
Enhancing Radiological Diagnosis: A Collaborative Approach Integrating AI and Human Expertise for Visual Miss Correction
di: Awasthi, Akash, et al.
Pubblicazione: (2024)
di: Awasthi, Akash, et al.
Pubblicazione: (2024)
BALM: A Model-Agnostic Framework for Balanced Multimodal Learning under Imbalanced Missing Rates
di: Nguyen, Phuong-Anh, et al.
Pubblicazione: (2026)
di: Nguyen, Phuong-Anh, et al.
Pubblicazione: (2026)
MissBench: Benchmarking Multimodal Affective Analysis under Imbalanced Missing Modalities
di: Pham, Tien Anh, et al.
Pubblicazione: (2026)
di: Pham, Tien Anh, et al.
Pubblicazione: (2026)
ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transport
di: Tran, Quoc-Khang, et al.
Pubblicazione: (2026)
di: Tran, Quoc-Khang, et al.
Pubblicazione: (2026)
Semi-Supervised Semantic Segmentation using Redesigned Self-Training for White Blood Cells
di: Luu, Vinh Quoc, et al.
Pubblicazione: (2024)
di: Luu, Vinh Quoc, et al.
Pubblicazione: (2024)
MADTempo: An Interactive System for Multi-Event Temporal Video Retrieval with Query Augmentation
di: Vu, Huu-An, et al.
Pubblicazione: (2025)
di: Vu, Huu-An, et al.
Pubblicazione: (2025)
A Two-Stage, Object-Centric Deep Learning Framework for Robust Exam Cheating Detection
di: Le, Van-Truong, et al.
Pubblicazione: (2026)
di: Le, Van-Truong, et al.
Pubblicazione: (2026)
Dynamic Multimodal Activation Steering for Hallucination Mitigation in Large Vision-Language Models
di: Yin, Jianghao, et al.
Pubblicazione: (2026)
di: Yin, Jianghao, et al.
Pubblicazione: (2026)
Federated Prompt-Tuning with Heterogeneous and Incomplete Multimodal Client Data
di: Phung, Thu Hang, et al.
Pubblicazione: (2026)
di: Phung, Thu Hang, et al.
Pubblicazione: (2026)
MedMax: Mixed-Modal Instruction Tuning for Training Biomedical Assistants
di: Bansal, Hritik, et al.
Pubblicazione: (2024)
di: Bansal, Hritik, et al.
Pubblicazione: (2024)
Improving Zero-Shot Object-Level Change Detection by Incorporating Visual Correspondence
di: Nguyen, Hung Huy, et al.
Pubblicazione: (2025)
di: Nguyen, Hung Huy, et al.
Pubblicazione: (2025)
Documenti analoghi
-
GazeQwen: Lightweight Gaze-Conditioned LLM Modulation for Streaming Video Understanding
di: Pham, Trong Thang, et al.
Pubblicazione: (2026) -
Interpreting Radiologist's Intention from Eye Movements in Chest X-ray Diagnosis
di: Pham, Trong-Thang, et al.
Pubblicazione: (2025) -
GazeSearch: Radiology Findings Search Benchmark
di: Pham, Trong Thang, et al.
Pubblicazione: (2024) -
DuFal: Dual-Frequency-Aware Learning for High-Fidelity Extremely Sparse-view CBCT Reconstruction
di: Van, Cuong Tran, et al.
Pubblicazione: (2026) -
InverFill: One-Step Inversion for Enhanced Few-Step Diffusion Inpainting
di: Vu, Duc, et al.
Pubblicazione: (2026)