Saved in:
| Main Authors: | Li, Tao, Cheng, Chin-Yi, Xie, Amber, Li, Gang, Li, Yang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2406.18559 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Lost in Edits? A $λ$-Compass for AIGC Provenance
by: You, Wenhao, et al.
Published: (2025)
by: You, Wenhao, et al.
Published: (2025)
Predicting and Explaining Mobile UI Tappability with Vision Modeling and Saliency Analysis
by: Schoop, Eldon, et al.
Published: (2022)
by: Schoop, Eldon, et al.
Published: (2022)
InstructEdit: Instruction-based Knowledge Editing for Large Language Models
by: Zhang, Ningyu, et al.
Published: (2024)
by: Zhang, Ningyu, et al.
Published: (2024)
Guided AbsoluteGrad: Magnitude of Gradients Matters to Explanation's Localization and Saliency
by: Huang, Jun, et al.
Published: (2024)
by: Huang, Jun, et al.
Published: (2024)
ExpressEdit: Fast Editing of Stylized Facial Expressions with Diffusion Models in Photoshop
by: Tang, Kenan, et al.
Published: (2026)
by: Tang, Kenan, et al.
Published: (2026)
Improving Prototypical Visual Explanations with Reward Reweighing, Reselection, and Retraining
by: Li, Aaron J., et al.
Published: (2023)
by: Li, Aaron J., et al.
Published: (2023)
EasyEdit2: An Easy-to-use Steering Framework for Editing Large Language Models
by: Xu, Ziwen, et al.
Published: (2025)
by: Xu, Ziwen, et al.
Published: (2025)
Efficient Personalization of Generative User Interfaces
by: Peng, Yi-Hao, et al.
Published: (2026)
by: Peng, Yi-Hao, et al.
Published: (2026)
A Foundational Generative Model for Breast Ultrasound Image Analysis
by: Yu, Haojun, et al.
Published: (2025)
by: Yu, Haojun, et al.
Published: (2025)
Fusing Forces: Deep-Human-Guided Refinement of Segmentation Masks
by: Sterzinger, Rafael, et al.
Published: (2024)
by: Sterzinger, Rafael, et al.
Published: (2024)
Dodgersort: Uncertainty-Aware VLM-Guided Human-in-the-Loop Pairwise Ranking
by: Park, Yujin, et al.
Published: (2026)
by: Park, Yujin, et al.
Published: (2026)
EyeFormer: Predicting Personalized Scanpaths with Transformer-Guided Reinforcement Learning
by: Jiang, Yue, et al.
Published: (2024)
by: Jiang, Yue, et al.
Published: (2024)
Beyond One-Size-Fits-All: A Survey of Personalized Affective Computing in Human-Agent Interaction
by: Li, Jialin, et al.
Published: (2023)
by: Li, Jialin, et al.
Published: (2023)
Less is More: Empowering GUI Agent with Context-Aware Simplification
by: Chen, Gongwei, et al.
Published: (2025)
by: Chen, Gongwei, et al.
Published: (2025)
OpenDriver: An Open-Road Driver State Detection Dataset
by: Liu, Delong, et al.
Published: (2023)
by: Liu, Delong, et al.
Published: (2023)
Generalization of CNNs on Relational Reasoning with Bar Charts
by: Cui, Zhenxing, et al.
Published: (2025)
by: Cui, Zhenxing, et al.
Published: (2025)
Semantic Approach to Quantifying the Consistency of Diffusion Model Image Generation
by: Bent, Brinnae
Published: (2024)
by: Bent, Brinnae
Published: (2024)
What's Producible May Not Be Reachable: Measuring the Steerability of Generative Models
by: Vafa, Keyon, et al.
Published: (2025)
by: Vafa, Keyon, et al.
Published: (2025)
AI Guide Dog: Egocentric Path Prediction on Smartphone
by: Jadhav, Aishwarya, et al.
Published: (2025)
by: Jadhav, Aishwarya, et al.
Published: (2025)
GazeTrack: High-Precision Eye Tracking Based on Regularization and Spatial Computing
by: Yang, Xiaoyin
Published: (2025)
by: Yang, Xiaoyin
Published: (2025)
A Survey on Trustworthiness in Foundation Models for Medical Image Analysis
by: Shi, Congzhen, et al.
Published: (2024)
by: Shi, Congzhen, et al.
Published: (2024)
Deep Generative Domain Adaptation with Temporal Attention for Cross-User Activity Recognition
by: Ye, Xiaozhou, et al.
Published: (2024)
by: Ye, Xiaozhou, et al.
Published: (2024)
Generating Synthetic Satellite Imagery for Rare Objects: An Empirical Comparison of Models and Metrics
by: Nguyen, Tuong Vy, et al.
Published: (2024)
by: Nguyen, Tuong Vy, et al.
Published: (2024)
EmoGene: Audio-Driven Emotional 3D Talking-Head Generation
by: Wang, Wenqing, et al.
Published: (2024)
by: Wang, Wenqing, et al.
Published: (2024)
Screen2AX: Vision-Based Approach for Automatic macOS Accessibility Generation
by: Muryn, Viktor, et al.
Published: (2025)
by: Muryn, Viktor, et al.
Published: (2025)
Deep Generative Domain Adaptation with Temporal Relation Knowledge for Cross-User Activity Recognition
by: Ye, Xiaozhou, et al.
Published: (2024)
by: Ye, Xiaozhou, et al.
Published: (2024)
FERGI: Automatic Scoring of User Preferences for Text-to-Image Generation from Spontaneous Facial Expression Reaction
by: Feng, Shuangquan, et al.
Published: (2023)
by: Feng, Shuangquan, et al.
Published: (2023)
Generating Synthetic Satellite Imagery With Deep-Learning Text-to-Image Models -- Technical Challenges and Implications for Monitoring and Verification
by: Nguyen, Tuong Vy, et al.
Published: (2024)
by: Nguyen, Tuong Vy, et al.
Published: (2024)
Efficient Retail Video Annotation: A Robust Key Frame Generation Approach for Product and Customer Interaction Analysis
by: Mannam, Varun, et al.
Published: (2025)
by: Mannam, Varun, et al.
Published: (2025)
Instruction-Guided Editing Controls for Images and Multimedia: A Survey in LLM era
by: Nguyen, Thanh Tam, et al.
Published: (2024)
by: Nguyen, Thanh Tam, et al.
Published: (2024)
MNIST-Gen: A Modular MNIST-Style Dataset Generation Using Hierarchical Semantics, Reinforcement Learning, and Category Theory
by: Shaeri, Pouya, et al.
Published: (2025)
by: Shaeri, Pouya, et al.
Published: (2025)
DebiasPI: Inference-time Debiasing by Prompt Iteration of a Text-to-Image Generative Model
by: Bonna, Sarah, et al.
Published: (2025)
by: Bonna, Sarah, et al.
Published: (2025)
DOTA: Distributional Test-Time Adaptation of Vision-Language Models
by: Han, Zongbo, et al.
Published: (2024)
by: Han, Zongbo, et al.
Published: (2024)
Gesture Matters: Pedestrian Gesture Recognition for AVs Through Skeleton Pose Evaluation
by: Mahdi, Alif Rizqullah, et al.
Published: (2026)
by: Mahdi, Alif Rizqullah, et al.
Published: (2026)
ChainReaction: Causal Chain-Guided Reasoning for Modular and Explainable Causal-Why Video Question Answering
by: Parmar, Paritosh, et al.
Published: (2025)
by: Parmar, Paritosh, et al.
Published: (2025)
Magic Insert: Style-Aware Drag-and-Drop
by: Ruiz, Nataniel, et al.
Published: (2024)
by: Ruiz, Nataniel, et al.
Published: (2024)
Human-like object concept representations emerge naturally in multimodal large language models
by: Du, Changde, et al.
Published: (2024)
by: Du, Changde, et al.
Published: (2024)
Human-Agent Joint Learning for Efficient Robot Manipulation Skill Acquisition
by: Luo, Shengcheng, et al.
Published: (2024)
by: Luo, Shengcheng, et al.
Published: (2024)
Explainable AI for Safe and Trustworthy Autonomous Driving: A Systematic Review
by: Kuznietsov, Anton, et al.
Published: (2024)
by: Kuznietsov, Anton, et al.
Published: (2024)
Analysis of the 2024 BraTS Meningioma Radiotherapy Planning Automated Segmentation Challenge
by: LaBella, Dominic, et al.
Published: (2024)
by: LaBella, Dominic, et al.
Published: (2024)
Similar Items
-
Lost in Edits? A $λ$-Compass for AIGC Provenance
by: You, Wenhao, et al.
Published: (2025) -
Predicting and Explaining Mobile UI Tappability with Vision Modeling and Saliency Analysis
by: Schoop, Eldon, et al.
Published: (2022) -
InstructEdit: Instruction-based Knowledge Editing for Large Language Models
by: Zhang, Ningyu, et al.
Published: (2024) -
Guided AbsoluteGrad: Magnitude of Gradients Matters to Explanation's Localization and Saliency
by: Huang, Jun, et al.
Published: (2024) -
ExpressEdit: Fast Editing of Stylized Facial Expressions with Diffusion Models in Photoshop
by: Tang, Kenan, et al.
Published: (2026)