Gespeichert in:
| Hauptverfasser: | Park, Ji-Hoon, Ju, Yeong-Joon, Lee, Seong-Whan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2402.10404 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MIRe: Enhancing Multimodal Queries Representation via Fusion-Free Modality Interaction for Multimodal Retrieval
von: Ju, Yeong-Joon, et al.
Veröffentlicht: (2024)
von: Ju, Yeong-Joon, et al.
Veröffentlicht: (2024)
DarSwin: Distortion Aware Radial Swin Transformer
von: Athwale, Akshaya, et al.
Veröffentlicht: (2023)
von: Athwale, Akshaya, et al.
Veröffentlicht: (2023)
Semantic Depth Matters: Explaining Errors of Deep Vision Networks through Perceived Class Similarities
von: Filus, Katarzyna, et al.
Veröffentlicht: (2025)
von: Filus, Katarzyna, et al.
Veröffentlicht: (2025)
Identity-free Artificial Emotional Intelligence via Micro-Gesture Understanding
von: Gao, Rong, et al.
Veröffentlicht: (2024)
von: Gao, Rong, et al.
Veröffentlicht: (2024)
Deepfakes on Demand: the rise of accessible non-consensual deepfake image generators
von: Hawkins, Will, et al.
Veröffentlicht: (2025)
von: Hawkins, Will, et al.
Veröffentlicht: (2025)
Matrix-Valued LogSumExp Approximation for Colour Morphology
von: Kahra, Marvin, et al.
Veröffentlicht: (2024)
von: Kahra, Marvin, et al.
Veröffentlicht: (2024)
Colour Morphological Distance Ordering based on the Log-Exp-Supremum
von: Kahra, Marvin, et al.
Veröffentlicht: (2025)
von: Kahra, Marvin, et al.
Veröffentlicht: (2025)
Sequence Transferability and Task Order Selection in Continual Learning
von: Nguyen, Thinh, et al.
Veröffentlicht: (2025)
von: Nguyen, Thinh, et al.
Veröffentlicht: (2025)
From Volume Rendering to 3D Gaussian Splatting: Theory and Applications
von: Matias, Vitor Pereira, et al.
Veröffentlicht: (2025)
von: Matias, Vitor Pereira, et al.
Veröffentlicht: (2025)
What is the Visual Cognition Gap between Humans and Multimodal LLMs?
von: Cao, Xu, et al.
Veröffentlicht: (2024)
von: Cao, Xu, et al.
Veröffentlicht: (2024)
Animation Needs Attention: A Holistic Approach to Slides Animation Comprehension with Visual-Language Models
von: Jiang, Yifan, et al.
Veröffentlicht: (2025)
von: Jiang, Yifan, et al.
Veröffentlicht: (2025)
DQE-CIR: Distinctive Query Embeddings through Learnable Attribute Weights and Target Relative Negative Sampling in Composed Image Retrieval
von: Park, Geon, et al.
Veröffentlicht: (2026)
von: Park, Geon, et al.
Veröffentlicht: (2026)
MedShapeNet -- A Large-Scale Dataset of 3D Medical Shapes for Computer Vision
von: Li, Jianning, et al.
Veröffentlicht: (2023)
von: Li, Jianning, et al.
Veröffentlicht: (2023)
A Hierarchically Feature Reconstructed Autoencoder for Unsupervised Anomaly Detection
von: Chen, Honghui, et al.
Veröffentlicht: (2024)
von: Chen, Honghui, et al.
Veröffentlicht: (2024)
Semantic Prioritization in Visual Counterfactual Explanations with Weighted Segmentation and Auto-Adaptive Region Selection
von: Zhang, Lintong, et al.
Veröffentlicht: (2025)
von: Zhang, Lintong, et al.
Veröffentlicht: (2025)
Universal Adversarial Perturbations for Vision-Language Pre-trained Models
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2024)
von: Zhang, Peng-Fei, et al.
Veröffentlicht: (2024)
AIDOVECL: AI-generated Dataset of Outpainted Vehicles for Eye-level Classification and Localization
von: Kazemi, Amir, et al.
Veröffentlicht: (2024)
von: Kazemi, Amir, et al.
Veröffentlicht: (2024)
Accelerating Large Kernel Convolutions with Nested Winograd Transformation.pdf
von: Jiang, Jingbo, et al.
Veröffentlicht: (2021)
von: Jiang, Jingbo, et al.
Veröffentlicht: (2021)
Empowering Manufacturers with Privacy-Preserving AI Tools: A Case Study in Privacy-Preserving Machine Learning to Solve Real-World Problems
von: Ji, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Ji, Xiaoyu, et al.
Veröffentlicht: (2025)
Pedestrian intention prediction in Adverse Weather Conditions with Spiking Neural Networks and Dynamic Vision Sensors
von: Sakhai, Mustafa, et al.
Veröffentlicht: (2024)
von: Sakhai, Mustafa, et al.
Veröffentlicht: (2024)
VACoDe: Visual Augmented Contrastive Decoding
von: Kim, Sihyeon, et al.
Veröffentlicht: (2024)
von: Kim, Sihyeon, et al.
Veröffentlicht: (2024)
HAAP: Vision-context Hierarchical Attention Autoregressive with Adaptive Permutation for Scene Text Recognition
von: Chen, Honghui, et al.
Veröffentlicht: (2024)
von: Chen, Honghui, et al.
Veröffentlicht: (2024)
Rethinking Evaluation of Multiple Sclerosis (MS) Lesion Segmentation Models
von: Basit, Abdul, et al.
Veröffentlicht: (2026)
von: Basit, Abdul, et al.
Veröffentlicht: (2026)
A lifted Bregman strategy for training unfolded proximal neural network Gaussian denoisers
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2024)
Socratic Planner: Self-QA-Based Zero-Shot Planning for Embodied Instruction Following
von: Shin, Suyeon, et al.
Veröffentlicht: (2024)
von: Shin, Suyeon, et al.
Veröffentlicht: (2024)
Upper-body free-breathing Magnetic Resonance Fingerprinting applied to the quantification of water T1 and fat fraction
von: Slioussarenko, Constantin, et al.
Veröffentlicht: (2024)
von: Slioussarenko, Constantin, et al.
Veröffentlicht: (2024)
African Gender Classification Using Clothing Identification Via Deep Learning
von: Ozechi, Samuel
Veröffentlicht: (2025)
von: Ozechi, Samuel
Veröffentlicht: (2025)
On Discrete Prompt Optimization for Diffusion Models
von: Wang, Ruochen, et al.
Veröffentlicht: (2024)
von: Wang, Ruochen, et al.
Veröffentlicht: (2024)
Representative Feature Extraction During Diffusion Process for Sketch Extraction with One Example
von: Yun, Kwan, et al.
Veröffentlicht: (2024)
von: Yun, Kwan, et al.
Veröffentlicht: (2024)
Semantic2Graph: Graph-based Multi-modal Feature Fusion for Action Segmentation in Videos
von: Zhang, Junbin, et al.
Veröffentlicht: (2022)
von: Zhang, Junbin, et al.
Veröffentlicht: (2022)
Leveraging CNN and IoT for Effective E-Waste Management
von: Nadar, Ajesh Thangaraj, et al.
Veröffentlicht: (2025)
von: Nadar, Ajesh Thangaraj, et al.
Veröffentlicht: (2025)
SASWISE-UE: Segmentation and Synthesis with Interpretable Scalable Ensembles for Uncertainty Estimation
von: Chen, Weijie, et al.
Veröffentlicht: (2024)
von: Chen, Weijie, et al.
Veröffentlicht: (2024)
Reconstructing Gridded Data from Higher Autocorrelations
von: Casper, W. Riley, et al.
Veröffentlicht: (2025)
von: Casper, W. Riley, et al.
Veröffentlicht: (2025)
FST.ai 2.0: An Explainable AI Ecosystem for Fair, Fast, and Inclusive Decision-Making in Olympic and Paralympic Taekwondo
von: Shariatmadar, Keivan, et al.
Veröffentlicht: (2025)
von: Shariatmadar, Keivan, et al.
Veröffentlicht: (2025)
Vision transformer-based multi-camera multi-object tracking framework for dairy cow monitoring
von: Abbas, Kumail, et al.
Veröffentlicht: (2025)
von: Abbas, Kumail, et al.
Veröffentlicht: (2025)
MCoT-RE: Multi-Faceted Chain-of-Thought and Re-Ranking for Training-Free Zero-Shot Composed Image Retrieval
von: Park, Jeong-Woo, et al.
Veröffentlicht: (2025)
von: Park, Jeong-Woo, et al.
Veröffentlicht: (2025)
Evaluation of (Un-)Supervised Machine Learning Methods for GNSS Interference Classification with Real-World Data Discrepancies
von: Heublein, Lucas, et al.
Veröffentlicht: (2025)
von: Heublein, Lucas, et al.
Veröffentlicht: (2025)
Representation Selection via Cross-Model Agreement using Canonical Correlation Analysis
von: Lewis, Dylan B., et al.
Veröffentlicht: (2026)
von: Lewis, Dylan B., et al.
Veröffentlicht: (2026)
Surrealistic-like Image Generation with Vision-Language Models
von: Ayten, Elif, et al.
Veröffentlicht: (2024)
von: Ayten, Elif, et al.
Veröffentlicht: (2024)
TIFu: Tri-directional Implicit Function for High-Fidelity 3D Character Reconstruction
von: Lim, Byoungsung, et al.
Veröffentlicht: (2024)
von: Lim, Byoungsung, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MIRe: Enhancing Multimodal Queries Representation via Fusion-Free Modality Interaction for Multimodal Retrieval
von: Ju, Yeong-Joon, et al.
Veröffentlicht: (2024) -
DarSwin: Distortion Aware Radial Swin Transformer
von: Athwale, Akshaya, et al.
Veröffentlicht: (2023) -
Semantic Depth Matters: Explaining Errors of Deep Vision Networks through Perceived Class Similarities
von: Filus, Katarzyna, et al.
Veröffentlicht: (2025) -
Identity-free Artificial Emotional Intelligence via Micro-Gesture Understanding
von: Gao, Rong, et al.
Veröffentlicht: (2024) -
Deepfakes on Demand: the rise of accessible non-consensual deepfake image generators
von: Hawkins, Will, et al.
Veröffentlicht: (2025)