Guardado en:
| Autores principales: | Storonkin, Daniil, Dziub, Ilia, Golyadkin, Maksim, Makarov, Ilya |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.07062 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CerberusDet: Unified Multi-Dataset Object Detection
por: Tolstykh, Irina, et al.
Publicado: (2024)
por: Tolstykh, Irina, et al.
Publicado: (2024)
Beyond Specialization: Assessing the Capabilities of MLLMs in Age and Gender Estimation
por: Kuprashevich, Maksim, et al.
Publicado: (2024)
por: Kuprashevich, Maksim, et al.
Publicado: (2024)
A Review of Pseudo-Labeling for Computer Vision
por: Kage, Patrick, et al.
Publicado: (2024)
por: Kage, Patrick, et al.
Publicado: (2024)
Synthetic Photography Detection: A Visual Guidance for Identifying Synthetic Images Created by AI
por: Mathys, Melanie, et al.
Publicado: (2024)
por: Mathys, Melanie, et al.
Publicado: (2024)
Designing UNICORN: a Unified Benchmark for Imaging in Computational Pathology, Radiology, and Natural Language
por: Stegeman, Michelle, et al.
Publicado: (2026)
por: Stegeman, Michelle, et al.
Publicado: (2026)
Synthetic Image Generation in Cyber Influence Operations: An Emergent Threat?
por: Mathys, Melanie, et al.
Publicado: (2024)
por: Mathys, Melanie, et al.
Publicado: (2024)
Stereo Vision Based Robot for Remote Monitoring with VR Support
por: S., Mohamed Fazil M., et al.
Publicado: (2024)
por: S., Mohamed Fazil M., et al.
Publicado: (2024)
COLORA: Efficient Fine-Tuning for Convolutional Models with a Study Case on Optical Coherence Tomography Image Classification
por: Rivera, Mariano, et al.
Publicado: (2025)
por: Rivera, Mariano, et al.
Publicado: (2025)
High-Entropy Tokens as Multimodal Failure Points in Vision-Language Models
por: He, Mengqi, et al.
Publicado: (2025)
por: He, Mengqi, et al.
Publicado: (2025)
Vision Transformer-based Model for Severity Quantification of Lung Pneumonia Using Chest X-ray Images
por: Slika, Bouthaina, et al.
Publicado: (2023)
por: Slika, Bouthaina, et al.
Publicado: (2023)
A Survey on Dynamic Neural Networks: from Computer Vision to Multi-modal Sensor Fusion
por: Montello, Fabio, et al.
Publicado: (2025)
por: Montello, Fabio, et al.
Publicado: (2025)
Unified Local and Global Attention Interaction Modeling for Vision Transformers
por: Nguyen, Tan, et al.
Publicado: (2024)
por: Nguyen, Tan, et al.
Publicado: (2024)
Efficient Diffusion Training through Parallelization with Truncated Karhunen-Loève Expansion
por: Ren, Yumeng, et al.
Publicado: (2025)
por: Ren, Yumeng, et al.
Publicado: (2025)
Proto-FG3D: Prototype-based Interpretable Fine-Grained 3D Shape Classification
por: Ma, Shuxian, et al.
Publicado: (2025)
por: Ma, Shuxian, et al.
Publicado: (2025)
Context-Aware Full Body Anonymization using Text-to-Image Diffusion Models
por: Zwick, Pascal, et al.
Publicado: (2024)
por: Zwick, Pascal, et al.
Publicado: (2024)
Foreground Focus: Enhancing Coherence and Fidelity in Camouflaged Image Generation
por: Chen, Pei-Chi, et al.
Publicado: (2025)
por: Chen, Pei-Chi, et al.
Publicado: (2025)
Event-based Solutions for Human-centered Applications: A Comprehensive Review
por: Adra, Mira, et al.
Publicado: (2025)
por: Adra, Mira, et al.
Publicado: (2025)
Collaborative Control for Geometry-Conditioned PBR Image Generation
por: Vainer, Shimon, et al.
Publicado: (2024)
por: Vainer, Shimon, et al.
Publicado: (2024)
BlanketGen2-Fit3D: Synthetic Blanket Augmentation Towards Improving Real-World In-Bed Blanket Occluded Human Pose Estimation
por: Karácsony, Tamás, et al.
Publicado: (2025)
por: Karácsony, Tamás, et al.
Publicado: (2025)
Doodle Your Keypoints: Sketch-Based Few-Shot Keypoint Detection
por: Maity, Subhajit, et al.
Publicado: (2025)
por: Maity, Subhajit, et al.
Publicado: (2025)
Geometry and Perception Guided Gaussians for Multiview-consistent 3D Generation from a Single Image
por: Li, Pufan, et al.
Publicado: (2025)
por: Li, Pufan, et al.
Publicado: (2025)
StereoCrafter: Diffusion-based Generation of Long and High-fidelity Stereoscopic 3D from Monocular Videos
por: Zhao, Sijie, et al.
Publicado: (2024)
por: Zhao, Sijie, et al.
Publicado: (2024)
FLOWING: Implicit Neural Flows for Structure-Preserving Morphing
por: Bizzi, Arthur, et al.
Publicado: (2025)
por: Bizzi, Arthur, et al.
Publicado: (2025)
Data-Augmented Multimodal Feature Fusion for Multiclass Visual Recognition of Oral Cancer Lesions
por: Naoum, Joy, et al.
Publicado: (2025)
por: Naoum, Joy, et al.
Publicado: (2025)
PlacidDreamer: Advancing Harmony in Text-to-3D Generation
por: Huang, Shuo, et al.
Publicado: (2024)
por: Huang, Shuo, et al.
Publicado: (2024)
Robust Self-calibration of Focal Lengths from the Fundamental Matrix
por: Kocur, Viktor, et al.
Publicado: (2023)
por: Kocur, Viktor, et al.
Publicado: (2023)
Textured-GS: Gaussian Splatting with Spatially Defined Color and Opacity
por: Huang, Zhentao, et al.
Publicado: (2024)
por: Huang, Zhentao, et al.
Publicado: (2024)
Robust automatic brain vessel segmentation in 3D CTA scans using dynamic 4D-CTA data
por: Ceballos-Arroyo, Alberto Mario, et al.
Publicado: (2026)
por: Ceballos-Arroyo, Alberto Mario, et al.
Publicado: (2026)
View-Consistent 3D Scene Editing via Dual-Path Structural Correspondense and Semantic Continuity
por: Li, Pufan, et al.
Publicado: (2026)
por: Li, Pufan, et al.
Publicado: (2026)
Goal-conditioned reinforcement learning for ultrasound navigation guidance
por: Amadou, Abdoul Aziz, et al.
Publicado: (2024)
por: Amadou, Abdoul Aziz, et al.
Publicado: (2024)
Autoregressive Omni-Aware Outpainting for Open-Vocabulary 360-Degree Image Generation
por: Lu, Zhuqiang, et al.
Publicado: (2023)
por: Lu, Zhuqiang, et al.
Publicado: (2023)
Ada-adapter:Fast Few-shot Style Personlization of Diffusion Model with Pre-trained Image Encoder
por: Liu, Jia, et al.
Publicado: (2024)
por: Liu, Jia, et al.
Publicado: (2024)
Multi-Objective Optimization for Synthetic-to-Real Style Transfer
por: Chigot, Estelle, et al.
Publicado: (2026)
por: Chigot, Estelle, et al.
Publicado: (2026)
Diffusion Lens: Interpreting Text Encoders in Text-to-Image Pipelines
por: Toker, Michael, et al.
Publicado: (2024)
por: Toker, Michael, et al.
Publicado: (2024)
Context-Enriched Contrastive Loss: Enhancing Presentation of Inherent Sample Connections in Contrastive Learning Framework
por: Deng, Haojin, et al.
Publicado: (2025)
por: Deng, Haojin, et al.
Publicado: (2025)
Enhancing Explainable AI: A Hybrid Approach Combining GradCAM and LRP for CNN Interpretability
por: Dhore, Vaibhav, et al.
Publicado: (2024)
por: Dhore, Vaibhav, et al.
Publicado: (2024)
Inference-Time Scaling for Visual AutoRegressive modeling by Searching Representative Samples
por: Tang, Weidong, et al.
Publicado: (2026)
por: Tang, Weidong, et al.
Publicado: (2026)
Composite Data Augmentations for Synthetic Image Detection Against Real-World Perturbations
por: Amarantidou, Efthymia, et al.
Publicado: (2025)
por: Amarantidou, Efthymia, et al.
Publicado: (2025)
RipVIS: Rip Currents Video Instance Segmentation Benchmark for Beach Monitoring and Safety
por: Dumitriu, Andrei, et al.
Publicado: (2025)
por: Dumitriu, Andrei, et al.
Publicado: (2025)
Label Delay in Online Continual Learning
por: Csaba, Botos, et al.
Publicado: (2023)
por: Csaba, Botos, et al.
Publicado: (2023)
Ejemplares similares
-
CerberusDet: Unified Multi-Dataset Object Detection
por: Tolstykh, Irina, et al.
Publicado: (2024) -
Beyond Specialization: Assessing the Capabilities of MLLMs in Age and Gender Estimation
por: Kuprashevich, Maksim, et al.
Publicado: (2024) -
A Review of Pseudo-Labeling for Computer Vision
por: Kage, Patrick, et al.
Publicado: (2024) -
Synthetic Photography Detection: A Visual Guidance for Identifying Synthetic Images Created by AI
por: Mathys, Melanie, et al.
Publicado: (2024) -
Designing UNICORN: a Unified Benchmark for Imaging in Computational Pathology, Radiology, and Natural Language
por: Stegeman, Michelle, et al.
Publicado: (2026)