Salvato in:
| Autori principali: | Lee, Jaewook, Park, Yoel, Lee, Seulki |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2408.03663 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices
di: Kim, Bosung, et al.
Pubblicazione: (2025)
di: Kim, Bosung, et al.
Pubblicazione: (2025)
On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices
di: Kim, Bosung, et al.
Pubblicazione: (2025)
di: Kim, Bosung, et al.
Pubblicazione: (2025)
Partial Large Kernel CNNs for Efficient Super-Resolution
di: Lee, Dongheon, et al.
Pubblicazione: (2024)
di: Lee, Dongheon, et al.
Pubblicazione: (2024)
Free-Grained Hierarchical Visual Recognition
di: Park, Seulki, et al.
Pubblicazione: (2025)
di: Park, Seulki, et al.
Pubblicazione: (2025)
Leveraging Programmatically Generated Synthetic Data for Differentially Private Diffusion Training
di: Choi, Yujin, et al.
Pubblicazione: (2024)
di: Choi, Yujin, et al.
Pubblicazione: (2024)
Investigating and unmasking feature-level vulnerabilities of CNNs to adversarial perturbations
di: Coppola, Davide, et al.
Pubblicazione: (2024)
di: Coppola, Davide, et al.
Pubblicazione: (2024)
Lightweight Channel Attention for Efficient CNNs
di: Kanaparthi, Prem Babu, et al.
Pubblicazione: (2026)
di: Kanaparthi, Prem Babu, et al.
Pubblicazione: (2026)
Visually Consistent Hierarchical Image Classification
di: Park, Seulki, et al.
Pubblicazione: (2024)
di: Park, Seulki, et al.
Pubblicazione: (2024)
XStreamVGGT: Extremely Memory-Efficient Streaming Vision Geometry Grounded Transformer with KV Cache Compression
di: Su, Zunhai, et al.
Pubblicazione: (2026)
di: Su, Zunhai, et al.
Pubblicazione: (2026)
XStreamVGGT: Extremely Memory-Efficient Streaming Vision Geometry Grounded Transformer with KV Cache Compression
di: Su, Zunhai, et al.
Pubblicazione: (2026)
di: Su, Zunhai, et al.
Pubblicazione: (2026)
ViscNet: Vision-Based In-line Viscometry for Fluid Mixing Process
di: Sohn, Jongwon, et al.
Pubblicazione: (2025)
di: Sohn, Jongwon, et al.
Pubblicazione: (2025)
ExBluRF: Efficient Radiance Fields for Extreme Motion Blurred Images
di: Lee, Dongwoo, et al.
Pubblicazione: (2023)
di: Lee, Dongwoo, et al.
Pubblicazione: (2023)
SHViT: Single-Head Vision Transformer with Memory Efficient Macro Design
di: Yun, Seokju, et al.
Pubblicazione: (2024)
di: Yun, Seokju, et al.
Pubblicazione: (2024)
B-cos Alignment for Inherently Interpretable CNNs and Vision Transformers
di: Böhle, Moritz, et al.
Pubblicazione: (2023)
di: Böhle, Moritz, et al.
Pubblicazione: (2023)
ZIP: An Efficient Zeroth-order Prompt Tuning for Black-box Vision-Language Models
di: Park, Seonghwan, et al.
Pubblicazione: (2025)
di: Park, Seonghwan, et al.
Pubblicazione: (2025)
Parameter-Efficient Architectural Modifications for Translation-Invariant CNNs
di: Alabau-Bosque, Nuria, et al.
Pubblicazione: (2026)
di: Alabau-Bosque, Nuria, et al.
Pubblicazione: (2026)
CNNs, Transformers, Hybrid, and Vision Language Models for Skin Cancer Detection
di: Dey, Durjoy, et al.
Pubblicazione: (2026)
di: Dey, Durjoy, et al.
Pubblicazione: (2026)
KOALA: Empirical Lessons Toward Memory-Efficient and Fast Diffusion Models for Text-to-Image Synthesis
di: Lee, Youngwan, et al.
Pubblicazione: (2023)
di: Lee, Youngwan, et al.
Pubblicazione: (2023)
MEIL-NeRF: Memory-Efficient Incremental Learning of Neural Radiance Fields
di: Chung, Jaeyoung, et al.
Pubblicazione: (2022)
di: Chung, Jaeyoung, et al.
Pubblicazione: (2022)
TADFormer : Task-Adaptive Dynamic Transformer for Efficient Multi-Task Learning
di: Baek, Seungmin, et al.
Pubblicazione: (2025)
di: Baek, Seungmin, et al.
Pubblicazione: (2025)
BEM: Training-Free Background Embedding Memory for False-Positive Suppression in Real-Time Fixed-Background Camera
di: Park, Junwoo, et al.
Pubblicazione: (2026)
di: Park, Junwoo, et al.
Pubblicazione: (2026)
OA-CNNs: Omni-Adaptive Sparse CNNs for 3D Semantic Segmentation
di: Peng, Bohao, et al.
Pubblicazione: (2024)
di: Peng, Bohao, et al.
Pubblicazione: (2024)
Vehicle Classification under Extreme Imbalance: A Comparative Study of Ensemble Learning and CNNs
di: Syarubany, Abu Hanif Muhammad
Pubblicazione: (2025)
di: Syarubany, Abu Hanif Muhammad
Pubblicazione: (2025)
ReDistill: Residual Encoded Distillation for Peak Memory Reduction of CNNs
di: Chen, Fang, et al.
Pubblicazione: (2024)
di: Chen, Fang, et al.
Pubblicazione: (2024)
CAD: Memory Efficient Convolutional Adapter for Segment Anything
di: Kim, Joohyeok, et al.
Pubblicazione: (2024)
di: Kim, Joohyeok, et al.
Pubblicazione: (2024)
Extreme Point Supervised Instance Segmentation
di: Lee, Hyeonjun, et al.
Pubblicazione: (2024)
di: Lee, Hyeonjun, et al.
Pubblicazione: (2024)
MoRGS: Efficient Per-Gaussian Motion Reasoning for Streamable Dynamic 3D Scenes
di: Lee, Wonjoon, et al.
Pubblicazione: (2026)
di: Lee, Wonjoon, et al.
Pubblicazione: (2026)
Promoting CNNs with Cross-Architecture Knowledge Distillation for Efficient Monocular Depth Estimation
di: Zheng, Zhimeng, et al.
Pubblicazione: (2024)
di: Zheng, Zhimeng, et al.
Pubblicazione: (2024)
How Does Vision-Language Adaptation Impact the Safety of Vision Language Models?
di: Lee, Seongyun, et al.
Pubblicazione: (2024)
di: Lee, Seongyun, et al.
Pubblicazione: (2024)
ACM-UNet: Adaptive Integration of CNNs and Mamba for Efficient Medical Image Segmentation
di: Huang, Jing, et al.
Pubblicazione: (2025)
di: Huang, Jing, et al.
Pubblicazione: (2025)
Combining Transformers and CNNs for Efficient Object Detection in High-Resolution Satellite Imagery
di: Drapier, Nicolas, et al.
Pubblicazione: (2025)
di: Drapier, Nicolas, et al.
Pubblicazione: (2025)
Continuous Memory Representation for Anomaly Detection
di: Lee, Joo Chan, et al.
Pubblicazione: (2024)
di: Lee, Joo Chan, et al.
Pubblicazione: (2024)
Self-Supervised Vision Transformers Are Efficient Segmentation Learners for Imperfect Labels
di: Lee, Seungho, et al.
Pubblicazione: (2024)
di: Lee, Seungho, et al.
Pubblicazione: (2024)
Efficient Hyperparameter Importance Assessment for CNNs
di: Wang, Ruinan, et al.
Pubblicazione: (2024)
di: Wang, Ruinan, et al.
Pubblicazione: (2024)
The Role of Masking for Efficient Supervised Knowledge Distillation of Vision Transformers
di: Son, Seungwoo, et al.
Pubblicazione: (2023)
di: Son, Seungwoo, et al.
Pubblicazione: (2023)
EfficientViM: Efficient Vision Mamba with Hidden State Mixer based State Space Duality
di: Lee, Sanghyeok, et al.
Pubblicazione: (2024)
di: Lee, Sanghyeok, et al.
Pubblicazione: (2024)
RoCOCO: Robustness Benchmark of MS-COCO to Stress-test Image-Text Matching Models
di: Park, Seulki, et al.
Pubblicazione: (2023)
di: Park, Seulki, et al.
Pubblicazione: (2023)
Bridging Vision and Language Spaces with Assignment Prediction
di: Park, Jungin, et al.
Pubblicazione: (2024)
di: Park, Jungin, et al.
Pubblicazione: (2024)
MoE-GRPO: Optimizing Mixture-of-Experts via Reinforcement Learning in Vision-Language Models
di: Ko, Dohwan, et al.
Pubblicazione: (2026)
di: Ko, Dohwan, et al.
Pubblicazione: (2026)
Exploring Synergistic Ensemble Learning: Uniting CNNs, MLP-Mixers, and Vision Transformers to Enhance Image Classification
di: Bashar, Mk, et al.
Pubblicazione: (2025)
di: Bashar, Mk, et al.
Pubblicazione: (2025)
Documenti analoghi
-
On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices
di: Kim, Bosung, et al.
Pubblicazione: (2025) -
On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices
di: Kim, Bosung, et al.
Pubblicazione: (2025) -
Partial Large Kernel CNNs for Efficient Super-Resolution
di: Lee, Dongheon, et al.
Pubblicazione: (2024) -
Free-Grained Hierarchical Visual Recognition
di: Park, Seulki, et al.
Pubblicazione: (2025) -
Leveraging Programmatically Generated Synthetic Data for Differentially Private Diffusion Training
di: Choi, Yujin, et al.
Pubblicazione: (2024)