Salvato in:
| Autori principali: | Lei, Xing, Liu, Longjun, Zhou, Zhiheng, Sun, Hongbin, Zheng, Nanning |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2403.06352 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PMT: Progressive Mean Teacher via Exploring Temporal Consistency for Semi-Supervised Medical Image Segmentation
di: Gao, Ning, et al.
Pubblicazione: (2024)
di: Gao, Ning, et al.
Pubblicazione: (2024)
Voxel or Pillar: Exploring Efficient Point Cloud Representation for 3D Object Detection
di: Huang, Yuhao, et al.
Pubblicazione: (2023)
di: Huang, Yuhao, et al.
Pubblicazione: (2023)
Exploring Architectures for CNN-Based Word Spotting
di: Rusakov, Eugen, et al.
Pubblicazione: (2018)
di: Rusakov, Eugen, et al.
Pubblicazione: (2018)
FreqPDE: Rethinking Positional Depth Embedding for Multi-View 3D Object Detection Transformers
di: Su, Haisheng, et al.
Pubblicazione: (2025)
di: Su, Haisheng, et al.
Pubblicazione: (2025)
Leveraging Anchor-based LiDAR 3D Object Detection via Point Assisted Sample Selection
di: Chen, Shitao, et al.
Pubblicazione: (2024)
di: Chen, Shitao, et al.
Pubblicazione: (2024)
Learning to Infer Unseen Single-/Multi-Attribute-Object Compositions with Graph Networks
di: Chen, Hui, et al.
Pubblicazione: (2020)
di: Chen, Hui, et al.
Pubblicazione: (2020)
Selective Transfer Learning of Cross-Modality Distillation for Monocular 3D Object Detection
di: Ding, Rui, et al.
Pubblicazione: (2026)
di: Ding, Rui, et al.
Pubblicazione: (2026)
Beyond the Embedding Bottleneck: Adaptive Retrieval-Augmented 3D CT Report Generation
di: Liang, Renjie, et al.
Pubblicazione: (2026)
di: Liang, Renjie, et al.
Pubblicazione: (2026)
Sketch-to-Architecture: Generative AI-aided Architectural Design
di: Li, Pengzhi, et al.
Pubblicazione: (2024)
di: Li, Pengzhi, et al.
Pubblicazione: (2024)
EVA: Mixture-of-Experts Semantic Variant Alignment for Compositional Zero-Shot Learning
di: Zhang, Xiao, et al.
Pubblicazione: (2025)
di: Zhang, Xiao, et al.
Pubblicazione: (2025)
HybridMIM: A Hybrid Masked Image Modeling Framework for 3D Medical Image Segmentation
di: Xing, Zhaohu, et al.
Pubblicazione: (2023)
di: Xing, Zhaohu, et al.
Pubblicazione: (2023)
Robust Noisy Label Learning via Two-Stream Sample Distillation
di: Bai, Sihan, et al.
Pubblicazione: (2024)
di: Bai, Sihan, et al.
Pubblicazione: (2024)
Hardware-Friendly Static Quantization Method for Video Diffusion Transformers
di: Yi, Sanghyun, et al.
Pubblicazione: (2025)
di: Yi, Sanghyun, et al.
Pubblicazione: (2025)
CNN-JEPA: Self-Supervised Pretraining Convolutional Neural Networks Using Joint Embedding Predictive Architecture
di: Kalapos, András, et al.
Pubblicazione: (2024)
di: Kalapos, András, et al.
Pubblicazione: (2024)
DAMap: Distance-aware MapNet for High Quality HD Map Construction
di: Dong, Jinpeng, et al.
Pubblicazione: (2025)
di: Dong, Jinpeng, et al.
Pubblicazione: (2025)
StructVPR++: Distill Structural and Semantic Knowledge with Weighting Samples for Visual Place Recognition
di: Shen, Yanqing, et al.
Pubblicazione: (2025)
di: Shen, Yanqing, et al.
Pubblicazione: (2025)
SeqTrack3D: Exploring Sequence Information for Robust 3D Point Cloud Tracking
di: Lin, Yu, et al.
Pubblicazione: (2024)
di: Lin, Yu, et al.
Pubblicazione: (2024)
Cross-Task Benchmarking of CNN Architectures
di: Sherawat, Kamal, et al.
Pubblicazione: (2026)
di: Sherawat, Kamal, et al.
Pubblicazione: (2026)
A Compact Hybrid Convolution--Frequency State Space Network for Learned Image Compression
di: Pan, Haodong, et al.
Pubblicazione: (2025)
di: Pan, Haodong, et al.
Pubblicazione: (2025)
GlobalPaint: Spatiotemporal Coherent Video Outpainting with Global Feature Guidance
di: Pan, Yueming, et al.
Pubblicazione: (2026)
di: Pan, Yueming, et al.
Pubblicazione: (2026)
GIC-DLC: Differentiable Logic Circuits for Hardware-Friendly Grayscale Image Compression
di: Aczel, Till, et al.
Pubblicazione: (2026)
di: Aczel, Till, et al.
Pubblicazione: (2026)
See Through Their Minds: Learning Transferable Neural Representation from Cross-Subject fMRI
di: Liu, Yulong, et al.
Pubblicazione: (2024)
di: Liu, Yulong, et al.
Pubblicazione: (2024)
Unsupervised Domain Adaption Harnessing Vision-Language Pre-training
di: Zhou, Wenlve, et al.
Pubblicazione: (2024)
di: Zhou, Wenlve, et al.
Pubblicazione: (2024)
A Hidden Semantic Bottleneck in Conditional Embeddings of Diffusion Transformers
di: Pham, Trung X., et al.
Pubblicazione: (2026)
di: Pham, Trung X., et al.
Pubblicazione: (2026)
VFM-VAE: Vision Foundation Models Can Be Good Tokenizers for Latent Diffusion Models
di: Bi, Tianci, et al.
Pubblicazione: (2025)
di: Bi, Tianci, et al.
Pubblicazione: (2025)
UniHOI: Unified Human-Object Interaction Understanding via Unified Token Space
di: Yang, Panqi, et al.
Pubblicazione: (2025)
di: Yang, Panqi, et al.
Pubblicazione: (2025)
FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training
di: Cao, Anjia, et al.
Pubblicazione: (2024)
di: Cao, Anjia, et al.
Pubblicazione: (2024)
LEFormer: A Hybrid CNN-Transformer Architecture for Accurate Lake Extraction from Remote Sensing Imagery
di: Chen, Ben, et al.
Pubblicazione: (2023)
di: Chen, Ben, et al.
Pubblicazione: (2023)
Iterative Filter Pruning for Concatenation-based CNN Architectures
di: Pavlitska, Svetlana, et al.
Pubblicazione: (2024)
di: Pavlitska, Svetlana, et al.
Pubblicazione: (2024)
ForestLPR: LiDAR Place Recognition in Forests Attentioning Multiple BEV Density Images
di: Shen, Yanqing, et al.
Pubblicazione: (2025)
di: Shen, Yanqing, et al.
Pubblicazione: (2025)
Analyzing the Mechanism of Attention Collapse in VGGT from a Dynamics Perspective
di: Li, Huan, et al.
Pubblicazione: (2025)
di: Li, Huan, et al.
Pubblicazione: (2025)
Efficient High-Resolution Visual Representation Learning with State Space Model for Human Pose Estimation
di: Zhang, Hao, et al.
Pubblicazione: (2024)
di: Zhang, Hao, et al.
Pubblicazione: (2024)
A Comparative Study of Adversarial Robustness in CNN and CNN-ANFIS Architectures
di: Shankar, Kaaustaaub, et al.
Pubblicazione: (2026)
di: Shankar, Kaaustaaub, et al.
Pubblicazione: (2026)
Efficient Hybrid CNN-GNN Architecture for Monocular Depth Estimation
di: Narayan, Ishan
Pubblicazione: (2026)
di: Narayan, Ishan
Pubblicazione: (2026)
Information Bottleneck Approach to Spatial Attention Learning
di: Lai, Qiuxia, et al.
Pubblicazione: (2021)
di: Lai, Qiuxia, et al.
Pubblicazione: (2021)
Editable Concept Bottleneck Models
di: Hu, Lijie, et al.
Pubblicazione: (2024)
di: Hu, Lijie, et al.
Pubblicazione: (2024)
Cascaded Robust Rectification for Arbitrary Document Images
di: Wang, Chaoyun, et al.
Pubblicazione: (2025)
di: Wang, Chaoyun, et al.
Pubblicazione: (2025)
Downscaling Intelligence: Exploring Perception and Reasoning Bottlenecks in Small Multimodal Models
di: Endo, Mark, et al.
Pubblicazione: (2025)
di: Endo, Mark, et al.
Pubblicazione: (2025)
MambaVesselNet++: A Hybrid CNN-Mamba Architecture for Medical Image Segmentation
di: Xu, Qing, et al.
Pubblicazione: (2025)
di: Xu, Qing, et al.
Pubblicazione: (2025)
One-Shot Multilingual Font Generation Via ViT
di: Wang, Zhiheng, et al.
Pubblicazione: (2024)
di: Wang, Zhiheng, et al.
Pubblicazione: (2024)
Documenti analoghi
-
PMT: Progressive Mean Teacher via Exploring Temporal Consistency for Semi-Supervised Medical Image Segmentation
di: Gao, Ning, et al.
Pubblicazione: (2024) -
Voxel or Pillar: Exploring Efficient Point Cloud Representation for 3D Object Detection
di: Huang, Yuhao, et al.
Pubblicazione: (2023) -
Exploring Architectures for CNN-Based Word Spotting
di: Rusakov, Eugen, et al.
Pubblicazione: (2018) -
FreqPDE: Rethinking Positional Depth Embedding for Multi-View 3D Object Detection Transformers
di: Su, Haisheng, et al.
Pubblicazione: (2025) -
Leveraging Anchor-based LiDAR 3D Object Detection via Point Assisted Sample Selection
di: Chen, Shitao, et al.
Pubblicazione: (2024)