Salvato in:
| Autori principali: | Lai, Qiuxia, Li, Yu, Zeng, Ailing, Liu, Minhao, Sun, Hanqiu, Xu, Qiang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2021
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2108.03418 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Vector Quantization Prompting for Continual Learning
di: Jiao, Li, et al.
Pubblicazione: (2024)
di: Jiao, Li, et al.
Pubblicazione: (2024)
Learning 3D Representations for Spatial Intelligence from Unposed Multi-View Images
di: Zhou, Bo, et al.
Pubblicazione: (2026)
di: Zhou, Bo, et al.
Pubblicazione: (2026)
Structural Teacher-Student Normality Learning for Multi-Class Anomaly Detection and Localization
di: Deng, Hanqiu, et al.
Pubblicazione: (2024)
di: Deng, Hanqiu, et al.
Pubblicazione: (2024)
TEASER: Token Enhanced Spatial Modeling for Expressions Reconstruction
di: Liu, Yunfei, et al.
Pubblicazione: (2025)
di: Liu, Yunfei, et al.
Pubblicazione: (2025)
Partially Shared Concept Bottleneck Models
di: Zhao, Delong, et al.
Pubblicazione: (2025)
di: Zhao, Delong, et al.
Pubblicazione: (2025)
Spatial Information Bottleneck for Interpretable Visual Recognition
di: Shu, Kaixiang, et al.
Pubblicazione: (2025)
di: Shu, Kaixiang, et al.
Pubblicazione: (2025)
SeCG: Semantic-Enhanced 3D Visual Grounding via Cross-modal Graph Attention
di: Xiao, Feng, et al.
Pubblicazione: (2024)
di: Xiao, Feng, et al.
Pubblicazione: (2024)
MotionCraft: Crafting Whole-Body Motion with Plug-and-Play Multimodal Controls
di: Bian, Yuxuan, et al.
Pubblicazione: (2024)
di: Bian, Yuxuan, et al.
Pubblicazione: (2024)
Information Bottleneck-based Causal Attention for Multi-label Medical Image Recognition
di: Cui, Xiaoxiao, et al.
Pubblicazione: (2025)
di: Cui, Xiaoxiao, et al.
Pubblicazione: (2025)
OmniMotion-X: Versatile Multimodal Whole-Body Motion Generation
di: Xu, Guowei, et al.
Pubblicazione: (2025)
di: Xu, Guowei, et al.
Pubblicazione: (2025)
Temporal Consistency-Aware Text-to-Motion Generation
di: Wang, Hongsong, et al.
Pubblicazione: (2026)
di: Wang, Hongsong, et al.
Pubblicazione: (2026)
STAA-SNN: Spatial-Temporal Attention Aggregator for Spiking Neural Networks
di: Zhang, Tianqing, et al.
Pubblicazione: (2025)
di: Zhang, Tianqing, et al.
Pubblicazione: (2025)
The Dawn of Video Generation: Preliminary Explorations with SORA-like Models
di: Zeng, Ailing, et al.
Pubblicazione: (2024)
di: Zeng, Ailing, et al.
Pubblicazione: (2024)
Poly Kernel Inception Network for Remote Sensing Detection
di: Cai, Xinhao, et al.
Pubblicazione: (2024)
di: Cai, Xinhao, et al.
Pubblicazione: (2024)
Evaluating Text-to-Image Generative Models: An Empirical Study on Human Image Synthesis
di: Chen, Muxi, et al.
Pubblicazione: (2024)
di: Chen, Muxi, et al.
Pubblicazione: (2024)
P3P: Pseudo-3D Pre-training for Scaling 3D Voxel-based Masked Autoencoders
di: Chen, Xuechao, et al.
Pubblicazione: (2024)
di: Chen, Xuechao, et al.
Pubblicazione: (2024)
Bootstrap Fine-Grained Vision-Language Alignment for Unified Zero-Shot Anomaly Localization
di: Deng, Hanqiu, et al.
Pubblicazione: (2023)
di: Deng, Hanqiu, et al.
Pubblicazione: (2023)
Dual-Image Enhanced CLIP for Zero-Shot Anomaly Detection
di: Zhang, Zhaoxiang, et al.
Pubblicazione: (2024)
di: Zhang, Zhaoxiang, et al.
Pubblicazione: (2024)
Information Bottleneck-Guided Heterogeneous Graph Learning for Interpretable Neurodevelopmental Disorder Diagnosis
di: Li, Yueyang, et al.
Pubblicazione: (2025)
di: Li, Yueyang, et al.
Pubblicazione: (2025)
MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions
di: Ju, Xuan, et al.
Pubblicazione: (2024)
di: Ju, Xuan, et al.
Pubblicazione: (2024)
SpatialFormer: Semantic and Target Aware Attentions for Few-Shot Learning
di: Lai, Jinxiang, et al.
Pubblicazione: (2023)
di: Lai, Jinxiang, et al.
Pubblicazione: (2023)
GPAvatar: Generalizable and Precise Head Avatar from Image(s)
di: Chu, Xuangeng, et al.
Pubblicazione: (2024)
di: Chu, Xuangeng, et al.
Pubblicazione: (2024)
X-Pose: Detecting Any Keypoints
di: Yang, Jie, et al.
Pubblicazione: (2023)
di: Yang, Jie, et al.
Pubblicazione: (2023)
Language Guided Concept Bottleneck Models for Interpretable Continual Learning
di: Yu, Lu, et al.
Pubblicazione: (2025)
di: Yu, Lu, et al.
Pubblicazione: (2025)
A Conditional Probability Framework for Compositional Zero-shot Learning
di: Wu, Peng, et al.
Pubblicazione: (2025)
di: Wu, Peng, et al.
Pubblicazione: (2025)
Open-World Human-Object Interaction Detection via Multi-modal Prompts
di: Yang, Jie, et al.
Pubblicazione: (2024)
di: Yang, Jie, et al.
Pubblicazione: (2024)
Why Multimodal In-Context Learning Lags Behind? Unveiling the Inner Mechanisms and Bottlenecks
di: Wang, Yu, et al.
Pubblicazione: (2026)
di: Wang, Yu, et al.
Pubblicazione: (2026)
SC-Net: Robust Correspondence Learning via Spatial and Cross-Channel Context
di: Lin, Shuyuan, et al.
Pubblicazione: (2025)
di: Lin, Shuyuan, et al.
Pubblicazione: (2025)
Learning Unsupervised Gaze Representation via Eye Mask Driven Information Bottleneck
di: Jiang, Yangzhou, et al.
Pubblicazione: (2024)
di: Jiang, Yangzhou, et al.
Pubblicazione: (2024)
Concept-wise Attention for Fine-grained Concept Bottleneck Models
di: Zhong, Minghong, et al.
Pubblicazione: (2026)
di: Zhong, Minghong, et al.
Pubblicazione: (2026)
On the Perception Bottleneck of VLMs for Chart Understanding
di: Liu, Junteng, et al.
Pubblicazione: (2025)
di: Liu, Junteng, et al.
Pubblicazione: (2025)
Prototypical Information Bottlenecking and Disentangling for Multimodal Cancer Survival Prediction
di: Zhang, Yilan, et al.
Pubblicazione: (2024)
di: Zhang, Yilan, et al.
Pubblicazione: (2024)
A Training-Free Approach for Multi-ID Customization via Attention Adjustment and Spatial Control
di: Lin, Jiawei, et al.
Pubblicazione: (2025)
di: Lin, Jiawei, et al.
Pubblicazione: (2025)
TAB: Transformer Attention Bottlenecks enable User Intervention and Debugging in Vision-Language Models
di: Rahmanzadehgervi, Pooyan, et al.
Pubblicazione: (2024)
di: Rahmanzadehgervi, Pooyan, et al.
Pubblicazione: (2024)
Attend to Evidence: Evidence-Anchored Spatial Attention Supervision for Multimodal RLVR
di: Hu, Ruina, et al.
Pubblicazione: (2026)
di: Hu, Ruina, et al.
Pubblicazione: (2026)
Cell Variational Information Bottleneck Network
di: Zhai, Zhonghua, et al.
Pubblicazione: (2024)
di: Zhai, Zhonghua, et al.
Pubblicazione: (2024)
Information-Bottleneck Driven Binary Neural Network for Change Detection
di: Yin, Kaijie, et al.
Pubblicazione: (2025)
di: Yin, Kaijie, et al.
Pubblicazione: (2025)
PAMD: Plausibility-Aware Motion Diffusion Model for Long Dance Generation
di: Wang, Hongsong, et al.
Pubblicazione: (2025)
di: Wang, Hongsong, et al.
Pubblicazione: (2025)
Disentangled Representation Learning with Transmitted Information Bottleneck
di: Dang, Zhuohang, et al.
Pubblicazione: (2023)
di: Dang, Zhuohang, et al.
Pubblicazione: (2023)
RGB-Sonar Tracking Benchmark and Spatial Cross-Attention Transformer Tracker
di: Li, Yunfeng, et al.
Pubblicazione: (2024)
di: Li, Yunfeng, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Vector Quantization Prompting for Continual Learning
di: Jiao, Li, et al.
Pubblicazione: (2024) -
Learning 3D Representations for Spatial Intelligence from Unposed Multi-View Images
di: Zhou, Bo, et al.
Pubblicazione: (2026) -
Structural Teacher-Student Normality Learning for Multi-Class Anomaly Detection and Localization
di: Deng, Hanqiu, et al.
Pubblicazione: (2024) -
TEASER: Token Enhanced Spatial Modeling for Expressions Reconstruction
di: Liu, Yunfei, et al.
Pubblicazione: (2025) -
Partially Shared Concept Bottleneck Models
di: Zhao, Delong, et al.
Pubblicazione: (2025)