Prediction-Feedback DETR for Temporal Action Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Jihwan, Lee, Miso, Cho, Cheol-Ho, Lee, Jihyun, Heo, Jae-Pil |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Long-term Pre-training for Temporal Action Detection with Transformers
von: Kim, Jihwan, et al.
Veröffentlicht: (2024)
von: Kim, Jihwan, et al.
Veröffentlicht: (2024)
Activating Self-Attention for Multi-Scene Absolute Pose Regression
von: Lee, Miso, et al.
Veröffentlicht: (2024)
von: Lee, Miso, et al.
Veröffentlicht: (2024)
Boundary-Recovering Network for Temporal Action Detection
von: Kim, Jihwan, et al.
Veröffentlicht: (2024)
von: Kim, Jihwan, et al.
Veröffentlicht: (2024)
Temporally Consistent Long-Term Memory for 3D Single Object Tracking
von: Yoo, Jaejoon, et al.
Veröffentlicht: (2026)
von: Yoo, Jaejoon, et al.
Veröffentlicht: (2026)
White Aggregation and Restoration for Few-shot 3D Point Cloud Semantic Segmentation
von: Im, Jiyun, et al.
Veröffentlicht: (2025)
von: Im, Jiyun, et al.
Veröffentlicht: (2025)
Disambiguating 2D-3D Correspondences in Gaussian Splatting-based Feature Fields for Visual Localization
von: Lee, Miso, et al.
Veröffentlicht: (2026)
von: Lee, Miso, et al.
Veröffentlicht: (2026)
Mutually-Aware Feature Learning for Few-Shot Object Counting
von: Jeon, Yerim, et al.
Veröffentlicht: (2024)
von: Jeon, Yerim, et al.
Veröffentlicht: (2024)
Masking Matters: Unlocking the Spatial Reasoning Capabilities of LLMs for 3D Scene-Language Understanding
von: Jeon, Yerim, et al.
Veröffentlicht: (2025)
von: Jeon, Yerim, et al.
Veröffentlicht: (2025)
Noise-free Optimization in Early Training Steps for Image Super-Resolution
von: Lee, MinKyu, et al.
Veröffentlicht: (2023)
von: Lee, MinKyu, et al.
Veröffentlicht: (2023)
Temporal Alignment-Free Video Matching for Few-shot Action Recognition
von: Lee, SuBeen, et al.
Veröffentlicht: (2025)
von: Lee, SuBeen, et al.
Veröffentlicht: (2025)
Correlation-Guided Query-Dependency Calibration for Video Temporal Grounding
von: Moon, WonJun, et al.
Veröffentlicht: (2023)
von: Moon, WonJun, et al.
Veröffentlicht: (2023)
PDF-GS: Progressive Distractor Filtering for Robust 3D Gaussian Splatting
von: Seo, Kangmin, et al.
Veröffentlicht: (2026)
von: Seo, Kangmin, et al.
Veröffentlicht: (2026)
Prototypes are Balanced Units for Efficient and Effective Partially Relevant Video Retrieval
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
GSGAN: Adversarial Learning for Hierarchical Generation of 3D Gaussian Splats
von: Hyun, Sangeek, et al.
Veröffentlicht: (2024)
von: Hyun, Sangeek, et al.
Veröffentlicht: (2024)
Cross-scale Aligned Supervision for Training GANs
von: Hyun, Sangeek, et al.
Veröffentlicht: (2026)
von: Hyun, Sangeek, et al.
Veröffentlicht: (2026)
MDS-DETR: DETR with Masked Duplicate Suppressor
von: Lee, Chanho, et al.
Veröffentlicht: (2026)
von: Lee, Chanho, et al.
Veröffentlicht: (2026)
DiGIT: Multi-Dilated Gated Encoder and Central-Adjacent Region Integrated Decoder for Temporal Action Detection Transformer
von: Kim, Ho-Joong, et al.
Veröffentlicht: (2025)
von: Kim, Ho-Joong, et al.
Veröffentlicht: (2025)
Mitigating Semantic Collapse in Partially Relevant Video Retrieval
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
Style Injection in Diffusion: A Training-free Approach for Adapting Large-scale Diffusion Models for Style Transfer
von: Chung, Jiwoo, et al.
Veröffentlicht: (2023)
von: Chung, Jiwoo, et al.
Veröffentlicht: (2023)
Scalable GANs with Transformers
von: Hyun, Sangeek, et al.
Veröffentlicht: (2025)
von: Hyun, Sangeek, et al.
Veröffentlicht: (2025)
Mitigating Background Shift in Class-Incremental Semantic Segmentation
von: Park, Gilhan, et al.
Veröffentlicht: (2024)
von: Park, Gilhan, et al.
Veröffentlicht: (2024)
Fine-Tuning Visual Autoregressive Models for Subject-Driven Generation
von: Chung, Jiwoo, et al.
Veröffentlicht: (2025)
von: Chung, Jiwoo, et al.
Veröffentlicht: (2025)
Analyzing the Training Dynamics of Image Restoration Transformers: A Revisit to Layer Normalization
von: Lee, MinKyu, et al.
Veröffentlicht: (2025)
von: Lee, MinKyu, et al.
Veröffentlicht: (2025)
VLCounter: Text-aware Visual Representation for Zero-Shot Object Counting
von: Kang, Seunggu, et al.
Veröffentlicht: (2023)
von: Kang, Seunggu, et al.
Veröffentlicht: (2023)
Progressive Proxy Anchor Propagation for Unsupervised Semantic Segmentation
von: Seong, Hyun Seok, et al.
Veröffentlicht: (2024)
von: Seong, Hyun Seok, et al.
Veröffentlicht: (2024)
TE-TAD: Towards Full End-to-End Temporal Action Detection via Time-Aligned Coordinate Expression
von: Kim, Ho-Joong, et al.
Veröffentlicht: (2024)
von: Kim, Ho-Joong, et al.
Veröffentlicht: (2024)
Looking Beyond the Window: Global-Local Aligned CLIP for Training-free Open-Vocabulary Semantic Segmentation
von: Lee, ByeongCheol, et al.
Veröffentlicht: (2026)
von: Lee, ByeongCheol, et al.
Veröffentlicht: (2026)
Transferable Model-agnostic Vision-Language Model Adaptation for Efficient Weak-to-Strong Generalization
von: Park, Jihwan, et al.
Veröffentlicht: (2025)
von: Park, Jihwan, et al.
Veröffentlicht: (2025)
BAM-DETR: Boundary-Aligned Moment Detection Transformer for Temporal Sentence Grounding in Videos
von: Lee, Pilhyeon, et al.
Veröffentlicht: (2023)
von: Lee, Pilhyeon, et al.
Veröffentlicht: (2023)
Foreground-Covering Prototype Generation and Matching for SAM-Aided Few-Shot Segmentation
von: Park, Suho, et al.
Veröffentlicht: (2025)
von: Park, Suho, et al.
Veröffentlicht: (2025)
Auto-Encoded Supervision for Perceptual Image Super-Resolution
von: Lee, MinKyu, et al.
Veröffentlicht: (2024)
von: Lee, MinKyu, et al.
Veröffentlicht: (2024)
Sim-DETR: Unlock DETR for Temporal Sentence Grounding
von: Tang, Jiajin, et al.
Veröffentlicht: (2025)
von: Tang, Jiajin, et al.
Veröffentlicht: (2025)
Active Prompt Learning in Vision Language Models
von: Bang, Jihwan, et al.
Veröffentlicht: (2023)
von: Bang, Jihwan, et al.
Veröffentlicht: (2023)
Translation of Text Embedding via Delta Vector to Suppress Strongly Entangled Content in Text-to-Image Diffusion Models
von: Koh, Eunseo, et al.
Veröffentlicht: (2025)
von: Koh, Eunseo, et al.
Veröffentlicht: (2025)
Diversity-aware Channel Pruning for StyleGAN Compression
von: Chung, Jiwoo, et al.
Veröffentlicht: (2024)
von: Chung, Jiwoo, et al.
Veröffentlicht: (2024)
Auxiliary Descriptive Knowledge for Few-Shot Adaptation of Vision-Language Model
von: Lee, SuBeen, et al.
Veröffentlicht: (2025)
von: Lee, SuBeen, et al.
Veröffentlicht: (2025)
Object-Centric Representation Learning for Enhanced 3D Semantic Scene Graph Prediction
von: Heo, KunHo, et al.
Veröffentlicht: (2025)
von: Heo, KunHo, et al.
Veröffentlicht: (2025)
Classification Matters: Improving Video Action Detection with Class-Specific Attention
von: Lee, Jinsung, et al.
Veröffentlicht: (2024)
von: Lee, Jinsung, et al.
Veröffentlicht: (2024)
From Vicious to Virtuous Cycles: Synergistic Representation Learning for Unsupervised Video Object-Centric Learning
von: Seong, Hyun Seok, et al.
Veröffentlicht: (2026)
von: Seong, Hyun Seok, et al.
Veröffentlicht: (2026)
Selective Contrastive Learning for Weakly Supervised Affordance Grounding
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
von: Moon, WonJun, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Long-term Pre-training for Temporal Action Detection with Transformers
von: Kim, Jihwan, et al.
Veröffentlicht: (2024) -
Activating Self-Attention for Multi-Scene Absolute Pose Regression
von: Lee, Miso, et al.
Veröffentlicht: (2024) -
Boundary-Recovering Network for Temporal Action Detection
von: Kim, Jihwan, et al.
Veröffentlicht: (2024) -
Temporally Consistent Long-Term Memory for 3D Single Object Tracking
von: Yoo, Jaejoon, et al.
Veröffentlicht: (2026) -
White Aggregation and Restoration for Few-shot 3D Point Cloud Semantic Segmentation
von: Im, Jiyun, et al.
Veröffentlicht: (2025)