DiGIT: Multi-Dilated Gated Encoder and Central-Adjacent Region Integrated Decoder for Temporal Action Detection Transformer
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Ho-Joong, Lee, Yearang, Hong, Jung-Ho, Lee, Seong-Whan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TE-TAD: Towards Full End-to-End Temporal Action Detection via Time-Aligned Coordinate Expression
by: Kim, Ho-Joong, et al.
Published: (2024)
by: Kim, Ho-Joong, et al.
Published: (2024)
Comprehensive Information Bottleneck for Unveiling Universal Attribution to Interpret Vision Transformers
by: Hong, Jung-Ho, et al.
Published: (2025)
by: Hong, Jung-Ho, et al.
Published: (2025)
ClipTBP: Clip-Pair based Temporal Boundary Prediction with Boundary-Aware Learning for Moment Retrieval
by: Kim, Ji-Hyeon, et al.
Published: (2026)
by: Kim, Ji-Hyeon, et al.
Published: (2026)
FIQ: Fundamental Question Generation with the Integration of Question Embeddings for Video Question Answering
by: Oh, Ju-Young, et al.
Published: (2025)
by: Oh, Ju-Young, et al.
Published: (2025)
MIRe: Enhancing Multimodal Queries Representation via Fusion-Free Modality Interaction for Multimodal Retrieval
by: Ju, Yeong-Joon, et al.
Published: (2024)
by: Ju, Yeong-Joon, et al.
Published: (2024)
Enhancing Spatio-Temporal Zero-shot Action Recognition with Language-driven Description Attributes
by: Kim, Yehna, et al.
Published: (2025)
by: Kim, Yehna, et al.
Published: (2025)
AM-SORT: Adaptable Motion Predictor with Historical Trajectory Embedding for Multi-Object Tracking
by: Kim, Vitaliy, et al.
Published: (2024)
by: Kim, Vitaliy, et al.
Published: (2024)
Towards Better Visualizing the Decision Basis of Networks via Unfold and Conquer Attribution Guidance
by: Hong, Jung-Ho, et al.
Published: (2023)
by: Hong, Jung-Ho, et al.
Published: (2023)
Edge Conditional Node Update Graph Neural Network for Multi-variate Time Series Anomaly Detection
by: Jo, Hayoung, et al.
Published: (2024)
by: Jo, Hayoung, et al.
Published: (2024)
Prediction-Feedback DETR for Temporal Action Detection
by: Kim, Jihwan, et al.
Published: (2024)
by: Kim, Jihwan, et al.
Published: (2024)
ACoRN: Noise-Robust Abstractive Compression in Retrieval-Augmented Language Models
by: Kim, Singon, et al.
Published: (2025)
by: Kim, Singon, et al.
Published: (2025)
Multi-Context Temporal Consistent Modeling for Referring Video Object Segmentation
by: Choi, Sun-Hyuk, et al.
Published: (2025)
by: Choi, Sun-Hyuk, et al.
Published: (2025)
ChemFixer: Correcting Invalid Molecules to Unlock Previously Unseen Chemical Space
by: Park, Jun-Hyoung, et al.
Published: (2025)
by: Park, Jun-Hyoung, et al.
Published: (2025)
Towards Dynamic Neural Communication and Speech Neuroprosthesis Based on Viseme Decoding
by: Park, Ji-Ha, et al.
Published: (2025)
by: Park, Ji-Ha, et al.
Published: (2025)
Meta-cognitive Multi-scale Hierarchical Reasoning for Motor Imagery Decoding
by: Kim, Si-Hyun, et al.
Published: (2025)
by: Kim, Si-Hyun, et al.
Published: (2025)
RaDL: Relation-aware Disentangled Learning for Multi-Instance Text-to-Image Generation
by: Park, Geon, et al.
Published: (2025)
by: Park, Geon, et al.
Published: (2025)
LUMINA-Net: Low-light Upgrade through Multi-stage Illumination and Noise Adaptation Network for Image Enhancement
by: Siddiqua, Namrah, et al.
Published: (2025)
by: Siddiqua, Namrah, et al.
Published: (2025)
Integrating Locality-Aware Attention with Transformers for General Geometry PDEs
by: Koh, Minsu, et al.
Published: (2025)
by: Koh, Minsu, et al.
Published: (2025)
Text-guided Weakly Supervised Framework for Dynamic Facial Expression Recognition
by: Jung, Gunho, et al.
Published: (2025)
by: Jung, Gunho, et al.
Published: (2025)
On the Correctness of the Generalized Isotonic Recursive Partitioning Algorithm
by: Won, Joong-Ho, et al.
Published: (2024)
by: Won, Joong-Ho, et al.
Published: (2024)
Semantic Prioritization in Visual Counterfactual Explanations with Weighted Segmentation and Auto-Adaptive Region Selection
by: Zhang, Lintong, et al.
Published: (2025)
by: Zhang, Lintong, et al.
Published: (2025)
MCoT-RE: Multi-Faceted Chain-of-Thought and Re-Ranking for Training-Free Zero-Shot Composed Image Retrieval
by: Park, Jeong-Woo, et al.
Published: (2025)
by: Park, Jeong-Woo, et al.
Published: (2025)
Optimal Multi-Task Learning at Regularization Horizon for Speech Translation Task
by: Jung, JungHo, et al.
Published: (2025)
by: Jung, JungHo, et al.
Published: (2025)
Long-term Pre-training for Temporal Action Detection with Transformers
by: Kim, Jihwan, et al.
Published: (2024)
by: Kim, Jihwan, et al.
Published: (2024)
DiEmo-TTS: Disentangled Emotion Representations via Self-Supervised Distillation for Cross-Speaker Emotion Transfer in Text-to-Speech
by: Cho, Deok-Hyeon, et al.
Published: (2025)
by: Cho, Deok-Hyeon, et al.
Published: (2025)
Personalized Continual EEG Decoding: Retaining and Transferring Knowledge
by: Li, Dan, et al.
Published: (2024)
by: Li, Dan, et al.
Published: (2024)
Diversify and Conquer: Open-set Disagreement for Robust Semi-supervised Learning with Outliers
by: Kong, Heejo, et al.
Published: (2025)
by: Kong, Heejo, et al.
Published: (2025)
FlipConcept: Tuning-Free Multi-Concept Personalization for Text-to-Image Generation
by: Woo, Young Beom, et al.
Published: (2025)
by: Woo, Young Beom, et al.
Published: (2025)
Restoring native posterior tibial slope within 4° leads to better clinical outcomes after cruciate‐retaining robot‐assisted total knee arthroplasty with functional alignment
by: Young Tak Cho, et al.
Published: (2025)
by: Young Tak Cho, et al.
Published: (2025)
TIFu: Tri-directional Implicit Function for High-Fidelity 3D Character Reconstruction
by: Lim, Byoungsung, et al.
Published: (2024)
by: Lim, Byoungsung, et al.
Published: (2024)
FAR-Net: Multi-Stage Fusion Network with Enhanced Semantic Alignment and Adaptive Reconciliation for Composed Image Retrieval
by: Park, Jeong-Woo, et al.
Published: (2025)
by: Park, Jeong-Woo, et al.
Published: (2025)
PeriodWave: Multi-Period Flow Matching for High-Fidelity Waveform Generation
by: Lee, Sang-Hoon, et al.
Published: (2024)
by: Lee, Sang-Hoon, et al.
Published: (2024)
Local Representative Token Guided Merging for Text-to-Image Generation
by: Lee, Min-Jeong, et al.
Published: (2025)
by: Lee, Min-Jeong, et al.
Published: (2025)
TranSentence: Speech-to-speech Translation via Language-agnostic Sentence-level Speech Encoding without Language-parallel Data
by: Kim, Seung-Bin, et al.
Published: (2024)
by: Kim, Seung-Bin, et al.
Published: (2024)
Bioassay of Chironomid Larvae (Diptera: Chironomidae) Using Chlorine Dioxide Solution
by: Jang Ho Lee, et al.
Published: (2025)
by: Jang Ho Lee, et al.
Published: (2025)
Integrating Model-Based Footstep Planning with Model-Free Reinforcement Learning for Dynamic Legged Locomotion
by: Lee, Ho Jae, et al.
Published: (2024)
by: Lee, Ho Jae, et al.
Published: (2024)
Aligning Humans and Robots via Reinforcement Learning from Implicit Human Feedback
by: Kim, Suzie, et al.
Published: (2025)
by: Kim, Suzie, et al.
Published: (2025)
Appearance Debiased Gaze Estimation via Stochastic Subject-Wise Adversarial Learning
by: Kim, Suneung, et al.
Published: (2024)
by: Kim, Suneung, et al.
Published: (2024)
Multi-stage Prompt Refinement for Mitigating Hallucinations in Large Language Models
by: Shim, Jung-Woo, et al.
Published: (2025)
by: Shim, Jung-Woo, et al.
Published: (2025)
Citation-Closure Retrieval and Per-Rule Attribution for Real-World Regulatory Compliance Question Answering
by: Ju, Yeong-Joon, et al.
Published: (2026)
by: Ju, Yeong-Joon, et al.
Published: (2026)
Similar Items
-
TE-TAD: Towards Full End-to-End Temporal Action Detection via Time-Aligned Coordinate Expression
by: Kim, Ho-Joong, et al.
Published: (2024) -
Comprehensive Information Bottleneck for Unveiling Universal Attribution to Interpret Vision Transformers
by: Hong, Jung-Ho, et al.
Published: (2025) -
ClipTBP: Clip-Pair based Temporal Boundary Prediction with Boundary-Aware Learning for Moment Retrieval
by: Kim, Ji-Hyeon, et al.
Published: (2026) -
FIQ: Fundamental Question Generation with the Integration of Question Embeddings for Video Question Answering
by: Oh, Ju-Young, et al.
Published: (2025) -
MIRe: Enhancing Multimodal Queries Representation via Fusion-Free Modality Interaction for Multimodal Retrieval
by: Ju, Yeong-Joon, et al.
Published: (2024)