Saved in:
| Main Authors: | Lin, Shubo, Kou, Yutong, Wu, Zirui, Wang, Shaoru, Li, Bing, Hu, Weiming, Gao, Jin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2411.06780 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Experimental Study on Exploring Strong Lightweight Vision Transformers via Masked Image Modeling Pre-Training
by: Gao, Jin, et al.
Published: (2024)
by: Gao, Jin, et al.
Published: (2024)
ADA-Track++: End-to-End Multi-Camera 3D Multi-Object Tracking with Alternating Detection and Association
by: Ding, Shuxiao, et al.
Published: (2024)
by: Ding, Shuxiao, et al.
Published: (2024)
SynSeg: Feature Synergy for Multi-Category Contrastive Learning in End-to-End Open-Vocabulary Semantic Segmentation
by: Zhang, Weichen, et al.
Published: (2025)
by: Zhang, Weichen, et al.
Published: (2025)
Online Segment Any 3D Thing as Instance Tracking
by: Wang, Hanshi, et al.
Published: (2025)
by: Wang, Hanshi, et al.
Published: (2025)
Tracking by Detection and Query: An Efficient End-to-End Framework for Multi-Object Tracking
by: Jia, Shukun, et al.
Published: (2024)
by: Jia, Shukun, et al.
Published: (2024)
End-to-End Human Instance Matting
by: Liu, Qinglin, et al.
Published: (2024)
by: Liu, Qinglin, et al.
Published: (2024)
MI-DETR: A Strong Baseline for Moving Infrared Small Target Detection with Bio-Inspired Motion Integration
by: Liu, Nian, et al.
Published: (2026)
by: Liu, Nian, et al.
Published: (2026)
Efficient Multi-Camera Tokenization with Triplanes for End-to-End Driving
by: Ivanovic, Boris, et al.
Published: (2025)
by: Ivanovic, Boris, et al.
Published: (2025)
Towards Efficient and Effective Multi-Camera Encoding for End-to-End Driving
by: Yang, Jiawei, et al.
Published: (2025)
by: Yang, Jiawei, et al.
Published: (2025)
End-to-End On-Device Quantization-Aware Training for LLMs at Inference Cost
by: Tan, Qitao, et al.
Published: (2025)
by: Tan, Qitao, et al.
Published: (2025)
SMTrack: End-to-End Trained Spiking Neural Networks for Multi-Object Tracking in RGB Videos
by: Zhong, Pengzhi, et al.
Published: (2025)
by: Zhong, Pengzhi, et al.
Published: (2025)
MMPhysVideo: Scaling Physical Plausibility in Video Generation via Joint Multimodal Modeling
by: Lin, Shubo, et al.
Published: (2026)
by: Lin, Shubo, et al.
Published: (2026)
FusionTrack: End-to-End Multi-Object Tracking in Arbitrary Multi-View Environment
by: Li, Xiaohe, et al.
Published: (2025)
by: Li, Xiaohe, et al.
Published: (2025)
STORM: End-to-End Referring Multi-Object Tracking in Videos
by: Lu, Zijia, et al.
Published: (2026)
by: Lu, Zijia, et al.
Published: (2026)
S2-Track: A Simple yet Strong Approach for End-to-End 3D Multi-Object Tracking
by: Tang, Tao, et al.
Published: (2024)
by: Tang, Tao, et al.
Published: (2024)
An End-to-End Real-World Camera Imaging Pipeline
by: Xu, Kepeng, et al.
Published: (2024)
by: Xu, Kepeng, et al.
Published: (2024)
BridgeSim: Unveiling the OL-CL Gap in End-to-End Autonomous Driving
by: Zhao, Seth Z., et al.
Published: (2026)
by: Zhao, Seth Z., et al.
Published: (2026)
MMAPS: End-to-End Multi-Grained Multi-Modal Attribute-Aware Product Summarization
by: Chen, Tao, et al.
Published: (2023)
by: Chen, Tao, et al.
Published: (2023)
X-World: Controllable Ego-Centric Multi-Camera World Models for Scalable End-to-End Driving
by: Zheng, Chaoda, et al.
Published: (2026)
by: Zheng, Chaoda, et al.
Published: (2026)
MsaMIL-Net: An End-to-End Multi-Scale Aware Multiple Instance Learning Network for Efficient Whole Slide Image Classification
by: Wen, Jiangping, et al.
Published: (2025)
by: Wen, Jiangping, et al.
Published: (2025)
RC-AutoCalib: An End-to-End Radar-Camera Automatic Calibration Network
by: Luu, Van-Tin, et al.
Published: (2025)
by: Luu, Van-Tin, et al.
Published: (2025)
Multi-Agent End-to-End Vulnerability Management for Mitigating Recurring Vulnerabilities
by: Zheng, Zelong, et al.
Published: (2026)
by: Zheng, Zelong, et al.
Published: (2026)
Train Ego-Path Detection on Railway Tracks Using End-to-End Deep Learning
by: Laurent, Thomas
Published: (2024)
by: Laurent, Thomas
Published: (2024)
Referring Expression Instance Retrieval and A Strong End-to-End Baseline
by: Hao, Xiangzhao, et al.
Published: (2025)
by: Hao, Xiangzhao, et al.
Published: (2025)
BEEP3D: Box-Supervised End-to-End Pseudo-Mask Generation for 3D Instance Segmentation
by: Yoo, Youngju, et al.
Published: (2025)
by: Yoo, Youngju, et al.
Published: (2025)
End-To-End Training and Testing Gamification Framework to Learn Human Highway Driving
by: Jaladi, Satya R., et al.
Published: (2024)
by: Jaladi, Satya R., et al.
Published: (2024)
Content-Aware Foveated Camera for Multi-Target Tracking
by: Zang, Zihan, et al.
Published: (2025)
by: Zang, Zihan, et al.
Published: (2025)
SynAD: Enhancing Real-World End-to-End Autonomous Driving Models through Synthetic Data Integration
by: Kim, Jongsuk, et al.
Published: (2025)
by: Kim, Jongsuk, et al.
Published: (2025)
Dynamic Residual Encoding with Slide-Level Contrastive Learning for End-to-End Whole Slide Image Representation
by: Jin, Jing, et al.
Published: (2025)
by: Jin, Jing, et al.
Published: (2025)
Joint Speech and Text Training for LLM-Based End-to-End Spoken Dialogue State Tracking
by: Vendrame, Katia, et al.
Published: (2025)
by: Vendrame, Katia, et al.
Published: (2025)
D-VAT: End-to-End Visual Active Tracking for Micro Aerial Vehicles
by: Dionigi, Alberto, et al.
Published: (2023)
by: Dionigi, Alberto, et al.
Published: (2023)
U-DECN: End-to-End Underwater Object Detection ConvNet with Improved DeNoising Training
by: Liu, Zhuoyan, et al.
Published: (2024)
by: Liu, Zhuoyan, et al.
Published: (2024)
Analyzing Mitigation Strategies for Catastrophic Forgetting in End-to-End Training of Spoken Language Models
by: Hsiao, Chi-Yuan, et al.
Published: (2025)
by: Hsiao, Chi-Yuan, et al.
Published: (2025)
HENet: Hybrid Encoding for End-to-end Multi-task 3D Perception from Multi-view Cameras
by: Xia, Zhongyu, et al.
Published: (2024)
by: Xia, Zhongyu, et al.
Published: (2024)
UniSTPA: A Safety Analysis Framework for End-to-End Autonomous Driving
by: Kou, Hongrui, et al.
Published: (2025)
by: Kou, Hongrui, et al.
Published: (2025)
Physics-Aware Video Instance Removal Benchmark
by: Li, Zirui, et al.
Published: (2026)
by: Li, Zirui, et al.
Published: (2026)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
by: Raoufi, Behnam, et al.
Published: (2025)
by: Raoufi, Behnam, et al.
Published: (2025)
Constraint-Aware Flow Matching: Decision Aligned End-to-End Training for Constrained Sampling
by: Christopher, Jacob K., et al.
Published: (2026)
by: Christopher, Jacob K., et al.
Published: (2026)
End-to-End Multi-Track Reconstruction using Graph Neural Networks at Belle II
by: Reuter, Lea, et al.
Published: (2024)
by: Reuter, Lea, et al.
Published: (2024)
WebAgent-R1: Training Web Agents via End-to-End Multi-Turn Reinforcement Learning
by: Wei, Zhepei, et al.
Published: (2025)
by: Wei, Zhepei, et al.
Published: (2025)
Similar Items
-
An Experimental Study on Exploring Strong Lightweight Vision Transformers via Masked Image Modeling Pre-Training
by: Gao, Jin, et al.
Published: (2024) -
ADA-Track++: End-to-End Multi-Camera 3D Multi-Object Tracking with Alternating Detection and Association
by: Ding, Shuxiao, et al.
Published: (2024) -
SynSeg: Feature Synergy for Multi-Category Contrastive Learning in End-to-End Open-Vocabulary Semantic Segmentation
by: Zhang, Weichen, et al.
Published: (2025) -
Online Segment Any 3D Thing as Instance Tracking
by: Wang, Hanshi, et al.
Published: (2025) -
Tracking by Detection and Query: An Efficient End-to-End Framework for Multi-Object Tracking
by: Jia, Shukun, et al.
Published: (2024)