Saved in:
| Main Authors: | Lu, Han, Xie, Yichen, Yang, Xiaokang, Yan, Junchi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2403.10069 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ActiveAD: Planning-Oriented Active Learning for End-to-End Autonomous Driving
by: Lu, Han, et al.
Published: (2024)
by: Lu, Han, et al.
Published: (2024)
Rethinking Classifier Re-Training in Long-Tailed Recognition: A Simple Logits Retargeting Approach
by: Lu, Han, et al.
Published: (2024)
by: Lu, Han, et al.
Published: (2024)
Unified Batch Normalization: Identifying and Alleviating the Feature Condensation in Batch Normalization and a Unified Framework
by: Wang, Shaobo, et al.
Published: (2023)
by: Wang, Shaobo, et al.
Published: (2023)
PointOBB-v3: Expanding Performance Boundaries of Single Point-Supervised Oriented Object Detection
by: Zhang, Peiyuan, et al.
Published: (2025)
by: Zhang, Peiyuan, et al.
Published: (2025)
MotionGPT: Finetuned LLMs Are General-Purpose Motion Generators
by: Zhang, Yaqi, et al.
Published: (2023)
by: Zhang, Yaqi, et al.
Published: (2023)
Wholly-WOOD: Wholly Leveraging Diversified-quality Labels for Weakly-supervised Oriented Object Detection
by: Yu, Yi, et al.
Published: (2025)
by: Yu, Yi, et al.
Published: (2025)
Rethinking Video Tokenization: A Conditioned Diffusion-based Approach
by: Yang, Nianzu, et al.
Published: (2025)
by: Yang, Nianzu, et al.
Published: (2025)
ARS-DETR: Aspect Ratio-Sensitive Detection Transformer for Aerial Oriented Object Detection
by: Zeng, Ying, et al.
Published: (2023)
by: Zeng, Ying, et al.
Published: (2023)
Video Finetuning Improves Reasoning Between Frames
by: Yang, Ruiqi, et al.
Published: (2025)
by: Yang, Ruiqi, et al.
Published: (2025)
MISS: A Generative Pretraining and Finetuning Approach for Med-VQA
by: Chen, Jiawei, et al.
Published: (2024)
by: Chen, Jiawei, et al.
Published: (2024)
FineRMoE: Dimension Expansion for Finer-Grained Expert with Its Upcycling Approach
by: Liao, Ning, et al.
Published: (2026)
by: Liao, Ning, et al.
Published: (2026)
Bi-Level Optimization for Single Domain Generalization
by: Heidari, Marzi, et al.
Published: (2026)
by: Heidari, Marzi, et al.
Published: (2026)
Augmentation Matters: A Mix-Paste Method for X-Ray Prohibited Item Detection under Noisy Annotations
by: Chen, Ruikang, et al.
Published: (2025)
by: Chen, Ruikang, et al.
Published: (2025)
SS-ADA: A Semi-Supervised Active Domain Adaptation Framework for Semantic Segmentation
by: Yan, Weihao, et al.
Published: (2024)
by: Yan, Weihao, et al.
Published: (2024)
Not All Samples Should Be Utilized Equally: Towards Understanding and Improving Dataset Distillation
by: Wang, Shaobo, et al.
Published: (2024)
by: Wang, Shaobo, et al.
Published: (2024)
RoboSense: Large-scale Dataset and Benchmark for Egocentric Robot Perception and Navigation in Crowded and Unstructured Environments
by: Su, Haisheng, et al.
Published: (2024)
by: Su, Haisheng, et al.
Published: (2024)
ViTree: Single-path Neural Tree for Step-wise Interpretable Fine-grained Visual Categorization
by: Lao, Danning, et al.
Published: (2024)
by: Lao, Danning, et al.
Published: (2024)
PostEdit: Posterior Sampling for Efficient Zero-Shot Image Editing
by: Tian, Feng, et al.
Published: (2024)
by: Tian, Feng, et al.
Published: (2024)
SUPClust: Active Learning at the Boundaries
by: Ono, Yuta, et al.
Published: (2024)
by: Ono, Yuta, et al.
Published: (2024)
Stabilizing Diffusion Posterior Sampling by Noise--Frequency Continuation
by: Tian, Feng, et al.
Published: (2026)
by: Tian, Feng, et al.
Published: (2026)
Raw2Drive: Reinforcement Learning with Aligned World Models for End-to-End Autonomous Driving (in CARLA v2)
by: Yang, Zhenjie, et al.
Published: (2025)
by: Yang, Zhenjie, et al.
Published: (2025)
Active Multimodal Distillation for Few-shot Action Recognition
by: Feng, Weijia, et al.
Published: (2025)
by: Feng, Weijia, et al.
Published: (2025)
Point2RBox: Combine Knowledge from Synthetic Visual Patterns for End-to-end Oriented Object Detection with Single Point Supervision
by: Yu, Yi, et al.
Published: (2023)
by: Yu, Yi, et al.
Published: (2023)
CAFe: Unifying Representation and Generation with Contrastive-Autoregressive Finetuning
by: Yu, Hao, et al.
Published: (2025)
by: Yu, Hao, et al.
Published: (2025)
Calibrating Biased Distribution in VFM-derived Latent Space via Cross-Domain Geometric Consistency
by: Ma, Yanbiao, et al.
Published: (2025)
by: Ma, Yanbiao, et al.
Published: (2025)
Theoretically Achieving Continuous Representation of Oriented Bounding Boxes
by: Xiao, Zi-Kai, et al.
Published: (2024)
by: Xiao, Zi-Kai, et al.
Published: (2024)
Open-Vocabulary Remote Sensing Image Semantic Segmentation
by: Cao, Qinglong, et al.
Published: (2024)
by: Cao, Qinglong, et al.
Published: (2024)
Tracing Copied Pixels and Regularizing Patch Affinity in Copy Detection
by: Lu, Yichen, et al.
Published: (2026)
by: Lu, Yichen, et al.
Published: (2026)
Latent Intuitive Physics: Learning to Transfer Hidden Physics from A 3D Video
by: Zhu, Xiangming, et al.
Published: (2024)
by: Zhu, Xiangming, et al.
Published: (2024)
ALF: Adaptive Label Finetuning for Scene Graph Generation
by: Chen, Qishen, et al.
Published: (2023)
by: Chen, Qishen, et al.
Published: (2023)
Prompt-Consistency Image Generation (PCIG): A Unified Framework Integrating LLMs, Knowledge Graphs, and Controllable Diffusion Models
by: Sun, Yichen, et al.
Published: (2024)
by: Sun, Yichen, et al.
Published: (2024)
Learning to Decode Against Compositional Hallucination in Video Multimodal Large Language Models
by: Xing, Wenbin, et al.
Published: (2026)
by: Xing, Wenbin, et al.
Published: (2026)
SparseFormer: Detecting Objects in HRW Shots via Sparse Vision Transformer
by: Li, Wenxi, et al.
Published: (2025)
by: Li, Wenxi, et al.
Published: (2025)
Infusion: Preventing Customized Text-to-Image Diffusion from Overfitting
by: Zeng, Weili, et al.
Published: (2024)
by: Zeng, Weili, et al.
Published: (2024)
Decision Boundary-aware Generation for Long-tailed Learning
by: Yang, Jiacheng, et al.
Published: (2026)
by: Yang, Jiacheng, et al.
Published: (2026)
Guidance Matters: Rethinking the Evaluation Pitfall for Text-to-Image Generation
by: Xie, Dian, et al.
Published: (2026)
by: Xie, Dian, et al.
Published: (2026)
Directing the Narrative: A Finetuning Method for Controlling Coherence and Style in Story Generation
by: Zhang, Jianzhang, et al.
Published: (2026)
by: Zhang, Jianzhang, et al.
Published: (2026)
A Hybrid Co-Finetuning Approach for Visual Bug Detection in Video Games
by: Yi, Faliu, et al.
Published: (2025)
by: Yi, Faliu, et al.
Published: (2025)
Overcoming the Pitfalls of Vision-Language Model Finetuning for OOD Generalization
by: Zang, Yuhang, et al.
Published: (2024)
by: Zang, Yuhang, et al.
Published: (2024)
Decouple-Then-Merge: Finetune Diffusion Models as Multi-Task Learning
by: Ma, Qianli, et al.
Published: (2024)
by: Ma, Qianli, et al.
Published: (2024)
Similar Items
-
ActiveAD: Planning-Oriented Active Learning for End-to-End Autonomous Driving
by: Lu, Han, et al.
Published: (2024) -
Rethinking Classifier Re-Training in Long-Tailed Recognition: A Simple Logits Retargeting Approach
by: Lu, Han, et al.
Published: (2024) -
Unified Batch Normalization: Identifying and Alleviating the Feature Condensation in Batch Normalization and a Unified Framework
by: Wang, Shaobo, et al.
Published: (2023) -
PointOBB-v3: Expanding Performance Boundaries of Single Point-Supervised Oriented Object Detection
by: Zhang, Peiyuan, et al.
Published: (2025) -
MotionGPT: Finetuned LLMs Are General-Purpose Motion Generators
by: Zhang, Yaqi, et al.
Published: (2023)