Strong but simple: A Baseline for Domain Generalized Dense Perception by CLIP-based Transfer Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Hümmer, Christoph, Schwonberg, Manuel, Zhou, Liangwei, Cao, Hu, Knoll, Alois, Gottschalk, Hanno |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Domain Generalization for Semantic Segmentation: A Survey
por: Schwonberg, Manuel, et al.
Publicado: (2025)
por: Schwonberg, Manuel, et al.
Publicado: (2025)
A Study on Unsupervised Domain Adaptation for Semantic Segmentation in the Era of Vision-Language Models
por: Schwonberg, Manuel, et al.
Publicado: (2024)
por: Schwonberg, Manuel, et al.
Publicado: (2024)
TUMTraf EMOT: Event-Based Multi-Object Tracking Dataset and Baseline for Traffic Scenarios
por: Li, Mengyu, et al.
Publicado: (2025)
por: Li, Mengyu, et al.
Publicado: (2025)
URNet: Uncertainty-aware Refinement Network for Event-based Stereo Depth Estimation
por: Cheng, Yifeng, et al.
Publicado: (2025)
por: Cheng, Yifeng, et al.
Publicado: (2025)
How Could Generative AI Support Compliance with the EU AI Act? A Review for Safe Automated Driving Perception
por: Keser, Mert, et al.
Publicado: (2024)
por: Keser, Mert, et al.
Publicado: (2024)
CLIP with Generative Latent Replay: a Strong Baseline for Incremental Learning
por: Frascaroli, Emanuele, et al.
Publicado: (2024)
por: Frascaroli, Emanuele, et al.
Publicado: (2024)
DeCLIP: Decoupled Learning for Open-Vocabulary Dense Perception
por: Wang, Junjie, et al.
Publicado: (2025)
por: Wang, Junjie, et al.
Publicado: (2025)
BiSeg-SAM: Weakly-Supervised Post-Processing Framework for Boosting Binary Segmentation in Segment Anything Models
por: Su, Encheng, et al.
Publicado: (2025)
por: Su, Encheng, et al.
Publicado: (2025)
GPT-4V as Traffic Assistant: An In-depth Look at Vision Language Model on Complex Traffic Events
por: Zhou, Xingcheng, et al.
Publicado: (2024)
por: Zhou, Xingcheng, et al.
Publicado: (2024)
Language-Guided Grasp Detection with Coarse-to-Fine Learning for Robotic Manipulation
por: Jiang, Zebin, et al.
Publicado: (2025)
por: Jiang, Zebin, et al.
Publicado: (2025)
SegRGB-X: General RGB-X Semantic Segmentation Model
por: Liu, Jiong, et al.
Publicado: (2026)
por: Liu, Jiong, et al.
Publicado: (2026)
LFA: Layer Feature Attention for Run-Time Introspection of 2D Object Detectors in Automated Driving
por: Keser, Mert, et al.
Publicado: (2026)
por: Keser, Mert, et al.
Publicado: (2026)
Learning Transient Convective Heat Transfer with Geometry Aware World Models
por: Doganay, Onur T., et al.
Publicado: (2026)
por: Doganay, Onur T., et al.
Publicado: (2026)
Exploring Weak-to-Strong Generalization for CLIP-based Classification
por: Li, Jinhao, et al.
Publicado: (2025)
por: Li, Jinhao, et al.
Publicado: (2025)
TUMTraf V2X Cooperative Perception Dataset
por: Zimmer, Walter, et al.
Publicado: (2024)
por: Zimmer, Walter, et al.
Publicado: (2024)
AffordanceGrasp-R1:Leveraging Reasoning-Based Affordance Segmentation with Reinforcement Learning for Robotic Grasping
por: Zhou, Dingyi, et al.
Publicado: (2026)
por: Zhou, Dingyi, et al.
Publicado: (2026)
ControlUDA: Controllable Diffusion-assisted Unsupervised Domain Adaptation for Cross-Weather Semantic Segmentation
por: Shen, Fengyi, et al.
Publicado: (2024)
por: Shen, Fengyi, et al.
Publicado: (2024)
Control-CLIP: Decoupling Category and Style Guidance in CLIP for Specific-Domain Generation
por: Jia, Zexi, et al.
Publicado: (2025)
por: Jia, Zexi, et al.
Publicado: (2025)
Harmonizing and Merging Source Models for CLIP-based Domain Generalization
por: Ding, Yuhe, et al.
Publicado: (2025)
por: Ding, Yuhe, et al.
Publicado: (2025)
VideoGAN-based Trajectory Proposal for Automated Vehicles
por: Mariani, Annajoyce, et al.
Publicado: (2025)
por: Mariani, Annajoyce, et al.
Publicado: (2025)
Fast-BEV: A Fast and Strong Bird's-Eye View Perception Baseline
por: Li, Yangguang, et al.
Publicado: (2023)
por: Li, Yangguang, et al.
Publicado: (2023)
Efficient and Deterministic Search Strategy Based on Residual Projections for Point Cloud Registration with Correspondences
por: Li, Xinyi, et al.
Publicado: (2023)
por: Li, Xinyi, et al.
Publicado: (2023)
ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference
por: Lan, Mengcheng, et al.
Publicado: (2024)
por: Lan, Mengcheng, et al.
Publicado: (2024)
Energy-Aware Imitation Learning for Steering Prediction Using Events and Frames
por: Cao, Hu, et al.
Publicado: (2026)
por: Cao, Hu, et al.
Publicado: (2026)
Towards Point Cloud Compression for Machine Perception: A Simple and Strong Baseline by Learning the Octree Depth Level Predictor
por: Liu, Lei, et al.
Publicado: (2024)
por: Liu, Lei, et al.
Publicado: (2024)
Rethinking Domain Adaptation and Generalization in the Era of CLIP
por: Feng, Ruoyu, et al.
Publicado: (2024)
por: Feng, Ruoyu, et al.
Publicado: (2024)
Learning Brenier Potentials with Convex Generative Adversarial Neural Networks
por: Drygala, Claudia, et al.
Publicado: (2025)
por: Drygala, Claudia, et al.
Publicado: (2025)
LEAD: Learning Decomposition for Source-free Universal Domain Adaptation
por: Qu, Sanqing, et al.
Publicado: (2024)
por: Qu, Sanqing, et al.
Publicado: (2024)
CCTVBench: Contrastive Consistency Traffic VideoQA Benchmark for Multimodal LLMs
por: Zhou, Xingcheng, et al.
Publicado: (2026)
por: Zhou, Xingcheng, et al.
Publicado: (2026)
Building a Strong Pre-Training Baseline for Universal 3D Large-Scale Perception
por: Chen, Haoming, et al.
Publicado: (2024)
por: Chen, Haoming, et al.
Publicado: (2024)
Adaptive Neural Networks for Intelligent Data-Driven Development
por: Shoeb, Youssef, et al.
Publicado: (2025)
por: Shoeb, Youssef, et al.
Publicado: (2025)
WARM-3D: A Weakly-Supervised Sim2Real Domain Adaptation Framework for Roadside Monocular 3D Object Detection
por: Zhou, Xingcheng, et al.
Publicado: (2024)
por: Zhou, Xingcheng, et al.
Publicado: (2024)
SGTA: Scene-Graph Based Multi-Modal Traffic Agent for Video Understanding
por: Zhou, Xingcheng, et al.
Publicado: (2026)
por: Zhou, Xingcheng, et al.
Publicado: (2026)
PCDepth: Pattern-based Complementary Learning for Monocular Depth Estimation by Best of Both Worlds
por: Liu, Haotian, et al.
Publicado: (2024)
por: Liu, Haotian, et al.
Publicado: (2024)
ERM++: An Improved Baseline for Domain Generalization
por: Teterwak, Piotr, et al.
Publicado: (2023)
por: Teterwak, Piotr, et al.
Publicado: (2023)
DiCLIP: Diffusion Model Enhances CLIP's Dense Knowledge for Weakly Supervised Semantic Segmentation
por: Yang, Zhiwei, et al.
Publicado: (2026)
por: Yang, Zhiwei, et al.
Publicado: (2026)
Phrase Grounding-based Style Transfer for Single-Domain Generalized Object Detection
por: Li, Hao, et al.
Publicado: (2024)
por: Li, Hao, et al.
Publicado: (2024)
Generalized Decoupled Learning for Enhancing Open-Vocabulary Dense Perception
por: Wang, Junjie, et al.
Publicado: (2025)
por: Wang, Junjie, et al.
Publicado: (2025)
A Hard-to-Beat Baseline for Training-free CLIP-based Adaptation
por: Wang, Zhengbo, et al.
Publicado: (2024)
por: Wang, Zhengbo, et al.
Publicado: (2024)
DeCLIP: Decoupled Prompting for CLIP-based Multi-Label Class-Incremental Learning
por: Du, Kaile, et al.
Publicado: (2025)
por: Du, Kaile, et al.
Publicado: (2025)
Ejemplares similares
-
Domain Generalization for Semantic Segmentation: A Survey
por: Schwonberg, Manuel, et al.
Publicado: (2025) -
A Study on Unsupervised Domain Adaptation for Semantic Segmentation in the Era of Vision-Language Models
por: Schwonberg, Manuel, et al.
Publicado: (2024) -
TUMTraf EMOT: Event-Based Multi-Object Tracking Dataset and Baseline for Traffic Scenarios
por: Li, Mengyu, et al.
Publicado: (2025) -
URNet: Uncertainty-aware Refinement Network for Event-based Stereo Depth Estimation
por: Cheng, Yifeng, et al.
Publicado: (2025) -
How Could Generative AI Support Compliance with the EU AI Act? A Review for Safe Automated Driving Perception
por: Keser, Mert, et al.
Publicado: (2024)