Image-Specific Adaptation of Transformer Encoders for Compute-Efficient Segmentation
Fuente:
arXiv
Salvato in:
| Autori principali: | Yao, Manyi, Aich, Abhishek, Suh, Yumin, Roy-Chowdhury, Amit, Shelton, Christian, Chandraker, Manmohan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Progressive Token Length Scaling in Transformer Encoders for Efficient Universal Segmentation
di: Aich, Abhishek, et al.
Pubblicazione: (2024)
di: Aich, Abhishek, et al.
Pubblicazione: (2024)
iFinder: Structured Zero-Shot Vision-Based LLM Grounding for Dash-Cam Video Reasoning
di: Yao, Manyi, et al.
Pubblicazione: (2025)
di: Yao, Manyi, et al.
Pubblicazione: (2025)
UDA-Bench: Revisiting Common Assumptions in Unsupervised Domain Adaptation Using a Standardized Framework
di: Kalluri, Tarun, et al.
Pubblicazione: (2024)
di: Kalluri, Tarun, et al.
Pubblicazione: (2024)
What to Test Next: Interpretable Coverage Gap Discovery in Driving VLMs
di: Aich, Abhishek, et al.
Pubblicazione: (2026)
di: Aich, Abhishek, et al.
Pubblicazione: (2026)
Tuned Contrastive Learning
di: Animesh, Chaitanya, et al.
Pubblicazione: (2023)
di: Animesh, Chaitanya, et al.
Pubblicazione: (2023)
Locally Orderless Images for Optimization in Differentiable Rendering
di: Mehta, Ishit, et al.
Pubblicazione: (2025)
di: Mehta, Ishit, et al.
Pubblicazione: (2025)
HorizonWeaver: Generalizable Multi-Level Semantic Editing for Driving Scenes
di: Soroco, Mauricio, et al.
Pubblicazione: (2026)
di: Soroco, Mauricio, et al.
Pubblicazione: (2026)
ST-VLM: Kinematic Instruction Tuning for Spatio-Temporal Reasoning in Vision-Language Models
di: Ko, Dohwan, et al.
Pubblicazione: (2025)
di: Ko, Dohwan, et al.
Pubblicazione: (2025)
Generating Enhanced Negatives for Training Language-Based Object Detectors
di: Zhao, Shiyu, et al.
Pubblicazione: (2023)
di: Zhao, Shiyu, et al.
Pubblicazione: (2023)
PhyCo: Learning Controllable Physical Priors for Generative Motion
di: Narayanan, Sriram, et al.
Pubblicazione: (2026)
di: Narayanan, Sriram, et al.
Pubblicazione: (2026)
Open-world Instance Segmentation: Top-down Learning with Bottom-up Supervision
di: Kalluri, Tarun, et al.
Pubblicazione: (2023)
di: Kalluri, Tarun, et al.
Pubblicazione: (2023)
RAD-LAD: Rule and Language Grounded Autonomous Driving in Real-Time
di: Ghosh, Anurag, et al.
Pubblicazione: (2026)
di: Ghosh, Anurag, et al.
Pubblicazione: (2026)
Taming Self-Training for Open-Vocabulary Object Detection
di: Zhao, Shiyu, et al.
Pubblicazione: (2023)
di: Zhao, Shiyu, et al.
Pubblicazione: (2023)
Tell, Don't Show!: Language Guidance Eases Transfer Across Domains in Images and Videos
di: Kalluri, Tarun, et al.
Pubblicazione: (2024)
di: Kalluri, Tarun, et al.
Pubblicazione: (2024)
Robust Disaster Assessment from Aerial Imagery Using Text-to-Image Synthetic Data
di: Kalluri, Tarun, et al.
Pubblicazione: (2024)
di: Kalluri, Tarun, et al.
Pubblicazione: (2024)
Rolling Sink: Bridging Limited-Horizon Training and Open-Ended Testing in Autoregressive Video Diffusion
di: Li, Haodong, et al.
Pubblicazione: (2026)
di: Li, Haodong, et al.
Pubblicazione: (2026)
Resolving Inconsistent Semantics in Multi-Dataset Image Segmentation
di: Zhangli, Qilong, et al.
Pubblicazione: (2024)
di: Zhangli, Qilong, et al.
Pubblicazione: (2024)
Instantaneous Perception of Moving Objects in 3D
di: Liu, Di, et al.
Pubblicazione: (2024)
di: Liu, Di, et al.
Pubblicazione: (2024)
ODES: Domain Adaptation with Expert Guidance for Online Medical Image Segmentation
di: Islam, Md Shazid, et al.
Pubblicazione: (2023)
di: Islam, Md Shazid, et al.
Pubblicazione: (2023)
Mapillary Vistas Validation for Fine-Grained Traffic Signs: A Benchmark Revealing Vision-Language Model Limitations
di: Garg, Sparsh, et al.
Pubblicazione: (2025)
di: Garg, Sparsh, et al.
Pubblicazione: (2025)
Efficient Parameter Adaptation for Multi-Modal Medical Image Segmentation and Prognosis
di: Saeed, Numan, et al.
Pubblicazione: (2025)
di: Saeed, Numan, et al.
Pubblicazione: (2025)
AIDE: An Automatic Data Engine for Object Detection in Autonomous Driving
di: Liang, Mingfu, et al.
Pubblicazione: (2024)
di: Liang, Mingfu, et al.
Pubblicazione: (2024)
Self-Training Large Language Models for Improved Visual Program Synthesis With Visual Reinforcement
di: Khan, Zaid, et al.
Pubblicazione: (2024)
di: Khan, Zaid, et al.
Pubblicazione: (2024)
LINGUAL: Language-INtegrated GUidance in Active Learning for Medical Image Segmentation
di: Islam, Md Shazid, et al.
Pubblicazione: (2025)
di: Islam, Md Shazid, et al.
Pubblicazione: (2025)
LLM-Assist: Enhancing Closed-Loop Planning with Language-Based Reasoning
di: Sharan, S P, et al.
Pubblicazione: (2023)
di: Sharan, S P, et al.
Pubblicazione: (2023)
LangDriveCTRL: Natural Language Controllable Driving Scene Editing with Multi-modal Agents
di: He, Yun, et al.
Pubblicazione: (2025)
di: He, Yun, et al.
Pubblicazione: (2025)
LidaRF: Delving into Lidar for Neural Radiance Field on Street Scenes
di: Sun, Shanlin, et al.
Pubblicazione: (2024)
di: Sun, Shanlin, et al.
Pubblicazione: (2024)
HorizonForge: Driving Scene Editing with Any Trajectories and Any Vehicles
di: Wang, Yifan, et al.
Pubblicazione: (2026)
di: Wang, Yifan, et al.
Pubblicazione: (2026)
Drive-1-to-3: Enriching Diffusion Priors for Novel View Synthesis of Real Vehicles
di: Lin, Chuang, et al.
Pubblicazione: (2024)
di: Lin, Chuang, et al.
Pubblicazione: (2024)
Learning from Unlabelled Data with Transformers: Domain Adaptation for Semantic Segmentation of High Resolution Aerial Images
di: Dionelis, Nikolaos, et al.
Pubblicazione: (2024)
di: Dionelis, Nikolaos, et al.
Pubblicazione: (2024)
From Segments to Concepts: Interpretable Image Classification via Concept-Guided Segmentation
di: Eisenberg, Ran, et al.
Pubblicazione: (2025)
di: Eisenberg, Ran, et al.
Pubblicazione: (2025)
POSTURE: Pose Guided Unsupervised Domain Adaptation for Human Body Part Segmentation
di: Dutta, Arindam, et al.
Pubblicazione: (2024)
di: Dutta, Arindam, et al.
Pubblicazione: (2024)
What You See is What You GAN: Rendering Every Pixel for High-Fidelity Geometry in 3D GANs
di: Trevithick, Alex, et al.
Pubblicazione: (2024)
di: Trevithick, Alex, et al.
Pubblicazione: (2024)
CART: Compositional Auto-Regressive Transformer for Image Generation
di: Roheda, Siddharth, et al.
Pubblicazione: (2024)
di: Roheda, Siddharth, et al.
Pubblicazione: (2024)
Repurposing SAM for User-Defined Semantics Aware Segmentation
di: Kundu, Rohit, et al.
Pubblicazione: (2023)
di: Kundu, Rohit, et al.
Pubblicazione: (2023)
SAFE-SIM: Safety-Critical Closed-Loop Traffic Simulation with Diffusion-Controllable Adversaries
di: Chang, Wei-Jer, et al.
Pubblicazione: (2023)
di: Chang, Wei-Jer, et al.
Pubblicazione: (2023)
Efficient Adaptation of Deep Neural Networks for Semantic Segmentation in Space Applications
di: Olivi, Leonardo, et al.
Pubblicazione: (2025)
di: Olivi, Leonardo, et al.
Pubblicazione: (2025)
Distilled Pooling Transformer Encoder for Efficient Realistic Image Dehazing
di: Tran, Le-Anh, et al.
Pubblicazione: (2024)
di: Tran, Le-Anh, et al.
Pubblicazione: (2024)
Broadband Wide Field of View Imaging with Computational Mirrors
di: Saragadam, Vishwanath, et al.
Pubblicazione: (2026)
di: Saragadam, Vishwanath, et al.
Pubblicazione: (2026)
PEMMA: Parameter-Efficient Multi-Modal Adaptation for Medical Image Segmentation
di: Saadi, Nada, et al.
Pubblicazione: (2024)
di: Saadi, Nada, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Progressive Token Length Scaling in Transformer Encoders for Efficient Universal Segmentation
di: Aich, Abhishek, et al.
Pubblicazione: (2024) -
iFinder: Structured Zero-Shot Vision-Based LLM Grounding for Dash-Cam Video Reasoning
di: Yao, Manyi, et al.
Pubblicazione: (2025) -
UDA-Bench: Revisiting Common Assumptions in Unsupervised Domain Adaptation Using a Standardized Framework
di: Kalluri, Tarun, et al.
Pubblicazione: (2024) -
What to Test Next: Interpretable Coverage Gap Discovery in Driving VLMs
di: Aich, Abhishek, et al.
Pubblicazione: (2026) -
Tuned Contrastive Learning
di: Animesh, Chaitanya, et al.
Pubblicazione: (2023)