Breaking the Resource Wall: Geometry-Guided Sequence Modeling for Efficient Semantic Segmentation
Fuente:
arXiv
Salvato in:
| Autori principali: | Chan, Sheng-Wei, Pan, Hsin-Jui, Shen, Chun-Po, Lin, Chia-Min, Wang, Yung-Che, Chiang, Jen-Shiun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FoR-Net: Learning to Focus on Hard Regions for Efficient Semantic Segmentation
di: Chan, Sheng-Wei, et al.
Pubblicazione: (2026)
di: Chan, Sheng-Wei, et al.
Pubblicazione: (2026)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
di: Raoufi, Behnam, et al.
Pubblicazione: (2025)
di: Raoufi, Behnam, et al.
Pubblicazione: (2025)
Conterfactual Generative Zero-Shot Semantic Segmentation
di: Shen, Feihong, et al.
Pubblicazione: (2021)
di: Shen, Feihong, et al.
Pubblicazione: (2021)
CoMoCAVs: Cohesive Decision-Guided Motion Planning for Connected and Autonomous Vehicles with Multi-Policy Reinforcement Learning
di: Hu, Pan
Pubblicazione: (2025)
di: Hu, Pan
Pubblicazione: (2025)
Evaluation of Attention Mechanisms in U-Net Architectures for Semantic Segmentation of Brazilian Rock Art Petroglyphs
di: Melo, Leonardi, et al.
Pubblicazione: (2025)
di: Melo, Leonardi, et al.
Pubblicazione: (2025)
Distinguishing Visually Similar Actions: Prompt-Guided Semantic Prototype Modulation for Few-Shot Action Recognition
di: Li, Xiaoyang, et al.
Pubblicazione: (2025)
di: Li, Xiaoyang, et al.
Pubblicazione: (2025)
Motion-Guided Semantic Alignment with Negative Prompts for Zero-Shot Video Action Recognition
di: Wang, Yiming, et al.
Pubblicazione: (2026)
di: Wang, Yiming, et al.
Pubblicazione: (2026)
Textual and Visual Guided Task Adaptation for Source-Free Cross-Domain Few-Shot Segmentation
di: Liu, Jianming, et al.
Pubblicazione: (2025)
di: Liu, Jianming, et al.
Pubblicazione: (2025)
Multi-Agent VLMs Guided Self-Training with PNU Loss for Low-Resource Offensive Content Detection
di: Wang, Han, et al.
Pubblicazione: (2025)
di: Wang, Han, et al.
Pubblicazione: (2025)
Break the Brake, Not the Wheel: Untargeted Jailbreak via Entropy Maximization
di: He, Mengqi, et al.
Pubblicazione: (2026)
di: He, Mengqi, et al.
Pubblicazione: (2026)
U-Net-Like Spiking Neural Networks for Single Image Dehazing
di: Li, Huibin, et al.
Pubblicazione: (2025)
di: Li, Huibin, et al.
Pubblicazione: (2025)
Enhancing Reinforcement Learning in 3D Environments through Semantic Segmentation: A Case Study in ViZDoom
di: Huang, Jin
Pubblicazione: (2025)
di: Huang, Jin
Pubblicazione: (2025)
Vision-based Situational Graphs Exploiting Fiducial Markers for the Integration of Semantic Entities
di: Tourani, Ali, et al.
Pubblicazione: (2023)
di: Tourani, Ali, et al.
Pubblicazione: (2023)
ClustViT: Clustering-based Token Merging for Semantic Segmentation
di: Montello, Fabio, et al.
Pubblicazione: (2025)
di: Montello, Fabio, et al.
Pubblicazione: (2025)
Semantic Prioritization in Visual Counterfactual Explanations with Weighted Segmentation and Auto-Adaptive Region Selection
di: Zhang, Lintong, et al.
Pubblicazione: (2025)
di: Zhang, Lintong, et al.
Pubblicazione: (2025)
Hierarchical Image-Guided 3D Point Cloud Segmentation in Industrial Scenes via Multi-View Bayesian Fusion
di: Zhu, Yu, et al.
Pubblicazione: (2025)
di: Zhu, Yu, et al.
Pubblicazione: (2025)
FocusedAD: Character-centric Movie Audio Description
di: Ye, Xiaojun, et al.
Pubblicazione: (2025)
di: Ye, Xiaojun, et al.
Pubblicazione: (2025)
Collaborative AI Enhances Image Understanding in Materials Science
di: Yin, Ruoyan Avery, et al.
Pubblicazione: (2025)
di: Yin, Ruoyan Avery, et al.
Pubblicazione: (2025)
EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents
di: Riva, Paolo, et al.
Pubblicazione: (2026)
di: Riva, Paolo, et al.
Pubblicazione: (2026)
Foreground Focus: Enhancing Coherence and Fidelity in Camouflaged Image Generation
di: Chen, Pei-Chi, et al.
Pubblicazione: (2025)
di: Chen, Pei-Chi, et al.
Pubblicazione: (2025)
Mask-Conditioned Voxel Diffusion for Joint Geometry and Color Inpainting
di: Sumuk, Aarya
Pubblicazione: (2026)
di: Sumuk, Aarya
Pubblicazione: (2026)
A Recipe for Geometry-Aware 3D Mesh Transformers
di: Farazi, Mohammad, et al.
Pubblicazione: (2024)
di: Farazi, Mohammad, et al.
Pubblicazione: (2024)
DesertFormer: Transformer-Based Semantic Segmentation for Off-Road Desert Terrain Classification in Autonomous Navigation Systems
di: Chebolu, Yasaswini
Pubblicazione: (2026)
di: Chebolu, Yasaswini
Pubblicazione: (2026)
A Segmented Robot Grasping Perception Neural Network for Edge AI
di: Bröcheler, Casper, et al.
Pubblicazione: (2025)
di: Bröcheler, Casper, et al.
Pubblicazione: (2025)
ActAlign: Zero-Shot Fine-Grained Video Classification via Language-Guided Sequence Alignment
di: Aghdam, Amir, et al.
Pubblicazione: (2025)
di: Aghdam, Amir, et al.
Pubblicazione: (2025)
SemanticHuman-HD: High-Resolution Semantic Disentangled 3D Human Generation
di: Zheng, Peng, et al.
Pubblicazione: (2024)
di: Zheng, Peng, et al.
Pubblicazione: (2024)
ERNet: Efficient Non-Rigid Registration Network for Point Sequences
di: He, Guangzhao, et al.
Pubblicazione: (2025)
di: He, Guangzhao, et al.
Pubblicazione: (2025)
Towards Cognitive Collaborative Robots: Semantic-Level Integration and Explainable Control for Human-Centric Cooperation
di: Oh, Jaehong
Pubblicazione: (2025)
di: Oh, Jaehong
Pubblicazione: (2025)
Pixel-Level Pavement Distress Assessment Using Instance Segmentation
di: Dewick, Logan, et al.
Pubblicazione: (2026)
di: Dewick, Logan, et al.
Pubblicazione: (2026)
CoachMe: Decoding Sport Elements with a Reference-Based Coaching Instruction Generation Model
di: Yeh, Wei-Hsin, et al.
Pubblicazione: (2025)
di: Yeh, Wei-Hsin, et al.
Pubblicazione: (2025)
A Scalable Pipeline Combining Procedural 3D Graphics and Guided Diffusion for Photorealistic Synthetic Training Data Generation in White Button Mushroom Segmentation
di: Károly, Artúr I., et al.
Pubblicazione: (2025)
di: Károly, Artúr I., et al.
Pubblicazione: (2025)
Learning Association via Track-Detection Matching for Multi-Object Tracking
di: Adžemović, Momir
Pubblicazione: (2025)
di: Adžemović, Momir
Pubblicazione: (2025)
Semi-Supervised Segmentation via Embedding Matching
di: Xie, Weiyi, et al.
Pubblicazione: (2024)
di: Xie, Weiyi, et al.
Pubblicazione: (2024)
A Guide to Structureless Visual Localization
di: Panek, Vojtech, et al.
Pubblicazione: (2025)
di: Panek, Vojtech, et al.
Pubblicazione: (2025)
See-through: Single-image Layer Decomposition for Anime Characters
di: Lin, Jian, et al.
Pubblicazione: (2026)
di: Lin, Jian, et al.
Pubblicazione: (2026)
SITransformer: Shared Information-Guided Transformer for Extreme Multimodal Summarization
di: Liu, Sicheng, et al.
Pubblicazione: (2024)
di: Liu, Sicheng, et al.
Pubblicazione: (2024)
Facial Attribute Based Text Guided Face Anonymization
di: Muştu, Mustafa İzzet, et al.
Pubblicazione: (2025)
di: Muştu, Mustafa İzzet, et al.
Pubblicazione: (2025)
MSTA3D: Multi-scale Twin-attention for 3D Instance Segmentation
di: Tran, Duc Dang Trung, et al.
Pubblicazione: (2024)
di: Tran, Duc Dang Trung, et al.
Pubblicazione: (2024)
PoseRefer: Pathway-Local Parameters for Semantically Grounded Reference Resolution
di: Deichler, Anna
Pubblicazione: (2026)
di: Deichler, Anna
Pubblicazione: (2026)
Motion Perceiver: Real-Time Occupancy Forecasting for Embedded Systems
di: Ferenczi, Bryce, et al.
Pubblicazione: (2023)
di: Ferenczi, Bryce, et al.
Pubblicazione: (2023)
Documenti analoghi
-
FoR-Net: Learning to Focus on Hard Regions for Efficient Semantic Segmentation
di: Chan, Sheng-Wei, et al.
Pubblicazione: (2026) -
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
di: Raoufi, Behnam, et al.
Pubblicazione: (2025) -
Conterfactual Generative Zero-Shot Semantic Segmentation
di: Shen, Feihong, et al.
Pubblicazione: (2021) -
CoMoCAVs: Cohesive Decision-Guided Motion Planning for Connected and Autonomous Vehicles with Multi-Policy Reinforcement Learning
di: Hu, Pan
Pubblicazione: (2025) -
Evaluation of Attention Mechanisms in U-Net Architectures for Semantic Segmentation of Brazilian Rock Art Petroglyphs
di: Melo, Leonardi, et al.
Pubblicazione: (2025)