DenVisCoM: Dense Vision Correspondence Mamba for Efficient and Real-time Optical Flow and Stereo Estimation
Fuente:
arXiv
Guardado en:
| Autores principales: | Anand, Tushar, Bora, Maheswar, Dantcheva, Antitza, Das, Abhijit |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DensePercept-NCSSD: Vision Mamba towards Real-time Dense Visual Perception with Non-Causal State Space Duality
por: Anand, Tushar, et al.
Publicado: (2025)
por: Anand, Tushar, et al.
Publicado: (2025)
Do You See What I Say? Generalizable Deepfake Detection based on Visual Speech Recognition
por: Bora, Maheswar, et al.
Publicado: (2025)
por: Bora, Maheswar, et al.
Publicado: (2025)
ViM-Disparity: Bridging the Gap of Speed, Accuracy and Memory for Disparity Map Generation
por: Bora, Maheswar, et al.
Publicado: (2024)
por: Bora, Maheswar, et al.
Publicado: (2024)
Beyond Real versus Fake Towards Intent-Aware Video Analysis
por: Atreya, Saurabh, et al.
Publicado: (2025)
por: Atreya, Saurabh, et al.
Publicado: (2025)
AM Flow: Adapters for Temporal Processing in Action Recognition
por: Agrawal, Tanay, et al.
Publicado: (2024)
por: Agrawal, Tanay, et al.
Publicado: (2024)
KDC-MAE: Knowledge Distilled Contrastive Mask Auto-Encoder
por: Bora, Maheswar, et al.
Publicado: (2024)
por: Bora, Maheswar, et al.
Publicado: (2024)
Enhancing 3D-Air Signature by Pen Tip Tail Trajectory Awareness: Dataset and Featuring by Novel Spatio-temporal CNN
por: Atreya, Saurabh, et al.
Publicado: (2024)
por: Atreya, Saurabh, et al.
Publicado: (2024)
Now You See Me, Now You Don't: A Unified Framework for Expression Consistent Anonymization in Talking Head Videos
por: Egin, Anil, et al.
Publicado: (2026)
por: Egin, Anil, et al.
Publicado: (2026)
Beyond the Visible: A Survey on Cross-spectral Face Recognition
por: Anghelone, David, et al.
Publicado: (2022)
por: Anghelone, David, et al.
Publicado: (2022)
THEval. Evaluation Framework for Talking Head Video Generation
por: Quignon, Nabyl, et al.
Publicado: (2025)
por: Quignon, Nabyl, et al.
Publicado: (2025)
HFNeRF: Learning Human Biomechanic Features with Neural Radiance Fields
por: Dey, Arnab, et al.
Publicado: (2024)
por: Dey, Arnab, et al.
Publicado: (2024)
AI killed the video star. Audio-driven diffusion model for expressive talking head generation
por: Chopin, Baptiste, et al.
Publicado: (2025)
por: Chopin, Baptiste, et al.
Publicado: (2025)
Dimitra: Audio-driven Diffusion model for Expressive Talking Head Generation
por: Chopin, Baptiste, et al.
Publicado: (2025)
por: Chopin, Baptiste, et al.
Publicado: (2025)
StereoMamba: Real-time and Robust Intraoperative Stereo Disparity Estimation via Long-range Spatial Dependencies
por: Wang, Xu, et al.
Publicado: (2025)
por: Wang, Xu, et al.
Publicado: (2025)
LIA-X: Interpretable Latent Portrait Animator
por: Wang, Yaohui, et al.
Publicado: (2025)
por: Wang, Yaohui, et al.
Publicado: (2025)
GHNeRF: Learning Generalizable Human Features with Efficient Neural Radiance Fields
por: Dey, Arnab, et al.
Publicado: (2024)
por: Dey, Arnab, et al.
Publicado: (2024)
CompactFlowNet: Efficient Real-time Optical Flow Estimation on Mobile Devices
por: Znobishchev, Andrei, et al.
Publicado: (2024)
por: Znobishchev, Andrei, et al.
Publicado: (2024)
Affine Correspondences in Stereo Vision: Theory, Practice, and Limitations
por: Hajder, Levente
Publicado: (2026)
por: Hajder, Levente
Publicado: (2026)
MambaFlow: A Mamba-Centric Architecture for End-to-End Optical Flow Estimation
por: Du, Juntian, et al.
Publicado: (2025)
por: Du, Juntian, et al.
Publicado: (2025)
LEO: Generative Latent Image Animator for Human Video Synthesis
por: Wang, Yaohui, et al.
Publicado: (2023)
por: Wang, Yaohui, et al.
Publicado: (2023)
LAC: Latent Action Composition for Skeleton-based Action Segmentation
por: Yang, Di, et al.
Publicado: (2023)
por: Yang, Di, et al.
Publicado: (2023)
DenSe-AdViT: A novel Vision Transformer for Dense SAR Object Detection
por: Zhang, Yang, et al.
Publicado: (2025)
por: Zhang, Yang, et al.
Publicado: (2025)
TrackingMiM: Efficient Mamba-in-Mamba Serialization for Real-time UAV Object Tracking
por: Liu, Bingxi, et al.
Publicado: (2025)
por: Liu, Bingxi, et al.
Publicado: (2025)
VisCoP: Visual Probing for Video Domain Adaptation of Vision Language Models
por: Reilly, Dominick, et al.
Publicado: (2025)
por: Reilly, Dominick, et al.
Publicado: (2025)
OuroMamba: A Data-Free Quantization Framework for Vision Mamba
por: Ramachandran, Akshat, et al.
Publicado: (2025)
por: Ramachandran, Akshat, et al.
Publicado: (2025)
Improving Optical Flow and Stereo Depth Estimation by Leveraging Uncertainty-Based Learning Difficulties
por: Jeong, Jisoo, et al.
Publicado: (2025)
por: Jeong, Jisoo, et al.
Publicado: (2025)
Self-Assessed Generation: Trustworthy Label Generation for Optical Flow and Stereo Matching in Real-world
por: Ling, Han, et al.
Publicado: (2024)
por: Ling, Han, et al.
Publicado: (2024)
VisFocus: Prompt-Guided Vision Encoders for OCR-Free Dense Document Understanding
por: Abramovich, Ofir, et al.
Publicado: (2024)
por: Abramovich, Ofir, et al.
Publicado: (2024)
Just Dance with $π$! A Poly-modal Inductor for Weakly-supervised Video Anomaly Detection
por: Majhi, Snehashis, et al.
Publicado: (2025)
por: Majhi, Snehashis, et al.
Publicado: (2025)
MonSter++: Unified Stereo Matching, Multi-view Stereo, and Real-time Stereo with Monodepth Priors
por: Cheng, Junda, et al.
Publicado: (2025)
por: Cheng, Junda, et al.
Publicado: (2025)
DenTab: A Dataset for Table Recognition and Visual QA on Real-World Dental Estimates
por: Hamdi, Laziz, et al.
Publicado: (2026)
por: Hamdi, Laziz, et al.
Publicado: (2026)
EDCFlow: Exploring Temporally Dense Difference Maps for Event-based Optical Flow Estimation
por: Liu, Daikun, et al.
Publicado: (2025)
por: Liu, Daikun, et al.
Publicado: (2025)
UFM: A Simple Path towards Unified Dense Correspondence with Flow
por: Zhang, Yuchen, et al.
Publicado: (2025)
por: Zhang, Yuchen, et al.
Publicado: (2025)
UltrasODM: A Dual Stream Optical Flow Mamba Network for 3D Freehand Ultrasound Reconstruction
por: Anand, Mayank, et al.
Publicado: (2025)
por: Anand, Mayank, et al.
Publicado: (2025)
CoMamba: Real-time Cooperative Perception Unlocked with State Space Models
por: Li, Jinlong, et al.
Publicado: (2024)
por: Li, Jinlong, et al.
Publicado: (2024)
Robust and Real-time Surface Normal Estimation from Stereo Disparities using Affine Transformations
por: Kariko, Csongor Csanad, et al.
Publicado: (2025)
por: Kariko, Csongor Csanad, et al.
Publicado: (2025)
NeuFlow: Real-time, High-accuracy Optical Flow Estimation on Robots Using Edge Devices
por: Zhang, Zhiyong, et al.
Publicado: (2024)
por: Zhang, Zhiyong, et al.
Publicado: (2024)
Den-SOFT: Dense Space-Oriented Light Field DataseT for 6-DOF Immersive Experience
por: Yu, Xiaohang, et al.
Publicado: (2024)
por: Yu, Xiaohang, et al.
Publicado: (2024)
RobustSpring: Benchmarking Robustness to Image Corruptions for Optical Flow, Scene Flow and Stereo
por: Oei, Victor, et al.
Publicado: (2025)
por: Oei, Victor, et al.
Publicado: (2025)
Foundation AI Models for Aerosol Optical Depth Estimation from PACE Satellite Data
por: Tushar, Zahid Hassan, et al.
Publicado: (2026)
por: Tushar, Zahid Hassan, et al.
Publicado: (2026)
Ejemplares similares
-
DensePercept-NCSSD: Vision Mamba towards Real-time Dense Visual Perception with Non-Causal State Space Duality
por: Anand, Tushar, et al.
Publicado: (2025) -
Do You See What I Say? Generalizable Deepfake Detection based on Visual Speech Recognition
por: Bora, Maheswar, et al.
Publicado: (2025) -
ViM-Disparity: Bridging the Gap of Speed, Accuracy and Memory for Disparity Map Generation
por: Bora, Maheswar, et al.
Publicado: (2024) -
Beyond Real versus Fake Towards Intent-Aware Video Analysis
por: Atreya, Saurabh, et al.
Publicado: (2025) -
AM Flow: Adapters for Temporal Processing in Action Recognition
por: Agrawal, Tanay, et al.
Publicado: (2024)