Exploring The Visual Feature Space for Multimodal Neural Decoding
Fuente:
arXiv
Guardado en:
| Autores principales: | Xia, Weihao, Oztireli, Cengiz |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Multigranular Evaluation for Brain Visual Decoding
por: Xia, Weihao, et al.
Publicado: (2025)
por: Xia, Weihao, et al.
Publicado: (2025)
UMBRAE: Unified Multimodal Brain Decoding
por: Xia, Weihao, et al.
Publicado: (2024)
por: Xia, Weihao, et al.
Publicado: (2024)
RETRO: REthinking Tactile Representation Learning with Material PriOrs
por: Xia, Weihao, et al.
Publicado: (2025)
por: Xia, Weihao, et al.
Publicado: (2025)
DREAM: Visual Decoding from Reversing Human Visual System
por: Xia, Weihao, et al.
Publicado: (2023)
por: Xia, Weihao, et al.
Publicado: (2023)
Quartet of Diffusions: Structure-Aware Point Cloud Generation through Part and Symmetry Guidance
por: Zhou, Chenliang, et al.
Publicado: (2026)
por: Zhou, Chenliang, et al.
Publicado: (2026)
FreNBRDF: A Frequency-Rectified Neural Material Representation
por: Zhou, Chenliang, et al.
Publicado: (2025)
por: Zhou, Chenliang, et al.
Publicado: (2025)
CLIP-PAE: Projection-Augmentation Embedding to Extract Relevant Features for a Disentangled, Interpretable, and Controllable Text-Guided Face Manipulation
por: Zhou, Chenliang, et al.
Publicado: (2022)
por: Zhou, Chenliang, et al.
Publicado: (2022)
SYM3D: Learning Symmetric Triplanes for Better 3D-Awareness of GANs
por: Yang, Jing, et al.
Publicado: (2024)
por: Yang, Jing, et al.
Publicado: (2024)
Physically Based Neural Bidirectional Reflectance Distribution Function
por: Zhou, Chenliang, et al.
Publicado: (2024)
por: Zhou, Chenliang, et al.
Publicado: (2024)
Gaussian Head & Shoulders: High Fidelity Neural Upper Body Avatars with Anchor Gaussian Guided Texture Warping
por: Wu, Tianhao, et al.
Publicado: (2024)
por: Wu, Tianhao, et al.
Publicado: (2024)
ARF-Plus: Controlling Perceptual Factors in Artistic Radiance Fields for 3D Scene Stylization
por: Li, Wenzhao, et al.
Publicado: (2023)
por: Li, Wenzhao, et al.
Publicado: (2023)
Twist and Compute: The Cost of Pose in 3D Generative Diffusion
por: Fogarty, Kyle, et al.
Publicado: (2025)
por: Fogarty, Kyle, et al.
Publicado: (2025)
Self-Supervised Implicit Attention Priors for Point Cloud Reconstruction
por: Fogarty, Kyle, et al.
Publicado: (2025)
por: Fogarty, Kyle, et al.
Publicado: (2025)
Decoding Visual Neural Representations by Multimodal with Dynamic Balancing
por: sun, Kaili, et al.
Publicado: (2025)
por: sun, Kaili, et al.
Publicado: (2025)
PoseCraft: Tokenized 3D Body Landmark and Camera Conditioning for Photorealistic Human Image Synthesis
por: Guo, Zhilin, et al.
Publicado: (2026)
por: Guo, Zhilin, et al.
Publicado: (2026)
Neural-MCRL: Neural Multimodal Contrastive Representation Learning for EEG-based Visual Decoding
por: Li, Yueyang, et al.
Publicado: (2024)
por: Li, Yueyang, et al.
Publicado: (2024)
$α$Surf: Implicit Surface Reconstruction for Semi-Transparent and Thin Objects with Decoupled Geometry and Opacity
por: Wu, Tianhao, et al.
Publicado: (2023)
por: Wu, Tianhao, et al.
Publicado: (2023)
FrePolad: Frequency-Rectified Point Latent Diffusion for Point Cloud Generation
por: Zhou, Chenliang, et al.
Publicado: (2023)
por: Zhou, Chenliang, et al.
Publicado: (2023)
Blue noise for diffusion models
por: Huang, Xingchang, et al.
Publicado: (2024)
por: Huang, Xingchang, et al.
Publicado: (2024)
Best Foot Forward: Robust Foot Reconstruction in-the-wild
por: Fogarty, Kyle, et al.
Publicado: (2025)
por: Fogarty, Kyle, et al.
Publicado: (2025)
Multispectral Fine-Grained Classification of Blackgrass in Wheat and Barley Crops
por: Darbyshire, Madeleine, et al.
Publicado: (2024)
por: Darbyshire, Madeleine, et al.
Publicado: (2024)
Evolutive Rendering Models
por: Zhan, Fangneng, et al.
Publicado: (2024)
por: Zhan, Fangneng, et al.
Publicado: (2024)
Mitigating Hallucination in Multimodal LLMs with Layer Contrastive Decoding
por: Tong, Bingkui, et al.
Publicado: (2025)
por: Tong, Bingkui, et al.
Publicado: (2025)
Matryoshka Gaussian Splatting
por: Guo, Zhilin, et al.
Publicado: (2026)
por: Guo, Zhilin, et al.
Publicado: (2026)
How Visual Representations Map to Language Feature Space in Multimodal LLMs
por: Venhoff, Constantin, et al.
Publicado: (2025)
por: Venhoff, Constantin, et al.
Publicado: (2025)
Multimodal Mamba: Decoder-only Multimodal State Space Model via Quadratic to Linear Distillation
por: Liao, Bencheng, et al.
Publicado: (2025)
por: Liao, Bencheng, et al.
Publicado: (2025)
Restereo: Diffusion stereo video generation and restoration
por: Huang, Xingchang, et al.
Publicado: (2025)
por: Huang, Xingchang, et al.
Publicado: (2025)
Are They the Same? Exploring Visual Correspondence Shortcomings of Multimodal LLMs
por: Zhou, Yikang, et al.
Publicado: (2025)
por: Zhou, Yikang, et al.
Publicado: (2025)
Feature-Based Dual Visual Feature Extraction Model for Compound Multimodal Emotion Recognition
por: Liu, Ran, et al.
Publicado: (2025)
por: Liu, Ran, et al.
Publicado: (2025)
VL-Mamba: Exploring State Space Models for Multimodal Learning
por: Qiao, Yanyuan, et al.
Publicado: (2024)
por: Qiao, Yanyuan, et al.
Publicado: (2024)
Fork-Merge Decoding: Enhancing Multimodal Understanding in Audio-Visual Large Language Models
por: Jung, Chaeyoung, et al.
Publicado: (2025)
por: Jung, Chaeyoung, et al.
Publicado: (2025)
Imagine while Reasoning in Space: Multimodal Visualization-of-Thought
por: Li, Chengzu, et al.
Publicado: (2025)
por: Li, Chengzu, et al.
Publicado: (2025)
Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
por: Tong, Shengbang, et al.
Publicado: (2024)
por: Tong, Shengbang, et al.
Publicado: (2024)
Feature Visualization in 3D Convolutional Neural Networks
por: Li, Chunpeng, et al.
Publicado: (2025)
por: Li, Chunpeng, et al.
Publicado: (2025)
SmokeBench: Evaluating Multimodal Large Language Models for Wildfire Smoke Detection
por: Qi, Tianye, et al.
Publicado: (2025)
por: Qi, Tianye, et al.
Publicado: (2025)
The NeRFect Match: Exploring NeRF Features for Visual Localization
por: Zhou, Qunjie, et al.
Publicado: (2024)
por: Zhou, Qunjie, et al.
Publicado: (2024)
Brain3D: EEG-to-3D Decoding of Visual Representations via Multimodal Reasoning
por: Balloni, Emanuele, et al.
Publicado: (2026)
por: Balloni, Emanuele, et al.
Publicado: (2026)
Visual Neural Decoding via Improved Visual-EEG Semantic Consistency
por: Chen, Hongzhou, et al.
Publicado: (2024)
por: Chen, Hongzhou, et al.
Publicado: (2024)
Robust Single-shot Structured Light 3D Imaging via Neural Feature Decoding
por: Li, Jiaheng, et al.
Publicado: (2025)
por: Li, Jiaheng, et al.
Publicado: (2025)
VmambaIR: Visual State Space Model for Image Restoration
por: Shi, Yuan, et al.
Publicado: (2024)
por: Shi, Yuan, et al.
Publicado: (2024)
Ejemplares similares
-
Multigranular Evaluation for Brain Visual Decoding
por: Xia, Weihao, et al.
Publicado: (2025) -
UMBRAE: Unified Multimodal Brain Decoding
por: Xia, Weihao, et al.
Publicado: (2024) -
RETRO: REthinking Tactile Representation Learning with Material PriOrs
por: Xia, Weihao, et al.
Publicado: (2025) -
DREAM: Visual Decoding from Reversing Human Visual System
por: Xia, Weihao, et al.
Publicado: (2023) -
Quartet of Diffusions: Structure-Aware Point Cloud Generation through Part and Symmetry Guidance
por: Zhou, Chenliang, et al.
Publicado: (2026)