Psychometry: An Omnifit Model for Image Reconstruction from Human Brain Activity
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Quan, Ruijie, Wang, Wenguan, Tian, Zhibo, Ma, Fan, Yang, Yi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BrainGuard: Privacy-Preserving Multisubject Image Reconstructions from Brain Activities
von: Tian, Zhibo, et al.
Veröffentlicht: (2025)
von: Tian, Zhibo, et al.
Veröffentlicht: (2025)
Moving Beyond Diffusion: Hierarchy-to-Hierarchy Autoregression for fMRI-to-Image Reconstruction
von: Zhang, Xu, et al.
Veröffentlicht: (2025)
von: Zhang, Xu, et al.
Veröffentlicht: (2025)
Shape2Scene: 3D Scene Representation Learning Through Pre-training on Shape Data
von: Feng, Tuo, et al.
Veröffentlicht: (2024)
von: Feng, Tuo, et al.
Veröffentlicht: (2024)
General and Task-Oriented Video Segmentation
von: Chen, Mu, et al.
Veröffentlicht: (2024)
von: Chen, Mu, et al.
Veröffentlicht: (2024)
Human-Object Interaction Detection Collaborated with Large Relation-driven Diffusion Models
von: Li, Liulei, et al.
Veröffentlicht: (2024)
von: Li, Liulei, et al.
Veröffentlicht: (2024)
LSK3DNet: Towards Effective and Efficient 3D Perception with Large Sparse Kernels
von: Feng, Tuo, et al.
Veröffentlicht: (2024)
von: Feng, Tuo, et al.
Veröffentlicht: (2024)
Navigation Instruction Generation with BEV Perception and Large Language Models
von: Fan, Sheng, et al.
Veröffentlicht: (2024)
von: Fan, Sheng, et al.
Veröffentlicht: (2024)
Clustering Propagation for Universal Medical Image Segmentation
von: Ding, Yuhang, et al.
Veröffentlicht: (2024)
von: Ding, Yuhang, et al.
Veröffentlicht: (2024)
A Survey of World Models for Autonomous Driving
von: Feng, Tuo, et al.
Veröffentlicht: (2025)
von: Feng, Tuo, et al.
Veröffentlicht: (2025)
Autonomous LLM-Enhanced Adversarial Attack for Text-to-Motion
von: Miao, Honglei, et al.
Veröffentlicht: (2024)
von: Miao, Honglei, et al.
Veröffentlicht: (2024)
Adversarial-Guided Diffusion for Multimodal LLM Attacks
von: Xia, Chengwei, et al.
Veröffentlicht: (2025)
von: Xia, Chengwei, et al.
Veröffentlicht: (2025)
UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models
von: Jiang, Hong, et al.
Veröffentlicht: (2026)
von: Jiang, Hong, et al.
Veröffentlicht: (2026)
Volumetric Environment Representation for Vision-Language Navigation
von: Liu, Rui, et al.
Veröffentlicht: (2024)
von: Liu, Rui, et al.
Veröffentlicht: (2024)
Vision-Language Navigation with Energy-Based Policy
von: Liu, Rui, et al.
Veröffentlicht: (2024)
von: Liu, Rui, et al.
Veröffentlicht: (2024)
Learning Human-Object Interaction as Groups
von: Hong, Jiajun, et al.
Veröffentlicht: (2025)
von: Hong, Jiajun, et al.
Veröffentlicht: (2025)
Insert Anything: Image Insertion via In-Context Editing in DiT
von: Song, Wensong, et al.
Veröffentlicht: (2025)
von: Song, Wensong, et al.
Veröffentlicht: (2025)
Echoes of ownership: Adversarial-guided dual injection for copyright protection in MLLMs
von: Xia, Chengwei, et al.
Veröffentlicht: (2026)
von: Xia, Chengwei, et al.
Veröffentlicht: (2026)
CLIP4STR: A Simple Baseline for Scene Text Recognition with Pre-trained Vision-Language Model
von: Zhao, Shuai, et al.
Veröffentlicht: (2023)
von: Zhao, Shuai, et al.
Veröffentlicht: (2023)
Neural Clustering based Visual Representation Learning
von: Chen, Guikun, et al.
Veröffentlicht: (2024)
von: Chen, Guikun, et al.
Veröffentlicht: (2024)
Hydra-SGG: Hybrid Relation Assignment for One-stage Scene Graph Generation
von: Chen, Minghan, et al.
Veröffentlicht: (2024)
von: Chen, Minghan, et al.
Veröffentlicht: (2024)
DIFFVSGG: Diffusion-Driven Online Video Scene Graph Generation
von: Chen, Mu, et al.
Veröffentlicht: (2025)
von: Chen, Mu, et al.
Veröffentlicht: (2025)
Visual Knowledge in the Big Model Era: Retrospect and Prospect
von: Wang, Wenguan, et al.
Veröffentlicht: (2024)
von: Wang, Wenguan, et al.
Veröffentlicht: (2024)
DoraemonGPT: Toward Understanding Dynamic Scenes with Large Language Models (Exemplified as A Video Agent)
von: Yang, Zongxin, et al.
Veröffentlicht: (2024)
von: Yang, Zongxin, et al.
Veröffentlicht: (2024)
SinkTrack: Attention Sink based Context Anchoring for Large Language Models
von: Liu, Xu, et al.
Veröffentlicht: (2026)
von: Liu, Xu, et al.
Veröffentlicht: (2026)
TarPro: Targeted Protection against Malicious Image Editing
von: Shen, Kaixin, et al.
Veröffentlicht: (2025)
von: Shen, Kaixin, et al.
Veröffentlicht: (2025)
Visual Image Reconstruction from Brain Activity via Latent Representation
von: Kamitani, Yukiyasu, et al.
Veröffentlicht: (2025)
von: Kamitani, Yukiyasu, et al.
Veröffentlicht: (2025)
Reconstructing Topology-Consistent Face Mesh by Volume Rendering from Multi-View Images
von: Wang, Yating, et al.
Veröffentlicht: (2024)
von: Wang, Yating, et al.
Veröffentlicht: (2024)
Scalable Video Object Segmentation with Identification Mechanism
von: Yang, Zongxin, et al.
Veröffentlicht: (2022)
von: Yang, Zongxin, et al.
Veröffentlicht: (2022)
Scene Graph Generation with Role-Playing Large Language Models
von: Chen, Guikun, et al.
Veröffentlicht: (2024)
von: Chen, Guikun, et al.
Veröffentlicht: (2024)
Towards Data-and Knowledge-Driven Artificial Intelligence: A Survey on Neuro-Symbolic Computing
von: Wang, Wenguan, et al.
Veröffentlicht: (2022)
von: Wang, Wenguan, et al.
Veröffentlicht: (2022)
OAHuman: Occlusion-Aware 3D Human Reconstruction from Monocular Images
von: Yang, Yuanwang, et al.
Veröffentlicht: (2026)
von: Yang, Yuanwang, et al.
Veröffentlicht: (2026)
Versatile Framework with Semantic and Structural guidance for Image Reconstruction from Brain Activity
von: Lu, Yizhuo, et al.
Veröffentlicht: (2026)
von: Lu, Yizhuo, et al.
Veröffentlicht: (2026)
AudioScenic: Audio-Driven Video Scene Editing
von: Shen, Kaixin, et al.
Veröffentlicht: (2024)
von: Shen, Kaixin, et al.
Veröffentlicht: (2024)
Nonverbal Interaction Detection
von: Wei, Jianan, et al.
Veröffentlicht: (2024)
von: Wei, Jianan, et al.
Veröffentlicht: (2024)
Snap-Snap: Taking Two Images to Reconstruct 3D Human Gaussians in Milliseconds
von: Lu, Jia, et al.
Veröffentlicht: (2025)
von: Lu, Jia, et al.
Veröffentlicht: (2025)
BrainMCLIP: Brain Image Decoding with Multi-Layer feature Fusion of CLIP
von: Xia, Tian, et al.
Veröffentlicht: (2025)
von: Xia, Tian, et al.
Veröffentlicht: (2025)
Image Segmentation in Foundation Model Era: A Survey
von: Zhou, Tianfei, et al.
Veröffentlicht: (2024)
von: Zhou, Tianfei, et al.
Veröffentlicht: (2024)
3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation
von: Gao, Jianzhe, et al.
Veröffentlicht: (2026)
von: Gao, Jianzhe, et al.
Veröffentlicht: (2026)
Rethinking Cross-modal Interaction from a Top-down Perspective for Referring Video Object Segmentation
von: Liang, Chen, et al.
Veröffentlicht: (2021)
von: Liang, Chen, et al.
Veröffentlicht: (2021)
Learning 3D Representations for Spatial Intelligence from Unposed Multi-View Images
von: Zhou, Bo, et al.
Veröffentlicht: (2026)
von: Zhou, Bo, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
BrainGuard: Privacy-Preserving Multisubject Image Reconstructions from Brain Activities
von: Tian, Zhibo, et al.
Veröffentlicht: (2025) -
Moving Beyond Diffusion: Hierarchy-to-Hierarchy Autoregression for fMRI-to-Image Reconstruction
von: Zhang, Xu, et al.
Veröffentlicht: (2025) -
Shape2Scene: 3D Scene Representation Learning Through Pre-training on Shape Data
von: Feng, Tuo, et al.
Veröffentlicht: (2024) -
General and Task-Oriented Video Segmentation
von: Chen, Mu, et al.
Veröffentlicht: (2024) -
Human-Object Interaction Detection Collaborated with Large Relation-driven Diffusion Models
von: Li, Liulei, et al.
Veröffentlicht: (2024)