SemVideo: Reconstructs What You Watch from Brain Activity via Hierarchical Semantic Guidance
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Minghan, Yang, Lan, Li, Ke, Zhang, Honggang, Pang, Kaiyue, Song, Yizhe |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SynMind: Reducing Semantic Hallucination in fMRI-Based Image Reconstruction
by: Yang, Lan, et al.
Published: (2026)
by: Yang, Lan, et al.
Published: (2026)
Annotation-Free Human Sketch Quality Assessment
by: Yang, Lan, et al.
Published: (2025)
by: Yang, Lan, et al.
Published: (2025)
Wired Perspectives: Multi-View Wire Art Embraces Generative AI
by: Qu, Zhiyu, et al.
Published: (2023)
by: Qu, Zhiyu, et al.
Published: (2023)
VersaGen: Unleashing Versatile Visual Control for Text-to-Image Synthesis
by: Chen, Zhipeng, et al.
Published: (2024)
by: Chen, Zhipeng, et al.
Published: (2024)
Bridging Brain and Semantics: A Hierarchical Framework for Semantically Enhanced fMRI-to-Video Reconstruction
by: Wei, Yujie, et al.
Published: (2026)
by: Wei, Yujie, et al.
Published: (2026)
Hierarchical Spatio-temporal Segmentation Network for Ejection Fraction Estimation in Echocardiography Videos
by: Wang, Dongfang, et al.
Published: (2025)
by: Wang, Dongfang, et al.
Published: (2025)
SketchAnimator: Animate Sketch via Motion Customization of Text-to-Video Diffusion Models
by: Yang, Ruolin, et al.
Published: (2025)
by: Yang, Ruolin, et al.
Published: (2025)
SemGeoMo: Dynamic Contextual Human Motion Generation with Semantic and Geometric Guidance
by: Cong, Peishan, et al.
Published: (2025)
by: Cong, Peishan, et al.
Published: (2025)
SocialMirror: Reconstructing 3D Human Interaction Behaviors from Monocular Videos with Semantic and Geometric Guidance
by: Xia, Qi, et al.
Published: (2026)
by: Xia, Qi, et al.
Published: (2026)
BrainGuard: Privacy-Preserving Multisubject Image Reconstructions from Brain Activities
by: Tian, Zhibo, et al.
Published: (2025)
by: Tian, Zhibo, et al.
Published: (2025)
UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models
by: Jiang, Hong, et al.
Published: (2026)
by: Jiang, Hong, et al.
Published: (2026)
HiSem: Hierarchical Semantic Disentangling for Remote Sensing Image Change Captioning
by: Wang, Man, et al.
Published: (2026)
by: Wang, Man, et al.
Published: (2026)
SemFlow: Binding Semantic Segmentation and Image Synthesis via Rectified Flow
by: Wang, Chaoyang, et al.
Published: (2024)
by: Wang, Chaoyang, et al.
Published: (2024)
iTryOn: Mastering Interactive Video Virtual Try-On with Spatial-Semantic Guidance
by: Zheng, Jun, et al.
Published: (2026)
by: Zheng, Jun, et al.
Published: (2026)
Psychometry: An Omnifit Model for Image Reconstruction from Human Brain Activity
by: Quan, Ruijie, et al.
Published: (2024)
by: Quan, Ruijie, et al.
Published: (2024)
Deformable One-shot Face Stylization via DINO Semantic Guidance
by: Zhou, Yang, et al.
Published: (2024)
by: Zhou, Yang, et al.
Published: (2024)
Consistent Video Colorization via Palette Guidance
by: Wang, Han, et al.
Published: (2025)
by: Wang, Han, et al.
Published: (2025)
SemGrasp: Semantic Grasp Generation via Language Aligned Discretization
by: Li, Kailin, et al.
Published: (2024)
by: Li, Kailin, et al.
Published: (2024)
UniSem: Generalizable Semantic 3D Reconstruction from Sparse Unposed Images
by: Liao, Guibiao, et al.
Published: (2026)
by: Liao, Guibiao, et al.
Published: (2026)
Knowledge-Refined Dual Context-Aware Network for Partially Relevant Video Retrieval
by: Yang, Junkai, et al.
Published: (2026)
by: Yang, Junkai, et al.
Published: (2026)
Focal Guidance: Unlocking Controllability from Semantic-Weak Layers in Video Diffusion Models
by: Yin, Yuanyang, et al.
Published: (2026)
by: Yin, Yuanyang, et al.
Published: (2026)
Pressure2Motion: Hierarchical Human Motion Reconstruction from Ground Pressure with Text Guidance
by: Li, Zhengxuan, et al.
Published: (2025)
by: Li, Zhengxuan, et al.
Published: (2025)
Brain-Streams: fMRI-to-Image Reconstruction with Multi-modal Guidance
by: Joo, Jaehoon, et al.
Published: (2024)
by: Joo, Jaehoon, et al.
Published: (2024)
FancyVideo: Towards Dynamic and Consistent Video Generation via Cross-frame Textual Guidance
by: Feng, Jiasong, et al.
Published: (2024)
by: Feng, Jiasong, et al.
Published: (2024)
FreeSpec: Training-Free Long Video Generation via Singular-Spectrum Reconstruction
by: Chen, Fangda, et al.
Published: (2026)
by: Chen, Fangda, et al.
Published: (2026)
Plane2Depth: Hierarchical Adaptive Plane Guidance for Monocular Depth Estimation
by: Liu, Li, et al.
Published: (2024)
by: Liu, Li, et al.
Published: (2024)
Versatile Framework with Semantic and Structural guidance for Image Reconstruction from Brain Activity
by: Lu, Yizhuo, et al.
Published: (2026)
by: Lu, Yizhuo, et al.
Published: (2026)
What You See Is What Matters: A Novel Visual and Physics-Based Metric for Evaluating Video Generation Quality
by: Wang, Zihan, et al.
Published: (2024)
by: Wang, Zihan, et al.
Published: (2024)
PS-CAD: Local Geometry Guidance via Prompting and Selection for CAD Reconstruction
by: Yang, Bingchen, et al.
Published: (2024)
by: Yang, Bingchen, et al.
Published: (2024)
UrbanCraft: Urban View Extrapolation via Hierarchical Sem-Geometric Priors
by: Wang, Tianhang, et al.
Published: (2025)
by: Wang, Tianhang, et al.
Published: (2025)
Watch and Learn: Learning to Use Computers from Online Videos
by: Song, Chan Hee, et al.
Published: (2025)
by: Song, Chan Hee, et al.
Published: (2025)
SpecSem-Net: Integrating Spectral and Semantic Features for Robust AI-generated Video Detection
by: Wei, Zixi, et al.
Published: (2026)
by: Wei, Zixi, et al.
Published: (2026)
Enhanced Semantic Extraction and Guidance for UGC Image Super Resolution
by: Wang, Yiwen, et al.
Published: (2025)
by: Wang, Yiwen, et al.
Published: (2025)
Video Streaming Thinking: VideoLLMs Can Watch and Think Simultaneously
by: Guan, Yiran, et al.
Published: (2026)
by: Guan, Yiran, et al.
Published: (2026)
Animatable 3D Gaussian: Fast and High-Quality Reconstruction of Multiple Human Avatars
by: Liu, Yang, et al.
Published: (2023)
by: Liu, Yang, et al.
Published: (2023)
Single-step Diffusion-based Video Coding with Semantic-Temporal Guidance
by: Xue, Naifu, et al.
Published: (2025)
by: Xue, Naifu, et al.
Published: (2025)
NEGATE: Constrained Semantic Guidance for Linguistic Negation in Text-to-Video Diffusion
by: Kang, Taewon, et al.
Published: (2026)
by: Kang, Taewon, et al.
Published: (2026)
GeoGuide: Hierarchical Geometric Guidance for Open-Vocabulary 3D Semantic Segmentation
by: Tao, Xujing, et al.
Published: (2026)
by: Tao, Xujing, et al.
Published: (2026)
Joint Learning for Scattered Point Cloud Understanding with Hierarchical Self-Distillation
by: Zhou, Kaiyue, et al.
Published: (2023)
by: Zhou, Kaiyue, et al.
Published: (2023)
SemHiTok: A Unified Image Tokenizer via Semantic-Guided Hierarchical Codebook for Multimodal Understanding and Generation
by: Chen, Zisheng, et al.
Published: (2025)
by: Chen, Zisheng, et al.
Published: (2025)
Similar Items
-
SynMind: Reducing Semantic Hallucination in fMRI-Based Image Reconstruction
by: Yang, Lan, et al.
Published: (2026) -
Annotation-Free Human Sketch Quality Assessment
by: Yang, Lan, et al.
Published: (2025) -
Wired Perspectives: Multi-View Wire Art Embraces Generative AI
by: Qu, Zhiyu, et al.
Published: (2023) -
VersaGen: Unleashing Versatile Visual Control for Text-to-Image Synthesis
by: Chen, Zhipeng, et al.
Published: (2024) -
Bridging Brain and Semantics: A Hierarchical Framework for Semantically Enhanced fMRI-to-Video Reconstruction
by: Wei, Yujie, et al.
Published: (2026)