NeuroCine: Decoding Vivid Video Sequences from Human Brain Activties
Fuente:
arXiv
Salvato in:
| Autori principali: | Sun, Jingyuan, Li, Mingxiao, Chen, Zijiao, Moens, Marie-Francine |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Animate Your Motion: Turning Still Images into Dynamic Videos
di: Li, Mingxiao, et al.
Pubblicazione: (2024)
di: Li, Mingxiao, et al.
Pubblicazione: (2024)
Consistent Story Generation: Unlocking the Potential of Zigzag Sampling
di: Li, Mingxiao, et al.
Pubblicazione: (2025)
di: Li, Mingxiao, et al.
Pubblicazione: (2025)
TS-LLaVA: Constructing Visual Tokens through Thumbnail-and-Sampling for Training-Free Video Large Language Models
di: Qu, Tingyu, et al.
Pubblicazione: (2024)
di: Qu, Tingyu, et al.
Pubblicazione: (2024)
Action-based image editing guided by human instructions
di: Trusca, Maria Mihaela, et al.
Pubblicazione: (2024)
di: Trusca, Maria Mihaela, et al.
Pubblicazione: (2024)
Towards More Accurate Personalized Image Generation: Addressing Overfitting and Evaluation Bias
di: Li, Mingxiao, et al.
Pubblicazione: (2025)
di: Li, Mingxiao, et al.
Pubblicazione: (2025)
Alleviating Exposure Bias in Diffusion Models through Sampling with Shifted Time Steps
di: Li, Mingxiao, et al.
Pubblicazione: (2023)
di: Li, Mingxiao, et al.
Pubblicazione: (2023)
NeurIPS: Neuro-anatomical Inductive Priors for Sphere-based Brain Decoding
di: Yu, Sijin, et al.
Pubblicazione: (2026)
di: Yu, Sijin, et al.
Pubblicazione: (2026)
Introducing Routing Functions to Vision-Language Parameter-Efficient Fine-Tuning with Low-Rank Bottlenecks
di: Qu, Tingyu, et al.
Pubblicazione: (2024)
di: Qu, Tingyu, et al.
Pubblicazione: (2024)
Visually-Aware Context Modeling for News Image Captioning
di: Qu, Tingyu, et al.
Pubblicazione: (2023)
di: Qu, Tingyu, et al.
Pubblicazione: (2023)
DM-Align: Leveraging the Power of Natural Language Instructions to Make Changes to Images
di: Trusca, Maria Mihaela, et al.
Pubblicazione: (2024)
di: Trusca, Maria Mihaela, et al.
Pubblicazione: (2024)
VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation
di: Wang, Qilin, et al.
Pubblicazione: (2024)
di: Wang, Qilin, et al.
Pubblicazione: (2024)
Vivid-VR: Distilling Concepts from Text-to-Video Diffusion Transformer for Photorealistic Video Restoration
di: Bai, Haoran, et al.
Pubblicazione: (2025)
di: Bai, Haoran, et al.
Pubblicazione: (2025)
Vivid-ZOO: Multi-View Video Generation with Diffusion Model
di: Li, Bing, et al.
Pubblicazione: (2024)
di: Li, Bing, et al.
Pubblicazione: (2024)
Automated Prompt Generation for Creative and Counterfactual Text-to-image Synthesis
di: Jelaca, Aleksa, et al.
Pubblicazione: (2025)
di: Jelaca, Aleksa, et al.
Pubblicazione: (2025)
MegActor: Harness the Power of Raw Video for Vivid Portrait Animation
di: Yang, Shurong, et al.
Pubblicazione: (2024)
di: Yang, Shurong, et al.
Pubblicazione: (2024)
PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement
di: Hu, Teng, et al.
Pubblicazione: (2025)
di: Hu, Teng, et al.
Pubblicazione: (2025)
GVA: Reconstructing Vivid 3D Gaussian Avatars from Monocular Videos
di: Liu, Xinqi, et al.
Pubblicazione: (2024)
di: Liu, Xinqi, et al.
Pubblicazione: (2024)
VividFace: High-Quality and Efficient One-Step Diffusion For Video Face Enhancement
di: Zhang, Shulian, et al.
Pubblicazione: (2025)
di: Zhang, Shulian, et al.
Pubblicazione: (2025)
Vivid4D: Improving 4D Reconstruction from Monocular Video by Video Inpainting
di: Huang, Jiaxin, et al.
Pubblicazione: (2025)
di: Huang, Jiaxin, et al.
Pubblicazione: (2025)
VividCam: Learning Unconventional Camera Motions from Virtual Synthetic Videos
di: Wu, Qiucheng, et al.
Pubblicazione: (2025)
di: Wu, Qiucheng, et al.
Pubblicazione: (2025)
From Flat to Round: Redefining Brain Decoding with Surface-Based fMRI and Cortex Structure
di: Yu, Sijin, et al.
Pubblicazione: (2025)
di: Yu, Sijin, et al.
Pubblicazione: (2025)
VividListener: Expressive and Controllable Listener Dynamics Modeling for Multi-Modal Responsive Interaction
di: Li, Shiying, et al.
Pubblicazione: (2025)
di: Li, Shiying, et al.
Pubblicazione: (2025)
VividAnimator: An End-to-End Audio and Pose-driven Half-Body Human Animation Framework
di: Huang, Donglin, et al.
Pubblicazione: (2025)
di: Huang, Donglin, et al.
Pubblicazione: (2025)
Object-Attribute Binding in Text-to-Image Generation: Evaluation and Control
di: Trusca, Maria Mihaela, et al.
Pubblicazione: (2024)
di: Trusca, Maria Mihaela, et al.
Pubblicazione: (2024)
Multimodal Adaptive Inference for Document Image Classification with Anytime Early Exiting
di: Hamed, Omar, et al.
Pubblicazione: (2024)
di: Hamed, Omar, et al.
Pubblicazione: (2024)
Deep Neural Networks and Brain Alignment: Brain Encoding and Decoding (Survey)
di: Oota, Subba Reddy, et al.
Pubblicazione: (2023)
di: Oota, Subba Reddy, et al.
Pubblicazione: (2023)
Make-It-Vivid: Dressing Your Animatable Biped Cartoon Characters from Text
di: Tang, Junshu, et al.
Pubblicazione: (2024)
di: Tang, Junshu, et al.
Pubblicazione: (2024)
WATCH: World-aware Allied Trajectory and pose reconstruction for Camera and Human
di: Ying, Qijun, et al.
Pubblicazione: (2025)
di: Ying, Qijun, et al.
Pubblicazione: (2025)
VividMed: Vision Language Model with Versatile Visual Grounding for Medicine
di: Luo, Lingxiao, et al.
Pubblicazione: (2024)
di: Luo, Lingxiao, et al.
Pubblicazione: (2024)
Preserving Old Memories in Vivid Detail: Human-Interactive Photo Restoration Framework
di: Back, Seung-Yeon, et al.
Pubblicazione: (2024)
di: Back, Seung-Yeon, et al.
Pubblicazione: (2024)
VividDreamer: Towards High-Fidelity and Efficient Text-to-3D Generation
di: Chen, Zixuan, et al.
Pubblicazione: (2024)
di: Chen, Zixuan, et al.
Pubblicazione: (2024)
Too Vivid to Be Real? Benchmarking and Calibrating Generative Color Fidelity
di: Fang, Zhengyao, et al.
Pubblicazione: (2026)
di: Fang, Zhengyao, et al.
Pubblicazione: (2026)
EMOVA: Empowering Language Models to See, Hear and Speak with Vivid Emotions
di: Chen, Kai, et al.
Pubblicazione: (2024)
di: Chen, Kai, et al.
Pubblicazione: (2024)
Textoon: Generating Vivid 2D Cartoon Characters from Text Descriptions
di: He, Chao, et al.
Pubblicazione: (2025)
di: He, Chao, et al.
Pubblicazione: (2025)
VividDream: Generating 3D Scene with Ambient Dynamics
di: Lee, Yao-Chih, et al.
Pubblicazione: (2024)
di: Lee, Yao-Chih, et al.
Pubblicazione: (2024)
Patient4D: Temporally Consistent Patient Body Mesh Recovery from Monocular Operating Room Video
di: Tu, Mingxiao, et al.
Pubblicazione: (2026)
di: Tu, Mingxiao, et al.
Pubblicazione: (2026)
Let Storytelling Tell Vivid Stories: An Expressive and Fluent Multimodal Storyteller
di: Zang, Chuanqi, et al.
Pubblicazione: (2024)
di: Zang, Chuanqi, et al.
Pubblicazione: (2024)
SVP: Style-Enhanced Vivid Portrait Talking Head Diffusion Model
di: Tan, Weipeng, et al.
Pubblicazione: (2024)
di: Tan, Weipeng, et al.
Pubblicazione: (2024)
Architect: Generating Vivid and Interactive 3D Scenes with Hierarchical 2D Inpainting
di: Wang, Yian, et al.
Pubblicazione: (2024)
di: Wang, Yian, et al.
Pubblicazione: (2024)
MVPortrait: Text-Guided Motion and Emotion Control for Multi-view Vivid Portrait Animation
di: Lin, Yukang, et al.
Pubblicazione: (2025)
di: Lin, Yukang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Animate Your Motion: Turning Still Images into Dynamic Videos
di: Li, Mingxiao, et al.
Pubblicazione: (2024) -
Consistent Story Generation: Unlocking the Potential of Zigzag Sampling
di: Li, Mingxiao, et al.
Pubblicazione: (2025) -
TS-LLaVA: Constructing Visual Tokens through Thumbnail-and-Sampling for Training-Free Video Large Language Models
di: Qu, Tingyu, et al.
Pubblicazione: (2024) -
Action-based image editing guided by human instructions
di: Trusca, Maria Mihaela, et al.
Pubblicazione: (2024) -
Towards More Accurate Personalized Image Generation: Addressing Overfitting and Evaluation Bias
di: Li, Mingxiao, et al.
Pubblicazione: (2025)