RAVEN: Rethinking Adversarial Video Generation with Efficient Tri-plane Networks
Fuente:
arXiv
Guardado en:
| Autores principales: | Ghosh, Partha, Sanyal, Soubhik, Schmid, Cordelia, Schölkopf, Bernhard |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SCULPT: Shape-Conditioned Unpaired Learning of Pose-dependent Clothed and Textured Human Meshes
por: Sanyal, Soubhik, et al.
Publicado: (2023)
por: Sanyal, Soubhik, et al.
Publicado: (2023)
Time-, Memory- and Parameter-Efficient Visual Adaptation
por: Mercea, Otniel-Bogdan, et al.
Publicado: (2024)
por: Mercea, Otniel-Bogdan, et al.
Publicado: (2024)
Generation is Required for Data-Efficient Perception
por: Brady, Jack, et al.
Publicado: (2025)
por: Brady, Jack, et al.
Publicado: (2025)
Generative Zoo
por: Niewiadomski, Tomasz, et al.
Publicado: (2024)
por: Niewiadomski, Tomasz, et al.
Publicado: (2024)
CAViAR: Critic-Augmented Video Agentic Reasoning
por: Menon, Sachit, et al.
Publicado: (2025)
por: Menon, Sachit, et al.
Publicado: (2025)
ViViDex: Learning Vision-based Dexterous Manipulation from Human Videos
por: Chen, Zerui, et al.
Publicado: (2024)
por: Chen, Zerui, et al.
Publicado: (2024)
DataDream: Few-shot Guided Dataset Generation
por: Kim, Jae Myung, et al.
Publicado: (2024)
por: Kim, Jae Myung, et al.
Publicado: (2024)
Hyperbolic Busemann Neural Networks
por: Chen, Ziheng, et al.
Publicado: (2026)
por: Chen, Ziheng, et al.
Publicado: (2026)
MoReVQA: Exploring Modular Reasoning Models for Video Question Answering
por: Min, Juhong, et al.
Publicado: (2024)
por: Min, Juhong, et al.
Publicado: (2024)
CaptionFormer: Unified Segmentation, Tracking, and Captioning for Spatio-Temporal Objects
por: Fiastre, Gabriel, et al.
Publicado: (2025)
por: Fiastre, Gabriel, et al.
Publicado: (2025)
A-I-RAVEN and I-RAVEN-Mesh: Two New Benchmarks for Abstract Visual Reasoning
por: Małkiński, Mikołaj, et al.
Publicado: (2024)
por: Małkiński, Mikołaj, et al.
Publicado: (2024)
Grounded Video Caption Generation
por: Kazakos, Evangelos, et al.
Publicado: (2024)
por: Kazakos, Evangelos, et al.
Publicado: (2024)
Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding
por: Nagrani, Arsha, et al.
Publicado: (2026)
por: Nagrani, Arsha, et al.
Publicado: (2026)
Video to Video Generative Adversarial Network for Few-shot Learning Based on Policy Gradient
por: Ma, Yintai, et al.
Publicado: (2024)
por: Ma, Yintai, et al.
Publicado: (2024)
MINERVA: Evaluating Complex Video Reasoning
por: Nagrani, Arsha, et al.
Publicado: (2025)
por: Nagrani, Arsha, et al.
Publicado: (2025)
Large-scale Pre-training for Grounded Video Caption Generation
por: Kazakos, Evangelos, et al.
Publicado: (2025)
por: Kazakos, Evangelos, et al.
Publicado: (2025)
Rethinking JEPA: Compute-Efficient Video SSL with Frozen Teachers
por: Li, Xianhang, et al.
Publicado: (2025)
por: Li, Xianhang, et al.
Publicado: (2025)
Rays as Pixels: Learning A Joint Distribution of Videos and Camera Trajectories
por: Jang, Wonbong, et al.
Publicado: (2026)
por: Jang, Wonbong, et al.
Publicado: (2026)
Evaluating the Adversarial Robustness of Semantic Segmentation: Trying Harder Pays Off
por: Halmosi, Levente, et al.
Publicado: (2024)
por: Halmosi, Levente, et al.
Publicado: (2024)
Diffusion-Based Representation Learning
por: Mittal, Sarthak, et al.
Publicado: (2021)
por: Mittal, Sarthak, et al.
Publicado: (2021)
Drifting Fields are not Conservative
por: Franz, Leonard T., et al.
Publicado: (2026)
por: Franz, Leonard T., et al.
Publicado: (2026)
RAVEN: Query-Guided Representation Alignment for Question Answering over Audio, Video, Embedded Sensors, and Natural Language
por: Biswas, Subrata, et al.
Publicado: (2025)
por: Biswas, Subrata, et al.
Publicado: (2025)
What Drives Compositional Generalization? The Importance of Continuous Training Objectives in Visual Generative Models
por: Farid, Karim, et al.
Publicado: (2025)
por: Farid, Karim, et al.
Publicado: (2025)
Chapter-Llama: Efficient Chaptering in Hour-Long Videos with LLMs
por: Ventura, Lucas, et al.
Publicado: (2025)
por: Ventura, Lucas, et al.
Publicado: (2025)
Language-Guided Image Tokenization for Generation
por: Zha, Kaiwen, et al.
Publicado: (2024)
por: Zha, Kaiwen, et al.
Publicado: (2024)
Verbalized Machine Learning: Revisiting Machine Learning with Language Models
por: Xiao, Tim Z., et al.
Publicado: (2024)
por: Xiao, Tim Z., et al.
Publicado: (2024)
Structure by Architecture: Structured Representations without Regularization
por: Leeb, Felix, et al.
Publicado: (2020)
por: Leeb, Felix, et al.
Publicado: (2020)
RECODE: Reasoning Through Code Generation for Visual Question Answering
por: Shen, Junhong, et al.
Publicado: (2025)
por: Shen, Junhong, et al.
Publicado: (2025)
Visual Lexicon: Rich Image Features in Language Space
por: Wang, XuDong, et al.
Publicado: (2024)
por: Wang, XuDong, et al.
Publicado: (2024)
GraphDreamer: Compositional 3D Scene Synthesis from Scene Graphs
por: Gao, Gege, et al.
Publicado: (2023)
por: Gao, Gege, et al.
Publicado: (2023)
Improving Generative Adversarial Networks with Self-Distillation
por: Nowinowski, Antoni, et al.
Publicado: (2026)
por: Nowinowski, Antoni, et al.
Publicado: (2026)
Ghost on the Shell: An Expressive Representation of General 3D Shapes
por: Liu, Zhen, et al.
Publicado: (2023)
por: Liu, Zhen, et al.
Publicado: (2023)
A Cost-Efficient Approach for Creating Virtual Fitting Room using Generative Adversarial Networks (GANs)
por: Attallah, Kirolos, et al.
Publicado: (2024)
por: Attallah, Kirolos, et al.
Publicado: (2024)
Virtual Fitting Room: Generating Arbitrarily Long Videos of Virtual Try-On from a Single Image -- Technical Preview
por: Chen, Jun-Kun, et al.
Publicado: (2025)
por: Chen, Jun-Kun, et al.
Publicado: (2025)
BrickNet: Graph-Backed Generative Brick Assembly
por: Kulits, Peter, et al.
Publicado: (2026)
por: Kulits, Peter, et al.
Publicado: (2026)
Generative Adversarial Networks Bridging Art and Machine Intelligence
por: Song, Junhao, et al.
Publicado: (2025)
por: Song, Junhao, et al.
Publicado: (2025)
Nested Annealed Training Scheme for Generative Adversarial Networks
por: Wan, Chang, et al.
Publicado: (2025)
por: Wan, Chang, et al.
Publicado: (2025)
Synthetic Medical Imaging Generation with Generative Adversarial Networks For Plain Radiographs
por: McNulty, John R., et al.
Publicado: (2024)
por: McNulty, John R., et al.
Publicado: (2024)
Synthetic Melanoma Image Generation and Evaluation Using Generative Adversarial Networks
por: Lin, Pei-Yu, et al.
Publicado: (2026)
por: Lin, Pei-Yu, et al.
Publicado: (2026)
Ego4OOD: Rethinking Egocentric Video Domain Generalization via Covariate Shift Scoring
por: Vaseqi, Zahra, et al.
Publicado: (2026)
por: Vaseqi, Zahra, et al.
Publicado: (2026)
Ejemplares similares
-
SCULPT: Shape-Conditioned Unpaired Learning of Pose-dependent Clothed and Textured Human Meshes
por: Sanyal, Soubhik, et al.
Publicado: (2023) -
Time-, Memory- and Parameter-Efficient Visual Adaptation
por: Mercea, Otniel-Bogdan, et al.
Publicado: (2024) -
Generation is Required for Data-Efficient Perception
por: Brady, Jack, et al.
Publicado: (2025) -
Generative Zoo
por: Niewiadomski, Tomasz, et al.
Publicado: (2024) -
CAViAR: Critic-Augmented Video Agentic Reasoning
por: Menon, Sachit, et al.
Publicado: (2025)