Quantised Global Autoencoder: A Holistic Approach to Representing Visual Data
Fuente:
arXiv
Saved in:
| Main Authors: | Elsner, Tim, Usinger, Paula, Czech, Victor, Kobsik, Gregor, He, Yanjiang, Lim, Isaak, Kobbelt, Leif |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multidimensional Byte Pair Encoding: Shortened Sequences for Improved Visual Data Generation
by: Elsner, Tim, et al.
Published: (2024)
by: Elsner, Tim, et al.
Published: (2024)
Learning Fine-to-Coarse Cuboid Shape Abstraction
by: Kobsik, Gregor, et al.
Published: (2025)
by: Kobsik, Gregor, et al.
Published: (2025)
Self‐supervised Learning of Fine‐to‐Coarse Cuboid Shape Abstraction
by: Gregor Kobsik, et al.
Published: (2026)
by: Gregor Kobsik, et al.
Published: (2026)
Retargeting Visual Data with Deformation Fields
by: Elsner, Tim, et al.
Published: (2023)
by: Elsner, Tim, et al.
Published: (2023)
DeferredGS: Decoupled and Editable Gaussian Splatting with Deferred Shading
by: Wu, Tong, et al.
Published: (2024)
by: Wu, Tong, et al.
Published: (2024)
4Dynamic: Text-to-4D Generation with Hybrid Priors
by: Yuan, Yu-Jie, et al.
Published: (2024)
by: Yuan, Yu-Jie, et al.
Published: (2024)
Global Context with Discrete Diffusion in Vector Quantised Modelling for Image Generation
by: Hu, Minghui, et al.
Published: (2021)
by: Hu, Minghui, et al.
Published: (2021)
Learning Quantised Structure-Preserving Motion Representations for Dance Fingerprinting
by: Kharlamova, Arina, et al.
Published: (2026)
by: Kharlamova, Arina, et al.
Published: (2026)
Product-Quantised Image Representation for High-Quality Image Synthesis
by: Zavadski, Denis, et al.
Published: (2025)
by: Zavadski, Denis, et al.
Published: (2025)
Divide-and-Conquer Approach to Holistic Cognition in High-Similarity Contexts with Limited Data
by: Wang, Shijie, et al.
Published: (2026)
by: Wang, Shijie, et al.
Published: (2026)
GViT: Representing Images as Gaussians for Visual Recognition
by: Hernandez, Jefferson, et al.
Published: (2025)
by: Hernandez, Jefferson, et al.
Published: (2025)
Hardware-Aware Feature Extraction Quantisation for Real-Time Visual Odometry on FPGA Platforms
by: Wasala, Mateusz, et al.
Published: (2025)
by: Wasala, Mateusz, et al.
Published: (2025)
On the Holistic Approach for Detecting Human Image Forgery
by: Guo, Xiao, et al.
Published: (2026)
by: Guo, Xiao, et al.
Published: (2026)
Towards Reliable and Holistic Visual In-Context Learning Prompt Selection
by: Wu, Wenxiao, et al.
Published: (2025)
by: Wu, Wenxiao, et al.
Published: (2025)
Visual Variational Autoencoder Prompt Tuning
by: Xiao, Xi, et al.
Published: (2025)
by: Xiao, Xi, et al.
Published: (2025)
Image Categorization and Search via a GAT Autoencoder and Representative Models
by: Sap, Duygu, et al.
Published: (2025)
by: Sap, Duygu, et al.
Published: (2025)
Holistic Visual-Textual Sentiment Analysis with Prior Models
by: Chen, Junyu, et al.
Published: (2022)
by: Chen, Junyu, et al.
Published: (2022)
InstructVEdit: A Holistic Approach for Instructional Video Editing
by: Zhang, Chi, et al.
Published: (2025)
by: Zhang, Chi, et al.
Published: (2025)
Artifacts of Idiosyncracy in Global Street View Data
by: Alpherts, Tim, et al.
Published: (2025)
by: Alpherts, Tim, et al.
Published: (2025)
A Note on Generalization in Variational Autoencoders: How Effective Is Synthetic Data & Overparameterization?
by: Xiao, Tim Z., et al.
Published: (2023)
by: Xiao, Tim Z., et al.
Published: (2023)
Frugal Incremental Generative Modeling using Variational Autoencoders
by: Enescu, Victor, et al.
Published: (2025)
by: Enescu, Victor, et al.
Published: (2025)
Masked Autoencoders are Robust Data Augmentors
by: Xu, Haohang, et al.
Published: (2022)
by: Xu, Haohang, et al.
Published: (2022)
Sen2Fire: A Challenging Benchmark Dataset for Wildfire Detection using Sentinel Data
by: Xu, Yonghao, et al.
Published: (2024)
by: Xu, Yonghao, et al.
Published: (2024)
Don't Just Chase "Highlighted Tokens" in MLLMs: Revisiting Visual Holistic Context Retention
by: Zou, Xin, et al.
Published: (2025)
by: Zou, Xin, et al.
Published: (2025)
Leave No Stone Unturned: Uncovering Holistic Audio-Visual Intrinsic Coherence for Deepfake Detection
by: Peng, Jielun, et al.
Published: (2026)
by: Peng, Jielun, et al.
Published: (2026)
Hi-VAE: Efficient Video Autoencoding with Global and Detailed Motion
by: Liu, Huaize, et al.
Published: (2025)
by: Liu, Huaize, et al.
Published: (2025)
Towards Open-Ended Visual Scientific Discovery with Sparse Autoencoders
by: Stevens, Samuel, et al.
Published: (2025)
by: Stevens, Samuel, et al.
Published: (2025)
VHS: High-Resolution Iterative Stereo Matching with Visual Hull Priors
by: Plack, Markus, et al.
Published: (2024)
by: Plack, Markus, et al.
Published: (2024)
Through-Foliage Surface-Temperature Reconstruction for early Wildfire Detection
by: Youssef, Mohamed, et al.
Published: (2025)
by: Youssef, Mohamed, et al.
Published: (2025)
Linear-Time Global Visual Modeling without Explicit Attention
by: He, Ruize, et al.
Published: (2026)
by: He, Ruize, et al.
Published: (2026)
Comparing Next-Day Wildfire Predictability of MODIS and VIIRS Satellite Data
by: Karlsson, Justus, et al.
Published: (2025)
by: Karlsson, Justus, et al.
Published: (2025)
Disentangling Visual Priors: Unsupervised Learning of Scene Interpretations with Compositional Autoencoder
by: Krawiec, Krzysztof, et al.
Published: (2024)
by: Krawiec, Krzysztof, et al.
Published: (2024)
Modeling Visual Memorability Assessment with Autoencoders Reveals Characteristics of Memorable Images
by: Bagheri, Elham, et al.
Published: (2024)
by: Bagheri, Elham, et al.
Published: (2024)
Masked Autoencoders are Parameter-Efficient Federated Continual Learners
by: He, Yuchen, et al.
Published: (2024)
by: He, Yuchen, et al.
Published: (2024)
From Objects to Anywhere: A Holistic Benchmark for Multi-level Visual Grounding in 3D Scenes
by: Wang, Tianxu, et al.
Published: (2025)
by: Wang, Tianxu, et al.
Published: (2025)
Virtual Full-stack Scanning of Brain MRI via Imputing Any Quantised Code
by: Wu, Yicheng, et al.
Published: (2025)
by: Wu, Yicheng, et al.
Published: (2025)
SENet: A Spectral Filtering Approach to Represent Exemplars for Few-shot Learning
by: Zhang, Tao, et al.
Published: (2023)
by: Zhang, Tao, et al.
Published: (2023)
Animation Needs Attention: A Holistic Approach to Slides Animation Comprehension with Visual-Language Models
by: Jiang, Yifan, et al.
Published: (2025)
by: Jiang, Yifan, et al.
Published: (2025)
Scalable Audio-Visual Masked Autoencoders for Efficient Affective Video Facial Analysis
by: Wu, Xuecheng, et al.
Published: (2025)
by: Wu, Xuecheng, et al.
Published: (2025)
Mixture-of-Modality-Experts with Holistic Token Learning for Fine-Grained Multimodal Visual Analytics in Driver Action Recognition
by: Liu, Tianyi, et al.
Published: (2026)
by: Liu, Tianyi, et al.
Published: (2026)
Similar Items
-
Multidimensional Byte Pair Encoding: Shortened Sequences for Improved Visual Data Generation
by: Elsner, Tim, et al.
Published: (2024) -
Learning Fine-to-Coarse Cuboid Shape Abstraction
by: Kobsik, Gregor, et al.
Published: (2025) -
Self‐supervised Learning of Fine‐to‐Coarse Cuboid Shape Abstraction
by: Gregor Kobsik, et al.
Published: (2026) -
Retargeting Visual Data with Deformation Fields
by: Elsner, Tim, et al.
Published: (2023) -
DeferredGS: Decoupled and Editable Gaussian Splatting with Deferred Shading
by: Wu, Tong, et al.
Published: (2024)