SIEDD: Shared-Implicit Encoder with Discrete Decoders
Fuente:
arXiv
Saved in:
| Main Authors: | Rangarajan, Vikram, Maiya, Shishira, Ehrlich, Max, Shrivastava, Abhinav |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Latent-INR: A Flexible Framework for Implicit Representations of Videos with Discriminative Semantics
by: Maiya, Shishira R, et al.
Published: (2024)
by: Maiya, Shishira R, et al.
Published: (2024)
Explaining the Implicit Neural Canvas: Connecting Pixels to Neurons by Tracing their Contributions
by: Padmanabhan, Namitha, et al.
Published: (2024)
by: Padmanabhan, Namitha, et al.
Published: (2024)
Continuous Video Process: Modeling Videos as Continuous Multi-Dimensional Processes for Video Prediction
by: Shrivastava, Gaurav, et al.
Published: (2024)
by: Shrivastava, Gaurav, et al.
Published: (2024)
LEIA: Latent View-invariant Embeddings for Implicit 3D Articulation
by: Swaminathan, Archana, et al.
Published: (2024)
by: Swaminathan, Archana, et al.
Published: (2024)
Implicit Inversion turns CLIP into a Decoder
by: D'Orazio, Antonio, et al.
Published: (2025)
by: D'Orazio, Antonio, et al.
Published: (2025)
Hierarchical Active Inference using Successor Representations
by: Rangarajan, Prashant, et al.
Published: (2026)
by: Rangarajan, Prashant, et al.
Published: (2026)
SENCA-st: Integrating Spatial Transcriptomics and Histopathology with Cross Attention Shared Encoder for Region Identification in Cancer Pathology
by: Liyanaarachchi, Shanaka, et al.
Published: (2025)
by: Liyanaarachchi, Shanaka, et al.
Published: (2025)
DisCoRD: Discrete Tokens to Continuous Motion via Rectified Flow Decoding
by: Cho, Jungbin, et al.
Published: (2024)
by: Cho, Jungbin, et al.
Published: (2024)
Multi-Task Learning with Additive U-Net for Image Denoising and Classification
by: Lakkavalli, Vikram, et al.
Published: (2026)
by: Lakkavalli, Vikram, et al.
Published: (2026)
V-VIPE: Variational View Invariant Pose Embedding
by: Levy, Mara, et al.
Published: (2024)
by: Levy, Mara, et al.
Published: (2024)
Rethinking Encoder-Decoder Flow Through Shared Structures
by: Laboyrie, Frederik, et al.
Published: (2025)
by: Laboyrie, Frederik, et al.
Published: (2025)
Quantum Implicit Neural Representations
by: Zhao, Jiaming, et al.
Published: (2024)
by: Zhao, Jiaming, et al.
Published: (2024)
Universal Algorithm-Implicit Learning
by: Woerner, Stefano, et al.
Published: (2026)
by: Woerner, Stefano, et al.
Published: (2026)
Video Decomposition Prior: A Methodology to Decompose Videos into Layers
by: Shrivastava, Gaurav, et al.
Published: (2024)
by: Shrivastava, Gaurav, et al.
Published: (2024)
Explaining Similarity in Vision-Language Encoders with Weighted Banzhaf Interactions
by: Baniecki, Hubert, et al.
Published: (2025)
by: Baniecki, Hubert, et al.
Published: (2025)
Foundation Visual Encoders Are Secretly Few-Shot Anomaly Detectors
by: Zhai, Guangyao, et al.
Published: (2025)
by: Zhai, Guangyao, et al.
Published: (2025)
Fine-tuning CLIP Text Encoders with Two-step Paraphrasing
by: Kim, Hyunjae, et al.
Published: (2024)
by: Kim, Hyunjae, et al.
Published: (2024)
An Empirical Study into Clustering of Unseen Datasets with Self-Supervised Encoders
by: Lowe, Scott C., et al.
Published: (2024)
by: Lowe, Scott C., et al.
Published: (2024)
ThermEval: A Structured Benchmark for Evaluation of Vision-Language Models on Thermal Imagery
by: Shrivastava, Ayush, et al.
Published: (2026)
by: Shrivastava, Ayush, et al.
Published: (2026)
Adaptive Object Detection for Indoor Navigation Assistance: A Performance Evaluation of Real-Time Algorithms
by: Pratap, Abhinav, et al.
Published: (2025)
by: Pratap, Abhinav, et al.
Published: (2025)
VIBE: Video-Input Brain Encoder for fMRI Response Modeling
by: Schad, Daniel Carlström, et al.
Published: (2025)
by: Schad, Daniel Carlström, et al.
Published: (2025)
UniFusion: Vision-Language Model as Unified Encoder in Image Generation
by: Li, Kevin, et al.
Published: (2025)
by: Li, Kevin, et al.
Published: (2025)
Robustness in Both Domains: CLIP Needs a Robust Text Encoder
by: Rocamora, Elias Abad, et al.
Published: (2025)
by: Rocamora, Elias Abad, et al.
Published: (2025)
Visual Encoders for Data-Efficient Imitation Learning in Modern Video Games
by: Schäfer, Lukas, et al.
Published: (2023)
by: Schäfer, Lukas, et al.
Published: (2023)
TextCraftor: Your Text Encoder Can be Image Quality Controller
by: Li, Yanyu, et al.
Published: (2024)
by: Li, Yanyu, et al.
Published: (2024)
EGA: Adapting Frozen Encoders for Vector Search with Bounded Out-of-Distribution Degradation
by: Zhao, Dongfang
Published: (2026)
by: Zhao, Dongfang
Published: (2026)
Buster: Implanting Semantic Backdoor into Text Encoder to Mitigate NSFW Content Generation
by: Zhao, Xin, et al.
Published: (2024)
by: Zhao, Xin, et al.
Published: (2024)
Implicit Contrastive Representation Learning with Guided Stop-gradient
by: Lee, Byeongchan, et al.
Published: (2025)
by: Lee, Byeongchan, et al.
Published: (2025)
MINR: Implicit Neural Representations with Masked Image Modelling
by: Lee, Sua, et al.
Published: (2025)
by: Lee, Sua, et al.
Published: (2025)
Denoising Fisher Training For Neural Implicit Samplers
by: Luo, Weijian, et al.
Published: (2024)
by: Luo, Weijian, et al.
Published: (2024)
GenMM: Geometrically and Temporally Consistent Multimodal Data Generation for Video and LiDAR
by: Singh, Bharat, et al.
Published: (2024)
by: Singh, Bharat, et al.
Published: (2024)
Equilibrium Matching: Generative Modeling with Implicit Energy-Based Models
by: Wang, Runqian, et al.
Published: (2025)
by: Wang, Runqian, et al.
Published: (2025)
SCFlow: Implicitly Learning Style and Content Disentanglement with Flow Models
by: Ma, Pingchuan, et al.
Published: (2025)
by: Ma, Pingchuan, et al.
Published: (2025)
One-Step Diffusion Distillation through Score Implicit Matching
by: Luo, Weijian, et al.
Published: (2024)
by: Luo, Weijian, et al.
Published: (2024)
Scale Space Diffusion
by: Mukhopadhyay, Soumik, et al.
Published: (2026)
by: Mukhopadhyay, Soumik, et al.
Published: (2026)
Skrr: Skip and Re-use Text Encoder Layers for Memory Efficient Text-to-Image Generation
by: Seo, Hoigi, et al.
Published: (2025)
by: Seo, Hoigi, et al.
Published: (2025)
ViTok-v2: Scaling Native Resolution Auto-Encoders to 5 Billion Parameters
by: Hansen-Estruch, Philippe, et al.
Published: (2026)
by: Hansen-Estruch, Philippe, et al.
Published: (2026)
IMEX-Reg: Implicit-Explicit Regularization in the Function Space for Continual Learning
by: Bhat, Prashant, et al.
Published: (2024)
by: Bhat, Prashant, et al.
Published: (2024)
Can Visual Encoder Learn to See Arrows?
by: Terashita, Naoyuki, et al.
Published: (2025)
by: Terashita, Naoyuki, et al.
Published: (2025)
Renaissance: Investigating the Pretraining of Vision-Language Encoders
by: Fields, Clayton, et al.
Published: (2024)
by: Fields, Clayton, et al.
Published: (2024)
Similar Items
-
Latent-INR: A Flexible Framework for Implicit Representations of Videos with Discriminative Semantics
by: Maiya, Shishira R, et al.
Published: (2024) -
Explaining the Implicit Neural Canvas: Connecting Pixels to Neurons by Tracing their Contributions
by: Padmanabhan, Namitha, et al.
Published: (2024) -
Continuous Video Process: Modeling Videos as Continuous Multi-Dimensional Processes for Video Prediction
by: Shrivastava, Gaurav, et al.
Published: (2024) -
LEIA: Latent View-invariant Embeddings for Implicit 3D Articulation
by: Swaminathan, Archana, et al.
Published: (2024) -
Implicit Inversion turns CLIP into a Decoder
by: D'Orazio, Antonio, et al.
Published: (2025)