Coarse-To-Fine Tensor Trains for Compact Visual Representations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Loeschcke, Sebastian, Wang, Dan, Leth-Espensen, Christian, Belongie, Serge, Kastoryano, Michael J., Benaim, Sagie |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Assessing Neural Network Robustness via Adversarial Pivotal Tuning
von: Christensen, Peter Ebert, et al.
Veröffentlicht: (2022)
von: Christensen, Peter Ebert, et al.
Veröffentlicht: (2022)
Designing a Conditional Prior Distribution for Flow-Based Generative Models
von: Issachar, Noam, et al.
Veröffentlicht: (2025)
von: Issachar, Noam, et al.
Veröffentlicht: (2025)
LoQT: Low-Rank Adapters for Quantized Pretraining
von: Loeschcke, Sebastian, et al.
Veröffentlicht: (2024)
von: Loeschcke, Sebastian, et al.
Veröffentlicht: (2024)
Generating Intermediate Representations for Compositional Text-To-Image Generation
von: Galun, Ran, et al.
Veröffentlicht: (2024)
von: Galun, Ran, et al.
Veröffentlicht: (2024)
LaMI: Augmenting Large Language Models via Late Multi-Image Fusion
von: Yariv, Guy, et al.
Veröffentlicht: (2024)
von: Yariv, Guy, et al.
Veröffentlicht: (2024)
RAD: Retrieval-Augmented Monocular Metric Depth Estimation for Underrepresented Classes
von: Baltaxe, Michael, et al.
Veröffentlicht: (2026)
von: Baltaxe, Michael, et al.
Veröffentlicht: (2026)
Familiarity-Based Open-Set Recognition Under Adversarial Attacks
von: Enevoldsen, Philip, et al.
Veröffentlicht: (2023)
von: Enevoldsen, Philip, et al.
Veröffentlicht: (2023)
RAIGen: Rare Attribute Identification in Text-to-Image Generative Models
von: Sreelatha, Silpa Vadakkeeveetil, et al.
Veröffentlicht: (2026)
von: Sreelatha, Silpa Vadakkeeveetil, et al.
Veröffentlicht: (2026)
MMEarth: Exploring Multi-Modal Pretext Tasks For Geospatial Representation Learning
von: Nedungadi, Vishal, et al.
Veröffentlicht: (2024)
von: Nedungadi, Vishal, et al.
Veröffentlicht: (2024)
Discriminative Class Tokens for Text-to-Image Diffusion Models
von: Schwartz, Idan, et al.
Veröffentlicht: (2023)
von: Schwartz, Idan, et al.
Veröffentlicht: (2023)
Stitch: Training-Free Position Control in Multimodal Diffusion Transformers
von: Bader, Jessica, et al.
Veröffentlicht: (2025)
von: Bader, Jessica, et al.
Veröffentlicht: (2025)
Unlearning-based Neural Interpretations
von: Choi, Ching Lam, et al.
Veröffentlicht: (2024)
von: Choi, Ching Lam, et al.
Veröffentlicht: (2024)
RespoDiff: Dual-Module Bottleneck Transformation for Responsible & Faithful T2I Generation
von: Sreelatha, Silpa Vadakkeeveetil, et al.
Veröffentlicht: (2025)
von: Sreelatha, Silpa Vadakkeeveetil, et al.
Veröffentlicht: (2025)
From Semantics to Pixels: Coarse-to-Fine Masked Autoencoders for Hierarchical Visual Understanding
von: Xiang, Wenzhao, et al.
Veröffentlicht: (2026)
von: Xiang, Wenzhao, et al.
Veröffentlicht: (2026)
Beyond Binary Success: A Diagnostic Meta-Evaluation Framework for Fine-Grained Manipulation
von: Xu, He-Yang, et al.
Veröffentlicht: (2026)
von: Xu, He-Yang, et al.
Veröffentlicht: (2026)
Colored Noise Diffusion Sampling
von: Davidson, Hadar, et al.
Veröffentlicht: (2026)
von: Davidson, Hadar, et al.
Veröffentlicht: (2026)
RewardSDS: Aligning Score Distillation via Reward-Weighted Sampling
von: Chachy, Itay, et al.
Veröffentlicht: (2025)
von: Chachy, Itay, et al.
Veröffentlicht: (2025)
Spherical Mask: Coarse-to-Fine 3D Point Cloud Instance Segmentation with Spherical Representation
von: Shin, Sangyun, et al.
Veröffentlicht: (2023)
von: Shin, Sangyun, et al.
Veröffentlicht: (2023)
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
von: Pach, Mateusz, et al.
Veröffentlicht: (2025)
von: Pach, Mateusz, et al.
Veröffentlicht: (2025)
The Latent Color Subspace: Emergent Order in High-Dimensional Chaos
von: Pach, Mateusz, et al.
Veröffentlicht: (2026)
von: Pach, Mateusz, et al.
Veröffentlicht: (2026)
SemanticMoments: Training-Free Motion Similarity via Third Moment Features
von: Huberman, Saar, et al.
Veröffentlicht: (2026)
von: Huberman, Saar, et al.
Veröffentlicht: (2026)
Through-The-Mask: Mask-based Motion Trajectories for Image-to-Video Generation
von: Yariv, Guy, et al.
Veröffentlicht: (2025)
von: Yariv, Guy, et al.
Veröffentlicht: (2025)
PhyGenHOI: Physically-Aware 4D Generation of Dynamic Human-Object Interactions
von: Benishu, Omer, et al.
Veröffentlicht: (2026)
von: Benishu, Omer, et al.
Veröffentlicht: (2026)
MV-RAG: Retrieval Augmented Multiview Diffusion
von: Dayani, Yosef, et al.
Veröffentlicht: (2025)
von: Dayani, Yosef, et al.
Veröffentlicht: (2025)
MMEarth-Bench: Global Model Adaptation via Multimodal Test-Time Training
von: Gordon, Lucia, et al.
Veröffentlicht: (2026)
von: Gordon, Lucia, et al.
Veröffentlicht: (2026)
Boosting Fine-Grained Visual Anomaly Detection with Coarse-Knowledge-Aware Adversarial Learning
von: Fang, Qingqing, et al.
Veröffentlicht: (2024)
von: Fang, Qingqing, et al.
Veröffentlicht: (2024)
Let it Snow! Animating 3D Gaussian Scenes with Dynamic Weather Effects via Physics-Guided Score Distillation
von: Fiebelman, Gal, et al.
Veröffentlicht: (2025)
von: Fiebelman, Gal, et al.
Veröffentlicht: (2025)
VideoFlexTok: Flexible-Length Coarse-to-Fine Video Tokenization
von: Atanov, Andrei, et al.
Veröffentlicht: (2026)
von: Atanov, Andrei, et al.
Veröffentlicht: (2026)
DGD: Dynamic 3D Gaussians Distillation
von: Labe, Isaac, et al.
Veröffentlicht: (2024)
von: Labe, Isaac, et al.
Veröffentlicht: (2024)
Structurally Disentangled Feature Fields Distillation for 3D Understanding and Editing
von: Levy, Yoel, et al.
Veröffentlicht: (2025)
von: Levy, Yoel, et al.
Veröffentlicht: (2025)
Explainable Adversarial Attacks on Coarse-to-Fine Classifiers
von: Heidarizadeh, Akram, et al.
Veröffentlicht: (2025)
von: Heidarizadeh, Akram, et al.
Veröffentlicht: (2025)
Geographical Context Matters: Bridging Fine and Coarse Spatial Information to Enhance Continental Land Cover Mapping
von: Ghassemi, Babak, et al.
Veröffentlicht: (2025)
von: Ghassemi, Babak, et al.
Veröffentlicht: (2025)
Extracting Symbolic Sequences from Visual Representations via Self-Supervised Learning
von: Pozos, Victor Sebastian Martinez, et al.
Veröffentlicht: (2025)
von: Pozos, Victor Sebastian Martinez, et al.
Veröffentlicht: (2025)
Hyperbolic Coarse-to-Fine Few-Shot Class-Incremental Learning
von: Dai, Jiaxin, et al.
Veröffentlicht: (2025)
von: Dai, Jiaxin, et al.
Veröffentlicht: (2025)
Lang3D-XL: Language Embedded 3D Gaussians for Large-scale Scenes
von: Krakovsky, Shai, et al.
Veröffentlicht: (2025)
von: Krakovsky, Shai, et al.
Veröffentlicht: (2025)
Compositional Adversarial Training for Robust Visual Watermarking
von: Satheesh, Anirudh, et al.
Veröffentlicht: (2026)
von: Satheesh, Anirudh, et al.
Veröffentlicht: (2026)
Contrastive Learning to Fine-Tune Feature Extraction Models for the Visual Cortex
von: Mulrooney, Alex, et al.
Veröffentlicht: (2024)
von: Mulrooney, Alex, et al.
Veröffentlicht: (2024)
Robust Data Clustering with Outliers via Transformed Tensor Low-Rank Representation
von: Wu, Tong
Veröffentlicht: (2023)
von: Wu, Tong
Veröffentlicht: (2023)
PhysConvex: Physics-Informed 3D Dynamic Convex Radiance Fields for Reconstruction and Simulation
von: Wang, Dan, et al.
Veröffentlicht: (2026)
von: Wang, Dan, et al.
Veröffentlicht: (2026)
CoatFusion: Controllable Material Coating in Images
von: Levy, Sagie, et al.
Veröffentlicht: (2025)
von: Levy, Sagie, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Assessing Neural Network Robustness via Adversarial Pivotal Tuning
von: Christensen, Peter Ebert, et al.
Veröffentlicht: (2022) -
Designing a Conditional Prior Distribution for Flow-Based Generative Models
von: Issachar, Noam, et al.
Veröffentlicht: (2025) -
LoQT: Low-Rank Adapters for Quantized Pretraining
von: Loeschcke, Sebastian, et al.
Veröffentlicht: (2024) -
Generating Intermediate Representations for Compositional Text-To-Image Generation
von: Galun, Ran, et al.
Veröffentlicht: (2024) -
LaMI: Augmenting Large Language Models via Late Multi-Image Fusion
von: Yariv, Guy, et al.
Veröffentlicht: (2024)