Geometry Fidelity for Spherical Images
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Christensen, Anders, Mojab, Nooshin, Patel, Khushman, Ahuja, Karan, Akata, Zeynep, Winther, Ole, Gonzalez-Franco, Mar, Colaco, Andrea |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DiffEnc: Variational Diffusion with a Learned Encoder
von: Nielsen, Beatrix M. G., et al.
Veröffentlicht: (2023)
von: Nielsen, Beatrix M. G., et al.
Veröffentlicht: (2023)
SurfaceXR: Fusing Smartwatch IMUs and Egocentric Hand Pose for Seamless Surface Interactions
von: Xu, Vasco, et al.
Veröffentlicht: (2026)
von: Xu, Vasco, et al.
Veröffentlicht: (2026)
Sparse Autoencoders are Topic Models
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
VGGRPO: Towards World-Consistent Video Generation with 4D Latent Reward
von: An, Zhaochong, et al.
Veröffentlicht: (2026)
von: An, Zhaochong, et al.
Veröffentlicht: (2026)
Explaining CLIP Zero-shot Predictions Through Concepts
von: Ozdemir, Onat, et al.
Veröffentlicht: (2026)
von: Ozdemir, Onat, et al.
Veröffentlicht: (2026)
DeLoRA: Decoupling Angles and Strength in Low-rank Adaptation
von: Bini, Massimo, et al.
Veröffentlicht: (2025)
von: Bini, Massimo, et al.
Veröffentlicht: (2025)
From Drop-off to Recovery: A Mechanistic Analysis of Segmentation in MLLMs
von: Wu, Boyong, et al.
Veröffentlicht: (2026)
von: Wu, Boyong, et al.
Veröffentlicht: (2026)
The Manifold Hypothesis for Gradient-Based Explanations
von: Bordt, Sebastian, et al.
Veröffentlicht: (2022)
von: Bordt, Sebastian, et al.
Veröffentlicht: (2022)
Diffuse, Attend, and Segment: Unsupervised Zero-Shot Segmentation using Stable Diffusion
von: Tian, Junjiao, et al.
Veröffentlicht: (2023)
von: Tian, Junjiao, et al.
Veröffentlicht: (2023)
Noise Hypernetworks: Amortizing Test-Time Compute in Diffusion Models
von: Eyring, Luca, et al.
Veröffentlicht: (2025)
von: Eyring, Luca, et al.
Veröffentlicht: (2025)
SUB: Benchmarking CBM Generalization via Synthetic Attribute Substitutions
von: Bader, Jessica, et al.
Veröffentlicht: (2025)
von: Bader, Jessica, et al.
Veröffentlicht: (2025)
ETHER: Efficient Finetuning of Large-Scale Models with Hyperplane Reflections
von: Bini, Massimo, et al.
Veröffentlicht: (2024)
von: Bini, Massimo, et al.
Veröffentlicht: (2024)
Vision-by-Language for Training-Free Compositional Image Retrieval
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2023)
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2023)
Vision Transformers Exhibit Human-Like Biases: Evidence of Orientation and Color Selectivity, Categorical Perception, and Phase Transitions
von: Bahador, Nooshin
Veröffentlicht: (2025)
von: Bahador, Nooshin
Veröffentlicht: (2025)
Reflecting on the State of Rehearsal-free Continual Learning with Pretrained Models
von: Thede, Lukas, et al.
Veröffentlicht: (2024)
von: Thede, Lukas, et al.
Veröffentlicht: (2024)
DataDream: Few-shot Guided Dataset Generation
von: Kim, Jae Myung, et al.
Veröffentlicht: (2024)
von: Kim, Jae Myung, et al.
Veröffentlicht: (2024)
Practical and Rich User Digitization
von: Ahuja, Karan
Veröffentlicht: (2024)
von: Ahuja, Karan
Veröffentlicht: (2024)
Improving Intervention Efficacy via Concept Realignment in Concept Bottleneck Models
von: Singhi, Nishad, et al.
Veröffentlicht: (2024)
von: Singhi, Nishad, et al.
Veröffentlicht: (2024)
Disentangled Representation Learning with the Gromov-Monge Gap
von: Uscidda, Théo, et al.
Veröffentlicht: (2024)
von: Uscidda, Théo, et al.
Veröffentlicht: (2024)
Time Series Representations for Classification Lie Hidden in Pretrained Vision Transformers
von: Roschmann, Simon, et al.
Veröffentlicht: (2025)
von: Roschmann, Simon, et al.
Veröffentlicht: (2025)
Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
von: Pach, Mateusz, et al.
Veröffentlicht: (2025)
von: Pach, Mateusz, et al.
Veröffentlicht: (2025)
The Latent Color Subspace: Emergent Order in High-Dimensional Chaos
von: Pach, Mateusz, et al.
Veröffentlicht: (2026)
von: Pach, Mateusz, et al.
Veröffentlicht: (2026)
Fantastic Gains and Where to Find Them: On the Existence and Prospect of General Knowledge Transfer between Any Pretrained Model
von: Roth, Karsten, et al.
Veröffentlicht: (2023)
von: Roth, Karsten, et al.
Veröffentlicht: (2023)
A Large Scale Analysis of Gender Biases in Text-to-Image Generative Models
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
ReNO: Enhancing One-step Text-to-Image Models through Reward-based Noise Optimization
von: Eyring, Luca, et al.
Veröffentlicht: (2024)
von: Eyring, Luca, et al.
Veröffentlicht: (2024)
Context-Aware Multimodal Pretraining
von: Roth, Karsten, et al.
Veröffentlicht: (2024)
von: Roth, Karsten, et al.
Veröffentlicht: (2024)
Stitch: Training-Free Position Control in Multimodal Diffusion Transformers
von: Bader, Jessica, et al.
Veröffentlicht: (2025)
von: Bader, Jessica, et al.
Veröffentlicht: (2025)
Fast Sphericity and Roundness approximation in 2D and 3D using Local Thickness
von: Pieta, Pawel Tomasz, et al.
Veröffentlicht: (2025)
von: Pieta, Pawel Tomasz, et al.
Veröffentlicht: (2025)
Concept-Guided Interpretability via Neural Chunking
von: Wu, Shuchen, et al.
Veröffentlicht: (2025)
von: Wu, Shuchen, et al.
Veröffentlicht: (2025)
MemLoRA: Distilling Expert Adapters for On-Device Memory Systems
von: Bini, Massimo, et al.
Veröffentlicht: (2025)
von: Bini, Massimo, et al.
Veröffentlicht: (2025)
Scalable Ranked Preference Optimization for Text-to-Image Generation
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2024)
von: Karthik, Shyamgopal, et al.
Veröffentlicht: (2024)
Post-hoc Probabilistic Vision-Language Models
von: Baumann, Anton, et al.
Veröffentlicht: (2024)
von: Baumann, Anton, et al.
Veröffentlicht: (2024)
Person-Centric Annotations of LAION-400M: Auditing Bias and Its Transfer to Models
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
von: Girrbach, Leander, et al.
Veröffentlicht: (2025)
EgoCVR: An Egocentric Benchmark for Fine-Grained Composed Video Retrieval
von: Hummel, Thomas, et al.
Veröffentlicht: (2024)
von: Hummel, Thomas, et al.
Veröffentlicht: (2024)
LoFT: LoRA-fused Training Dataset Generation with Few-shot Guidance
von: Kim, Jae Myung, et al.
Veröffentlicht: (2025)
von: Kim, Jae Myung, et al.
Veröffentlicht: (2025)
COSMOS: Cross-Modality Self-Distillation for Vision Language Pre-training
von: Kim, Sanghwan, et al.
Veröffentlicht: (2024)
von: Kim, Sanghwan, et al.
Veröffentlicht: (2024)
How to Merge Your Multimodal Models Over Time?
von: Dziadzio, Sebastian, et al.
Veröffentlicht: (2024)
von: Dziadzio, Sebastian, et al.
Veröffentlicht: (2024)
Unbalancedness in Neural Monge Maps Improves Unpaired Domain Translation
von: Eyring, Luca, et al.
Veröffentlicht: (2023)
von: Eyring, Luca, et al.
Veröffentlicht: (2023)
FLAIR: VLM with Fine-grained Language-informed Image Representations
von: Xiao, Rui, et al.
Veröffentlicht: (2024)
von: Xiao, Rui, et al.
Veröffentlicht: (2024)
SphereDrag: Spherical Geometry-Aware Panoramic Image Editing
von: Feng, Zhiao, et al.
Veröffentlicht: (2025)
von: Feng, Zhiao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DiffEnc: Variational Diffusion with a Learned Encoder
von: Nielsen, Beatrix M. G., et al.
Veröffentlicht: (2023) -
SurfaceXR: Fusing Smartwatch IMUs and Egocentric Hand Pose for Seamless Surface Interactions
von: Xu, Vasco, et al.
Veröffentlicht: (2026) -
Sparse Autoencoders are Topic Models
von: Girrbach, Leander, et al.
Veröffentlicht: (2025) -
VGGRPO: Towards World-Consistent Video Generation with 4D Latent Reward
von: An, Zhaochong, et al.
Veröffentlicht: (2026) -
Explaining CLIP Zero-shot Predictions Through Concepts
von: Ozdemir, Onat, et al.
Veröffentlicht: (2026)