Universal representations:The missing link between faces, text, planktons, and cat breeds
Fuente:
arXiv
Saved in:
| Main Authors: | Bilen, Hakan, Vedaldi, Andrea |
|---|---|
| Format: | Preprint |
| Published: |
2017
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tracktention: Leveraging Point Tracking to Attend Videos Faster and Better
by: Lai, Zihang, et al.
Published: (2025)
by: Lai, Zihang, et al.
Published: (2025)
RenderDiffusion: Image Diffusion for 3D Reconstruction, Inpainting and Generation
by: Anciukevičius, Titas, et al.
Published: (2022)
by: Anciukevičius, Titas, et al.
Published: (2022)
HumMorph: Generalized Dynamic Human Neural Fields from Few Views
by: Zadrożny, Jakub, et al.
Published: (2025)
by: Zadrożny, Jakub, et al.
Published: (2025)
DSO: Aligning 3D Generators with Simulation Feedback for Physical Soundness
by: Li, Ruining, et al.
Published: (2025)
by: Li, Ruining, et al.
Published: (2025)
Learning segmentation from point trajectories
by: Karazija, Laurynas, et al.
Published: (2025)
by: Karazija, Laurynas, et al.
Published: (2025)
VFMF: World Modeling by Forecasting Vision Foundation Model Features
by: Boduljak, Gabrijel, et al.
Published: (2025)
by: Boduljak, Gabrijel, et al.
Published: (2025)
What Happens Next? Anticipating Future Motion by Generating Point Trajectories
by: Boduljak, Gabrijel, et al.
Published: (2025)
by: Boduljak, Gabrijel, et al.
Published: (2025)
Realiz3D: 3D Generation Made Photorealistic via Domain-Aware Learning
by: Sobol, Ido, et al.
Published: (2026)
by: Sobol, Ido, et al.
Published: (2026)
Beyond Pixel Histories: World Models with Persistent 3D State
by: Garcin, Samuel, et al.
Published: (2026)
by: Garcin, Samuel, et al.
Published: (2026)
PAOLI: Pose-free Articulated Object Learning from Sparse-view Images
by: Deng, Jianning, et al.
Published: (2025)
by: Deng, Jianning, et al.
Published: (2025)
Visually Interpretable Subtask Reasoning for Visual Question Answering
by: Cheng, Yu, et al.
Published: (2025)
by: Cheng, Yu, et al.
Published: (2025)
Odd-One-Out: Anomaly Detection by Comparing with Neighbors
by: Bhunia, Ankan, et al.
Published: (2024)
by: Bhunia, Ankan, et al.
Published: (2024)
Articulate your NeRF: Unsupervised articulated object modeling via conditional view synthesis
by: Deng, Jianning, et al.
Published: (2024)
by: Deng, Jianning, et al.
Published: (2024)
Looking 3D: Anomaly Detection with 2D-3D Alignment
by: Bhunia, Ankan, et al.
Published: (2024)
by: Bhunia, Ankan, et al.
Published: (2024)
Buffer replay enhances the robustness of multimodal learning under missing-modality
by: Zhu, Hongye, et al.
Published: (2025)
by: Zhu, Hongye, et al.
Published: (2025)
3D-Aware Instance Segmentation and Tracking in Egocentric Videos
by: Bhalgat, Yash, et al.
Published: (2024)
by: Bhalgat, Yash, et al.
Published: (2024)
LaVR: Scene Latent Conditioned Generative Video Trajectory Re-Rendering using Large 4D Reconstruction Models
by: Xie, Mingyang, et al.
Published: (2026)
by: Xie, Mingyang, et al.
Published: (2026)
GES: Generalized Exponential Splatting for Efficient Radiance Field Rendering
by: Hamdi, Abdullah, et al.
Published: (2024)
by: Hamdi, Abdullah, et al.
Published: (2024)
N2F2: Hierarchical Scene Understanding with Nested Neural Feature Fields
by: Bhalgat, Yash, et al.
Published: (2024)
by: Bhalgat, Yash, et al.
Published: (2024)
TurboEdit: Instant text-based image editing
by: Wu, Zongze, et al.
Published: (2024)
by: Wu, Zongze, et al.
Published: (2024)
Secure and reversible face anonymization with diffusion models
by: Labarbarie, Pol, et al.
Published: (2025)
by: Labarbarie, Pol, et al.
Published: (2025)
Improving Semantic Correspondence with Viewpoint-Guided Spherical Maps
by: Mariotti, Octave, et al.
Published: (2023)
by: Mariotti, Octave, et al.
Published: (2023)
CHAI: CacHe Attention Inference for text2video
by: Cherian, Joel Mathew, et al.
Published: (2026)
by: Cherian, Joel Mathew, et al.
Published: (2026)
IM-3D: Iterative Multiview Diffusion and Reconstruction for High-Quality 3D Generation
by: Melas-Kyriazi, Luke, et al.
Published: (2024)
by: Melas-Kyriazi, Luke, et al.
Published: (2024)
Towards a text-based quantitative and explainable histopathology image analysis
by: Nguyen, Anh Tien, et al.
Published: (2024)
by: Nguyen, Anh Tien, et al.
Published: (2024)
Improving face generation quality and prompt following with synthetic captions
by: Tarasiou, Michail, et al.
Published: (2024)
by: Tarasiou, Michail, et al.
Published: (2024)
Personalised aesthetics with residual adapters
by: Rodríguez-Pardo, Carlos, et al.
Published: (2019)
by: Rodríguez-Pardo, Carlos, et al.
Published: (2019)
$R_\text{dm}$: Re-conceptualizing Distribution Matching as a Reward for Diffusion Distillation
by: Fan, Linqian, et al.
Published: (2026)
by: Fan, Linqian, et al.
Published: (2026)
Meta 3D TextureGen: Fast and Consistent Texture Generation for 3D Objects
by: Bensadoun, Raphael, et al.
Published: (2024)
by: Bensadoun, Raphael, et al.
Published: (2024)
Spatially-Adaptive Hash Encodings For Neural Surface Reconstruction
by: Walker, Thomas, et al.
Published: (2024)
by: Walker, Thomas, et al.
Published: (2024)
Closing the gap in multimodal medical representation alignment
by: Grassucci, Eleonora, et al.
Published: (2026)
by: Grassucci, Eleonora, et al.
Published: (2026)
Spectral regularization for adversarially-robust representation learning
by: Yang, Sheng, et al.
Published: (2024)
by: Yang, Sheng, et al.
Published: (2024)
Disentangling representations of retinal images with generative models
by: Müller, Sarah, et al.
Published: (2024)
by: Müller, Sarah, et al.
Published: (2024)
Transferring disentangled representations: bridging the gap between synthetic and real images
by: Dapueto, Jacopo, et al.
Published: (2024)
by: Dapueto, Jacopo, et al.
Published: (2024)
Free3D: Consistent Novel View Synthesis without 3D Representation
by: Zheng, Chuanxia, et al.
Published: (2023)
by: Zheng, Chuanxia, et al.
Published: (2023)
Disentangled representations via score-based variational autoencoders
by: Lyo, Benjamin S. H., et al.
Published: (2025)
by: Lyo, Benjamin S. H., et al.
Published: (2025)
DeCLIP: Decoding CLIP representations for deepfake localization
by: Smeu, Stefan, et al.
Published: (2024)
by: Smeu, Stefan, et al.
Published: (2024)
A comparison between humans and AI at recognizing objects in unusual poses
by: Ollikka, Netta, et al.
Published: (2024)
by: Ollikka, Netta, et al.
Published: (2024)
A supervised discriminant data representation: application to pattern classification
by: Dornaika, Fadi, et al.
Published: (2025)
by: Dornaika, Fadi, et al.
Published: (2025)
Understanding normalization in contrastive representation learning and out-of-distribution detection
by: Le-Gia, Tai, et al.
Published: (2023)
by: Le-Gia, Tai, et al.
Published: (2023)
Similar Items
-
Tracktention: Leveraging Point Tracking to Attend Videos Faster and Better
by: Lai, Zihang, et al.
Published: (2025) -
RenderDiffusion: Image Diffusion for 3D Reconstruction, Inpainting and Generation
by: Anciukevičius, Titas, et al.
Published: (2022) -
HumMorph: Generalized Dynamic Human Neural Fields from Few Views
by: Zadrożny, Jakub, et al.
Published: (2025) -
DSO: Aligning 3D Generators with Simulation Feedback for Physical Soundness
by: Li, Ruining, et al.
Published: (2025) -
Learning segmentation from point trajectories
by: Karazija, Laurynas, et al.
Published: (2025)