Guardado en:
| Autores principales: | An, Zhaochong, Kupyn, Orest, Uscidda, Théo, Colaco, Andrea, Ahuja, Karan, Belongie, Serge, Gonzalez-Franco, Mar, Gazulla, Marta Tintore |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2603.26599 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Dataset Enhancement with Instance-Level Augmentations
por: Kupyn, Orest, et al.
Publicado: (2024)
por: Kupyn, Orest, et al.
Publicado: (2024)
S3OD: Towards Generalizable Salient Object Detection with Synthetic Data
por: Kupyn, Orest, et al.
Publicado: (2025)
por: Kupyn, Orest, et al.
Publicado: (2025)
VGGHeads: 3D Multi Head Alignment with a Large-Scale Synthetic Dataset
por: Kupyn, Orest, et al.
Publicado: (2024)
por: Kupyn, Orest, et al.
Publicado: (2024)
Epipolar Geometry Improves Video Generation Models
por: Kupyn, Orest, et al.
Publicado: (2025)
por: Kupyn, Orest, et al.
Publicado: (2025)
SurfaceXR: Fusing Smartwatch IMUs and Egocentric Hand Pose for Seamless Surface Interactions
por: Xu, Vasco, et al.
Publicado: (2026)
por: Xu, Vasco, et al.
Publicado: (2026)
Geometry Fidelity for Spherical Images
por: Christensen, Anders, et al.
Publicado: (2024)
por: Christensen, Anders, et al.
Publicado: (2024)
Augmented Object Intelligence with XR-Objects
por: Dogan, Mustafa Doga, et al.
Publicado: (2024)
por: Dogan, Mustafa Doga, et al.
Publicado: (2024)
MOBIUS: Big-to-Mobile Universal Instance Segmentation via Multi-modal Bottleneck Fusion and Calibrated Decoder Pruning
por: Segu, Mattia, et al.
Publicado: (2025)
por: Segu, Mattia, et al.
Publicado: (2025)
Generalized Few-shot 3D Point Cloud Segmentation with Vision-Language Model
por: An, Zhaochong, et al.
Publicado: (2025)
por: An, Zhaochong, et al.
Publicado: (2025)
PARSE-Ego4D: Personal Action Recommendation Suggestions for Egocentric Videos
por: Abreu, Steven, et al.
Publicado: (2024)
por: Abreu, Steven, et al.
Publicado: (2024)
DAD-3DHeads: A Large-scale Dense, Accurate and Diverse Dataset for 3D Head Alignment from a Single Image
por: Martyniuk, Tetiana, et al.
Publicado: (2022)
por: Martyniuk, Tetiana, et al.
Publicado: (2022)
Rethinking Few-shot 3D Point Cloud Semantic Segmentation
por: An, Zhaochong, et al.
Publicado: (2024)
por: An, Zhaochong, et al.
Publicado: (2024)
Multimodality Helps Few-shot 3D Point Cloud Semantic Segmentation
por: An, Zhaochong, et al.
Publicado: (2024)
por: An, Zhaochong, et al.
Publicado: (2024)
Shifting the Breaking Point of Flow Matching for Multi-Instance Editing
por: Zaccagnino, Carmine, et al.
Publicado: (2026)
por: Zaccagnino, Carmine, et al.
Publicado: (2026)
GeOT: A spatially explicit framework for evaluating spatio-temporal predictions
por: Wiedemann, Nina, et al.
Publicado: (2024)
por: Wiedemann, Nina, et al.
Publicado: (2024)
ChatMotion: A Multimodal Multi-Agent for Human Motion Analysis
por: Li, Lei, et al.
Publicado: (2025)
por: Li, Lei, et al.
Publicado: (2025)
Revisiting the Perception-Distortion Trade-off with Spatial-Semantic Guided Super-Resolution
por: Wang, Dan, et al.
Publicado: (2026)
por: Wang, Dan, et al.
Publicado: (2026)
PoseDreamer: Scalable and Photorealistic Human Data Generation Pipeline with Diffusion Models
por: Prospero, Lorenza, et al.
Publicado: (2026)
por: Prospero, Lorenza, et al.
Publicado: (2026)
Symbiotic AI: Augmenting Human Cognition from PCs to Cars
por: Bovo, Riccardo, et al.
Publicado: (2025)
por: Bovo, Riccardo, et al.
Publicado: (2025)
Diffuse, Attend, and Segment: Unsupervised Zero-Shot Segmentation using Stable Diffusion
por: Tian, Junjiao, et al.
Publicado: (2023)
por: Tian, Junjiao, et al.
Publicado: (2023)
Practical and Rich User Digitization
por: Ahuja, Karan
Publicado: (2024)
por: Ahuja, Karan
Publicado: (2024)
EmBARDiment: an Embodied AI Agent for Productivity in XR
por: Bovo, Riccardo, et al.
Publicado: (2024)
por: Bovo, Riccardo, et al.
Publicado: (2024)
The Latent Color Subspace: Emergent Order in High-Dimensional Chaos
por: Pach, Mateusz, et al.
Publicado: (2026)
por: Pach, Mateusz, et al.
Publicado: (2026)
Generalized Discrete Diffusion from Snapshots
por: Zekri, Oussama, et al.
Publicado: (2026)
por: Zekri, Oussama, et al.
Publicado: (2026)
GENOT: Entropic (Gromov) Wasserstein Flow Matching with Applications to Single-Cell Genomics
por: Klein, Dominik, et al.
Publicado: (2023)
por: Klein, Dominik, et al.
Publicado: (2023)
Thinking in Frames: How Visual Context and Test-Time Scaling Empower Video Reasoning
por: Li, Chengzu, et al.
Publicado: (2026)
por: Li, Chengzu, et al.
Publicado: (2026)
Noise-Coded Illumination for Forensic and Photometric Video Analysis
por: Michael, Peter F., et al.
Publicado: (2025)
por: Michael, Peter F., et al.
Publicado: (2025)
Video Understanding: From Geometry and Semantics to Unified Models
por: An, Zhaochong, et al.
Publicado: (2026)
por: An, Zhaochong, et al.
Publicado: (2026)
PhysConvex: Physics-Informed 3D Dynamic Convex Radiance Fields for Reconstruction and Simulation
por: Wang, Dan, et al.
Publicado: (2026)
por: Wang, Dan, et al.
Publicado: (2026)
Text Entry for XR Trove (TEXT): Collecting and Analyzing Techniques for Text Input in XR
por: Bhatia, Arpit, et al.
Publicado: (2025)
por: Bhatia, Arpit, et al.
Publicado: (2025)
Unlearning-based Neural Interpretations
por: Choi, Ching Lam, et al.
Publicado: (2024)
por: Choi, Ching Lam, et al.
Publicado: (2024)
Large Vision-Language Models for Knowledge-Grounded Data Annotation of Memes
por: Deng, Shiling, et al.
Publicado: (2025)
por: Deng, Shiling, et al.
Publicado: (2025)
Panel on Synthesizing BPs .
por: Tintore, Joaquin
Publicado: (2019)
por: Tintore, Joaquin
Publicado: (2019)
Real-time surface current data in the Ibiza Channel from January to October 2016
por: Tintore, Joaquín
Publicado: (2016)
por: Tintore, Joaquín
Publicado: (2016)
Stitched Value Model for Diffusion Alignment
por: Go, Hyojun, et al.
Publicado: (2026)
por: Go, Hyojun, et al.
Publicado: (2026)
OneStory: Coherent Multi-Shot Video Generation with Adaptive Memory
por: An, Zhaochong, et al.
Publicado: (2025)
por: An, Zhaochong, et al.
Publicado: (2025)
MMEarth-Bench: Global Model Adaptation via Multimodal Test-Time Training
por: Gordon, Lucia, et al.
Publicado: (2026)
por: Gordon, Lucia, et al.
Publicado: (2026)
Assessing Neural Network Robustness via Adversarial Pivotal Tuning
por: Christensen, Peter Ebert, et al.
Publicado: (2022)
por: Christensen, Peter Ebert, et al.
Publicado: (2022)
From Videos to Conversations: Egocentric Instructions for Task Assistance
por: Aggarwal, Lavisha, et al.
Publicado: (2026)
por: Aggarwal, Lavisha, et al.
Publicado: (2026)
Disentangled Representation Learning with the Gromov-Monge Gap
por: Uscidda, Théo, et al.
Publicado: (2024)
por: Uscidda, Théo, et al.
Publicado: (2024)
Ejemplares similares
-
Dataset Enhancement with Instance-Level Augmentations
por: Kupyn, Orest, et al.
Publicado: (2024) -
S3OD: Towards Generalizable Salient Object Detection with Synthetic Data
por: Kupyn, Orest, et al.
Publicado: (2025) -
VGGHeads: 3D Multi Head Alignment with a Large-Scale Synthetic Dataset
por: Kupyn, Orest, et al.
Publicado: (2024) -
Epipolar Geometry Improves Video Generation Models
por: Kupyn, Orest, et al.
Publicado: (2025) -
SurfaceXR: Fusing Smartwatch IMUs and Egocentric Hand Pose for Seamless Surface Interactions
por: Xu, Vasco, et al.
Publicado: (2026)