Enregistré dans:
| Auteurs principaux: | Cao, Ang, Arnaud, Sergio, Maksymets, Oleksandr, Yang, Jianing, Jain, Ayush, Yenamandra, Sriram, Martin, Ada, Berges, Vincent-Pierre, McVay, Paul, Partsey, Ruslan, Rajeswaran, Aravind, Meier, Franziska, Johnson, Justin, Park, Jeong Joon, Sax, Alexander |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | https://arxiv.org/abs/2502.20389 |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Locate 3D: Real-World Object Localization via Self-Supervised Learning in 3D
par: Arnaud, Sergio, et autres
Publié: (2025)
par: Arnaud, Sergio, et autres
Publié: (2025)
What do we learn from a large-scale study of pre-trained visual representations in sim and real environments?
par: Silwal, Sneha, et autres
Publié: (2023)
par: Silwal, Sneha, et autres
Publié: (2023)
Unifying 2D and 3D Vision-Language Understanding
par: Jain, Ayush, et autres
Publié: (2025)
par: Jain, Ayush, et autres
Publié: (2025)
Where are we in the search for an Artificial Visual Cortex for Embodied Intelligence?
par: Majumdar, Arjun, et autres
Publié: (2023)
par: Majumdar, Arjun, et autres
Publié: (2023)
An Exploratory Study of Undergraduate Students' Perceptions of Visualization and Visualization Ability in Biochemistry
par: Andrew McVay, et autres
Publié: (2026)
par: Andrew McVay, et autres
Publié: (2026)
On Linear Separability under Linear Compression with Applications to Hard Support Vector Machine
par: McVay, Paul, et autres
Publié: (2022)
par: McVay, Paul, et autres
Publié: (2022)
HomeRobot Open Vocabulary Mobile Manipulation Challenge 2023 Participant Report (Team KuzHum)
par: Kuzma, Volodymyr, et autres
Publié: (2024)
par: Kuzma, Volodymyr, et autres
Publié: (2024)
Fast3R: Towards 3D Reconstruction of 1000+ Images in One Forward Pass
par: Yang, Jianing, et autres
Publié: (2025)
par: Yang, Jianing, et autres
Publié: (2025)
From LLMs to Actions: Latent Codes as Bridges in Hierarchical Robot Control
par: Shentu, Yide, et autres
Publié: (2024)
par: Shentu, Yide, et autres
Publié: (2024)
MoDem-V2: Visuo-Motor World Models for Real-World Robot Manipulation
par: Lancaster, Patrick, et autres
Publié: (2023)
par: Lancaster, Patrick, et autres
Publié: (2023)
RetinaGS: Scalable Training for Dense Scene Rendering with Billion-Scale 3D Gaussians
par: Li, Bingling, et autres
Publié: (2024)
par: Li, Bingling, et autres
Publié: (2024)
Exploring potential reach and representativeness of a self‐weighing weight gain prevention intervention in adults with overweight and obesity
par: Kellie B. Scotti, et autres
Publié: (2024)
par: Kellie B. Scotti, et autres
Publié: (2024)
Are VLMs Really Blind
par: Singh, Ayush, et autres
Publié: (2024)
par: Singh, Ayush, et autres
Publié: (2024)
Semi-Supervised One-Shot Imitation Learning
par: Wu, Philipp, et autres
Publié: (2024)
par: Wu, Philipp, et autres
Publié: (2024)
Exploring the Acceptability of the STOP Method for Addressing Weight Loss Misinformation on Social Media: An Interview Study
par: Danielle E. Jake‐Schoffman, et autres
Publié: (2025)
par: Danielle E. Jake‐Schoffman, et autres
Publié: (2025)
An Image Is Worth Ten Thousand Words: Verbose-Text Induction Attacks on VLMs
par: Luo, Zhi, et autres
Publié: (2025)
par: Luo, Zhi, et autres
Publié: (2025)
On Place, Well-Being, and Illness in the Andes
par: Marieka Sax
Publié: (2015)
par: Marieka Sax
Publié: (2015)
The Boy Problem: Many Boys Think School Is Stupid and Reading Stinks--Is There a Remedy?
par: Sax, Leonard
Publié: (2007)
par: Sax, Leonard
Publié: (2007)
Efficient Camera-Controlled Video Generation of Static Scenes via Sparse Diffusion and 3D Rendering
par: Chen, Jieying, et autres
Publié: (2026)
par: Chen, Jieying, et autres
Publié: (2026)
MoE3D: A Mixture-of-Experts Module for 3D Reconstruction
par: Wang, Zichen, et autres
Publié: (2026)
par: Wang, Zichen, et autres
Publié: (2026)
TouchMap-OR: Multi-View 3D Mapping of Hand-Surface Contacts
par: Ktistakis, Sophokles, et autres
Publié: (2026)
par: Ktistakis, Sophokles, et autres
Publié: (2026)
The Green ring of a family of copointed Hopf algebras
par: Vay, Cristian
Publié: (2022)
par: Vay, Cristian
Publié: (2022)
Linkage principle for small quantum groups
par: Vay, Cristian
Publié: (2023)
par: Vay, Cristian
Publié: (2023)
Polyploidy in Enkianthus (Ericaceae)
par: Sax, Hally Jolivette
Publié: (1960)
par: Sax, Hally Jolivette
Publié: (1960)
Bias in the Picture: Benchmarking VLMs with Social-Cue News Images and LLM-as-Judge Assessment
par: Narayanan, Aravind, et autres
Publié: (2025)
par: Narayanan, Aravind, et autres
Publié: (2025)
Agents Play Thousands of 3D Video Games
par: Xu, Zhongwen, et autres
Publié: (2025)
par: Xu, Zhongwen, et autres
Publié: (2025)
Probing Visual Language Priors in VLMs
par: Luo, Tiange, et autres
Publié: (2024)
par: Luo, Tiange, et autres
Publié: (2024)
Chitrarth: Bridging Vision and Language for a Billion People
par: Khan, Shaharukh, et autres
Publié: (2025)
par: Khan, Shaharukh, et autres
Publié: (2025)
Importance of Developing Emotional Intelligence in Preventing Addiction Syndrome
par: Viktoriia Mendelo, et autres
Publié: (2024)
par: Viktoriia Mendelo, et autres
Publié: (2024)
Stimpack: An Adaptive Rendering Optimization System for Scalable Cloud Gaming
par: Heo, Jin, et autres
Publié: (2024)
par: Heo, Jin, et autres
Publié: (2024)
The Prediction of Training Proficiency in Firefighters: A Study of Predictive Validity in Spain
par: Alfredo Berges
Publié: (2018)
par: Alfredo Berges
Publié: (2018)
Justice and Righteousness in the Old Testament
par: Berges, Ulrich
Publié: (2025)
par: Berges, Ulrich
Publié: (2025)
Pattern-Based Phase-Separation of Tracer and Dispersed Phase Particles in Two-Phase Defocusing Particle Tracking Velocimetry
par: Sax, Christian, et autres
Publié: (2025)
par: Sax, Christian, et autres
Publié: (2025)
On the Particle Image Overlap in Single Camera Defocusing Approaches
par: Sax, Christian, et autres
Publié: (2025)
par: Sax, Christian, et autres
Publié: (2025)
Off-Diagonal Continuous Rado Numbers $x_1 + x_2 + \dots + x_k = x_0$
par: Vestal, Don, et autres
Publié: (2025)
par: Vestal, Don, et autres
Publié: (2025)
R4: Retrieval-Augmented Reasoning for Vision-Language Models in 4D Spatio-Temporal Space
par: Sohn, Tin Stribor, et autres
Publié: (2025)
par: Sohn, Tin Stribor, et autres
Publié: (2025)
Does Your 3D Encoder Really Work? When Pretrain-SFT from 2D VLMs Meets 3D VLMs
par: Li, Haoyuan, et autres
Publié: (2025)
par: Li, Haoyuan, et autres
Publié: (2025)
GOAT-Bench: A Benchmark for Multi-Modal Lifelong Navigation
par: Khanna, Mukul, et autres
Publié: (2024)
par: Khanna, Mukul, et autres
Publié: (2024)
GaussRender: Learning 3D Occupancy with Gaussian Rendering
par: Chambon, Loïck, et autres
Publié: (2025)
par: Chambon, Loïck, et autres
Publié: (2025)
RenderOcc: Vision-Centric 3D Occupancy Prediction with 2D Rendering Supervision
par: Pan, Mingjie, et autres
Publié: (2023)
par: Pan, Mingjie, et autres
Publié: (2023)
Documents similaires
-
Locate 3D: Real-World Object Localization via Self-Supervised Learning in 3D
par: Arnaud, Sergio, et autres
Publié: (2025) -
What do we learn from a large-scale study of pre-trained visual representations in sim and real environments?
par: Silwal, Sneha, et autres
Publié: (2023) -
Unifying 2D and 3D Vision-Language Understanding
par: Jain, Ayush, et autres
Publié: (2025) -
Where are we in the search for an Artificial Visual Cortex for Embodied Intelligence?
par: Majumdar, Arjun, et autres
Publié: (2023) -
An Exploratory Study of Undergraduate Students' Perceptions of Visualization and Visualization Ability in Biochemistry
par: Andrew McVay, et autres
Publié: (2026)