Swiss DINO: Efficient and Versatile Vision Framework for On-device Personal Object Search
Fuente:
arXiv
Saved in:
| Main Authors: | Paramonov, Kirill, Zhong, Jia-Xing, Michieli, Umberto, Moon, Jijoong, Ozay, Mete |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Object-conditioned Bag of Instances for Few-Shot Personalized Instance Recognition
by: Michieli, Umberto, et al.
Published: (2024)
by: Michieli, Umberto, et al.
Published: (2024)
Controllable Forgetting Mechanism for Few-Shot Class-Incremental Learning
by: Paramonov, Kirill, et al.
Published: (2025)
by: Paramonov, Kirill, et al.
Published: (2025)
Cross-Architecture Auxiliary Feature Space Translation for Efficient Few-Shot Personalized Object Detection
by: Barbato, Francesco, et al.
Published: (2024)
by: Barbato, Francesco, et al.
Published: (2024)
FFT-based Selection and Optimization of Statistics for Robust Recognition of Severely Corrupted Images
by: Camuffo, Elena, et al.
Published: (2024)
by: Camuffo, Elena, et al.
Published: (2024)
Enhanced Model Robustness to Input Corruptions by Per-corruption Adaptation of Normalization Statistics
by: Camuffo, Elena, et al.
Published: (2024)
by: Camuffo, Elena, et al.
Published: (2024)
Feature-Space Generative Models for One-Shot Class-Incremental Learning
by: Foster, Jack, et al.
Published: (2026)
by: Foster, Jack, et al.
Published: (2026)
MOCHA: Multi-modal Objects-aware Cross-arcHitecture Alignment
by: Camuffo, Elena, et al.
Published: (2025)
by: Camuffo, Elena, et al.
Published: (2025)
DreamCache: Finetuning-Free Lightweight Personalized Image Generation via Feature Caching
by: Aiello, Emanuele, et al.
Published: (2024)
by: Aiello, Emanuele, et al.
Published: (2024)
Continual Error Correction on Low-Resource Devices
by: Paramonov, Kirill, et al.
Published: (2025)
by: Paramonov, Kirill, et al.
Published: (2025)
A Modular System for Enhanced Robustness of Multimedia Understanding Networks via Deep Parametric Estimation
by: Barbato, Francesco, et al.
Published: (2024)
by: Barbato, Francesco, et al.
Published: (2024)
Deep Neural Network Models Trained With A Fixed Random Classifier Transfer Better Across Domains
by: Ali, Hafiz Tiomoko, et al.
Published: (2024)
by: Ali, Hafiz Tiomoko, et al.
Published: (2024)
LoRA.rar: Learning to Merge LoRAs via Hypernetworks for Subject-Style Conditioned Image Generation
by: Shenaj, Donald, et al.
Published: (2024)
by: Shenaj, Donald, et al.
Published: (2024)
Efficient Compositional Multi-tasking for On-device Large Language Models
by: Bohdal, Ondrej, et al.
Published: (2025)
by: Bohdal, Ondrej, et al.
Published: (2025)
MemLoRA: Distilling Expert Adapters for On-Device Memory Systems
by: Bini, Massimo, et al.
Published: (2025)
by: Bini, Massimo, et al.
Published: (2025)
DINO Pre-training for Vision-based End-to-end Autonomous Driving
by: Juneja, Shubham, et al.
Published: (2024)
by: Juneja, Shubham, et al.
Published: (2024)
DINO-VO: Learning Where to Focus for Enhanced State Estimation
by: Chen, Qi, et al.
Published: (2026)
by: Chen, Qi, et al.
Published: (2026)
DINO-SD: Champion Solution for ICRA 2024 RoboDepth Challenge
by: Mao, Yifan, et al.
Published: (2024)
by: Mao, Yifan, et al.
Published: (2024)
DINO-CVA: A Multimodal Goal-Conditioned Vision-to-Action Model for Autonomous Catheter Navigation
by: Fekri, Pedram, et al.
Published: (2025)
by: Fekri, Pedram, et al.
Published: (2025)
DINO-Explorer: Active Underwater Discovery via Ego-Motion Compensated Semantic Predictive Coding
by: Jin, Yuhan, et al.
Published: (2026)
by: Jin, Yuhan, et al.
Published: (2026)
VLNVerse: A Benchmark for Vision-Language Navigation with Versatile, Embodied, Realistic Simulation and Evaluation
by: Lin, Sihao, et al.
Published: (2025)
by: Lin, Sihao, et al.
Published: (2025)
Personalized Embodied Navigation for Portable Object Finding
by: Dorbala, Vishnu Sashank, et al.
Published: (2024)
by: Dorbala, Vishnu Sashank, et al.
Published: (2024)
On Extending Semantic Abstraction for Efficient Search of Hidden Objects
by: Pais, Tasha, et al.
Published: (2025)
by: Pais, Tasha, et al.
Published: (2025)
HOP to the Next Tasks and Domains for Continual Learning in NLP
by: Michieli, Umberto, et al.
Published: (2024)
by: Michieli, Umberto, et al.
Published: (2024)
D$^3$FlowSLAM: Self-Supervised Dynamic SLAM with Flow Motion Decomposition and DINO Guidance
by: Yu, Xingyuan, et al.
Published: (2022)
by: Yu, Xingyuan, et al.
Published: (2022)
Data-driven Clustering and Merging of Adapters for On-device Large Language Models
by: Bohdal, Ondrej, et al.
Published: (2026)
by: Bohdal, Ondrej, et al.
Published: (2026)
WildOS: Open-Vocabulary Object Search in the Wild
by: Shah, Hardik, et al.
Published: (2026)
by: Shah, Hardik, et al.
Published: (2026)
Stop Wandering: Efficient Vision-Language Navigation via Metacognitive Reasoning
by: Li, Xueying, et al.
Published: (2026)
by: Li, Xueying, et al.
Published: (2026)
Personalized Instance-based Navigation Toward User-Specific Objects in Realistic Environments
by: Barsellotti, Luca, et al.
Published: (2024)
by: Barsellotti, Luca, et al.
Published: (2024)
BOP-ASK: Object-Interaction Reasoning for Vision-Language Models
by: Bhat, Vineet, et al.
Published: (2025)
by: Bhat, Vineet, et al.
Published: (2025)
Understanding Spatio-Temporal Relations in Human-Object Interaction using Pyramid Graph Convolutional Network
by: Xing, Hao, et al.
Published: (2024)
by: Xing, Hao, et al.
Published: (2024)
PNAS-MOT: Multi-Modal Object Tracking with Pareto Neural Architecture Search
by: Peng, Chensheng, et al.
Published: (2024)
by: Peng, Chensheng, et al.
Published: (2024)
A Model for Every User and Budget: Label-Free and Personalized Mixed-Precision Quantization
by: Fish, Edward, et al.
Published: (2023)
by: Fish, Edward, et al.
Published: (2023)
DegustaBot: Zero-Shot Visual Preference Estimation for Personalized Multi-Object Rearrangement
by: Newman, Benjamin A., et al.
Published: (2024)
by: Newman, Benjamin A., et al.
Published: (2024)
Contrastive Learning for Enhancing Robust Scene Transfer in Vision-based Agile Flight
by: Xing, Jiaxu, et al.
Published: (2023)
by: Xing, Jiaxu, et al.
Published: (2023)
Benchmarking Vision-Based Object Tracking for USVs in Complex Maritime Environments
by: Din, Muhayy Ud, et al.
Published: (2024)
by: Din, Muhayy Ud, et al.
Published: (2024)
DOPE: Dual Object Perception-Enhancement Network for Vision-and-Language Navigation
by: Yu, Yinfeng, et al.
Published: (2025)
by: Yu, Yinfeng, et al.
Published: (2025)
Vision-Based Autonomous UAV Navigation and Landing for Urban Search and Rescue
by: Mittal, Mayank, et al.
Published: (2019)
by: Mittal, Mayank, et al.
Published: (2019)
MonoDINO-DETR: Depth-Enhanced Monocular 3D Object Detection Using a Vision Foundation Model
by: Kim, Jihyeok, et al.
Published: (2025)
by: Kim, Jihyeok, et al.
Published: (2025)
Object-Scene-Camera Decomposition and Recomposition for Data-Efficient Monocular 3D Object Detection
by: Kuang, Zhaonian, et al.
Published: (2026)
by: Kuang, Zhaonian, et al.
Published: (2026)
WMNav: Integrating Vision-Language Models into World Models for Object Goal Navigation
by: Nie, Dujun, et al.
Published: (2025)
by: Nie, Dujun, et al.
Published: (2025)
Similar Items
-
Object-conditioned Bag of Instances for Few-Shot Personalized Instance Recognition
by: Michieli, Umberto, et al.
Published: (2024) -
Controllable Forgetting Mechanism for Few-Shot Class-Incremental Learning
by: Paramonov, Kirill, et al.
Published: (2025) -
Cross-Architecture Auxiliary Feature Space Translation for Efficient Few-Shot Personalized Object Detection
by: Barbato, Francesco, et al.
Published: (2024) -
FFT-based Selection and Optimization of Statistics for Robust Recognition of Severely Corrupted Images
by: Camuffo, Elena, et al.
Published: (2024) -
Enhanced Model Robustness to Input Corruptions by Per-corruption Adaptation of Normalization Statistics
by: Camuffo, Elena, et al.
Published: (2024)