Guardado en:
| Autores principales: | Douze, Matthijs, Guzhva, Alexandr, Deng, Chengqi, Johnson, Jeff, Szilvasy, Gergely, Mazaré, Pierre-Emmanuel, Lomeli, Maria, Hosseini, Lucas, Jégou, Hervé |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2401.08281 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Vector search with small radiuses
por: Szilvasy, Gergely, et al.
Publicado: (2024)
por: Szilvasy, Gergely, et al.
Publicado: (2024)
Inference-time sparse attention with asymmetric indexing
por: Mazaré, Pierre-Emmanuel, et al.
Publicado: (2025)
por: Mazaré, Pierre-Emmanuel, et al.
Publicado: (2025)
Self-Pruned Key-Value Attention: Learning When to Write by Predicting Future Utility
por: Szilvasy, Gergely, et al.
Publicado: (2026)
por: Szilvasy, Gergely, et al.
Publicado: (2026)
Short window attention enables long-term memorization
por: Cabannes, Loïc, et al.
Publicado: (2025)
por: Cabannes, Loïc, et al.
Publicado: (2025)
Stochastic activations
por: Lomeli, Maria, et al.
Publicado: (2025)
por: Lomeli, Maria, et al.
Publicado: (2025)
evclust: Python library for evidential clustering
por: Soubeiga, Armel, et al.
Publicado: (2025)
por: Soubeiga, Armel, et al.
Publicado: (2025)
Cross-Breed Pig Identification Using Auricular Vein Pattern Recognition: A Machine Learning Approach for Small-Scale Farming Applications
por: Nsengiyumvaa, Emmanuel, et al.
Publicado: (2025)
por: Nsengiyumvaa, Emmanuel, et al.
Publicado: (2025)
GUing: A Mobile GUI Search Engine using a Vision-Language Model
por: Wei, Jialiang, et al.
Publicado: (2024)
por: Wei, Jialiang, et al.
Publicado: (2024)
Can Vision-Language Models Handle Long-Context Code? An Empirical Study on Visual Compression
por: Zhong, Jianping, et al.
Publicado: (2026)
por: Zhong, Jianping, et al.
Publicado: (2026)
BLIP-FusePPO: A Vision-Language Deep Reinforcement Learning Framework for Lane Keeping in Autonomous Vehicles
por: Miangoleh, Seyed Ahmad Hosseini, et al.
Publicado: (2025)
por: Miangoleh, Seyed Ahmad Hosseini, et al.
Publicado: (2025)
Terrain characterisation for online adaptability of automated sonar processing: Lessons learnt from operationally applying ATR to sidescan sonar in MCM applications
por: Guerneve, Thomas, et al.
Publicado: (2024)
por: Guerneve, Thomas, et al.
Publicado: (2024)
SWAN -- Enabling Fast and Mobile Histopathology Image Annotation through Swipeable Interfaces
por: Banerjee, Sweta, et al.
Publicado: (2025)
por: Banerjee, Sweta, et al.
Publicado: (2025)
DD-CAM: Minimal Sufficient Explanations for Vision Models Using Delta Debugging
por: Khadka, Krishna, et al.
Publicado: (2026)
por: Khadka, Krishna, et al.
Publicado: (2026)
Foundation Models in Remote Sensing: Evolving from Unimodality to Multimodality
por: Hong, Danfeng, et al.
Publicado: (2026)
por: Hong, Danfeng, et al.
Publicado: (2026)
Technical Report for Argoverse2 Scenario Mining Challenges on Iterative Error Correction and Spatially-Aware Prompting
por: Chen, Yifei, et al.
Publicado: (2025)
por: Chen, Yifei, et al.
Publicado: (2025)
CUARewardBench: A Benchmark for Evaluating Reward Models on Computer-using Agent
por: Lin, Haojia, et al.
Publicado: (2025)
por: Lin, Haojia, et al.
Publicado: (2025)
How Far Can VLMs Go for Visual Bug Detection? Studying 19,738 Keyframes from 41 Hours of Gameplay Videos
por: Lu, Wentao, et al.
Publicado: (2026)
por: Lu, Wentao, et al.
Publicado: (2026)
Natural Adversaries: Fuzzing Autonomous Vehicles with Realistic Roadside Object Placements
por: Sun, Yang, et al.
Publicado: (2024)
por: Sun, Yang, et al.
Publicado: (2024)
Earth Embeddings as Products: Taxonomy, Ecosystem, and Standardized Access
por: Fang, Heng, et al.
Publicado: (2026)
por: Fang, Heng, et al.
Publicado: (2026)
Interpretable Gallbladder Ultrasound Diagnosis: A Lightweight Web-Mobile Software Platform with Real-Time XAI
por: Bhoyan, Fuyad Hasan, et al.
Publicado: (2025)
por: Bhoyan, Fuyad Hasan, et al.
Publicado: (2025)
What to Test Next: Interpretable Coverage Gap Discovery in Driving VLMs
por: Aich, Abhishek, et al.
Publicado: (2026)
por: Aich, Abhishek, et al.
Publicado: (2026)
ROMAN: Reward-Orchestrated Multi-Head Attention Network for Autonomous Driving System Testing
por: Chi, Jianlei, et al.
Publicado: (2026)
por: Chi, Jianlei, et al.
Publicado: (2026)
Effort-Optimized, Accuracy-Driven Labelling and Validation of Test Inputs for DL Systems: A Mixed-Integer Linear Programming Approach
por: Amini, Mohammad Hossein, et al.
Publicado: (2025)
por: Amini, Mohammad Hossein, et al.
Publicado: (2025)
Ear-Keeper: A Cross-Platform AI System for Rapid and Accurate Ear Disease Diagnosis
por: Lu, Feiyan, et al.
Publicado: (2023)
por: Lu, Feiyan, et al.
Publicado: (2023)
ClawMark: A Living-World Benchmark for Multi-Turn, Multi-Day, Multimodal Coworker Agents
por: Meng, Fanqing, et al.
Publicado: (2026)
por: Meng, Fanqing, et al.
Publicado: (2026)
Benchmarking Image Perturbations for Testing Automated Driving Assistance Systems
por: Lambertenghi, Stefano Carlo, et al.
Publicado: (2025)
por: Lambertenghi, Stefano Carlo, et al.
Publicado: (2025)
A Highly Efficient Diversity-based Input Selection for DNN Improvement Using VLMs
por: Abbasishahkoo, Amin, et al.
Publicado: (2026)
por: Abbasishahkoo, Amin, et al.
Publicado: (2026)
Do Existing Testing Tools Really Uncover Gender Bias in Text-to-Image Models?
por: Lyu, Yunbo, et al.
Publicado: (2025)
por: Lyu, Yunbo, et al.
Publicado: (2025)
Evaluating and Enhancing Segmentation Model Robustness with Metamorphic Testing
por: Mzoughi, Seif, et al.
Publicado: (2025)
por: Mzoughi, Seif, et al.
Publicado: (2025)
TigAug: Data Augmentation for Testing Traffic Light Detection in Autonomous Driving Systems
por: Lu, You, et al.
Publicado: (2025)
por: Lu, You, et al.
Publicado: (2025)
VEglue: Testing Visual Entailment Systems via Object-Aligned Joint Erasing
por: Chang, Zhiyuan, et al.
Publicado: (2024)
por: Chang, Zhiyuan, et al.
Publicado: (2024)
ARI3D: A Software for Interactive Quantification of Regions in X-Ray CT 3D Images
por: Albrecht, Jan Phillipp, et al.
Publicado: (2025)
por: Albrecht, Jan Phillipp, et al.
Publicado: (2025)
A Plausibility Study of Using Augmented Reality in the Ventriculoperitoneal Shunt Operations
por: Dorji, Tandin, et al.
Publicado: (2024)
por: Dorji, Tandin, et al.
Publicado: (2024)
VideoGameBunny: Towards vision assistants for video games
por: Taesiri, Mohammad Reza, et al.
Publicado: (2024)
por: Taesiri, Mohammad Reza, et al.
Publicado: (2024)
Open-source automatic pipeline for efficient conversion of large-scale point clouds to IFC format
por: Zbirovský, Slávek, et al.
Publicado: (2025)
por: Zbirovský, Slávek, et al.
Publicado: (2025)
DOne: Decoupling Structure and Rendering for High-Fidelity Design-to-Code Generation
por: Huang, Xinhao, et al.
Publicado: (2026)
por: Huang, Xinhao, et al.
Publicado: (2026)
Investigating Traffic Accident Detection Using Multimodal Large Language Models
por: Skender, Ilhan, et al.
Publicado: (2025)
por: Skender, Ilhan, et al.
Publicado: (2025)
A Retrieval-Augmented Generation Approach to Extracting Algorithmic Logic from Neural Networks
por: Khalid, Waleed, et al.
Publicado: (2025)
por: Khalid, Waleed, et al.
Publicado: (2025)
MVOS_HSI: A Python Library for Preprocessing Agricultural Crop Hyperspectral Data
por: Aggarwal, Rishik, et al.
Publicado: (2026)
por: Aggarwal, Rishik, et al.
Publicado: (2026)
ITKIT: Feasible CT Image Analysis based on SimpleITK and MMEngine
por: Zhang, Yiqin, et al.
Publicado: (2026)
por: Zhang, Yiqin, et al.
Publicado: (2026)
Ejemplares similares
-
Vector search with small radiuses
por: Szilvasy, Gergely, et al.
Publicado: (2024) -
Inference-time sparse attention with asymmetric indexing
por: Mazaré, Pierre-Emmanuel, et al.
Publicado: (2025) -
Self-Pruned Key-Value Attention: Learning When to Write by Predicting Future Utility
por: Szilvasy, Gergely, et al.
Publicado: (2026) -
Short window attention enables long-term memorization
por: Cabannes, Loïc, et al.
Publicado: (2025) -
Stochastic activations
por: Lomeli, Maria, et al.
Publicado: (2025)