Saved in:
| Main Authors: | Yang, Yunhao, Hu, Yuxin, Ye, Mao, Zhang, Zaiwei, Lu, Zhichao, Xu, Yi, Topcu, Ufuk, Snyder, Ben |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.01144 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multimodal Pretrained Models for Verifiable Sequential Decision-Making: Planning, Grounding, and Perception
by: Yang, Yunhao, et al.
Published: (2023)
by: Yang, Yunhao, et al.
Published: (2023)
Know Where You're Uncertain When Planning with Multimodal Foundation Models: A Formal Framework
by: Bhatt, Neel P., et al.
Published: (2024)
by: Bhatt, Neel P., et al.
Published: (2024)
Reasoning, Memorization, and Fine-Tuning Language Models for Non-Cooperative Games
by: Yang, Yunhao, et al.
Published: (2024)
by: Yang, Yunhao, et al.
Published: (2024)
Any2Any: Incomplete Multimodal Retrieval with Conformal Prediction
by: Li, Po-han, et al.
Published: (2024)
by: Li, Po-han, et al.
Published: (2024)
Real-Time Privacy Preservation for Robot Visual Perception
by: Choi, Minkyu, et al.
Published: (2025)
by: Choi, Minkyu, et al.
Published: (2025)
VLM-AD: End-to-End Autonomous Driving through Vision-Language Model Supervision
by: Xu, Yi, et al.
Published: (2024)
by: Xu, Yi, et al.
Published: (2024)
VLN-Zero: Rapid Exploration and Cache-Enabled Neurosymbolic Vision-Language Planning for Zero-Shot Transfer in Robot Navigation
by: Bhatt, Neel P., et al.
Published: (2025)
by: Bhatt, Neel P., et al.
Published: (2025)
Foundation Models for Logistics: Toward Certifiable, Conversational Planning Interfaces
by: Yang, Yunhao, et al.
Published: (2025)
by: Yang, Yunhao, et al.
Published: (2025)
AesExpert: Towards Multi-modality Foundation Model for Image Aesthetics Perception
by: Huang, Yipo, et al.
Published: (2024)
by: Huang, Yipo, et al.
Published: (2024)
Noncooperative Virtual Queue Coordination via Uncertainty-Aware Correlated Equilibria
by: Im, Jaehan, et al.
Published: (2026)
by: Im, Jaehan, et al.
Published: (2026)
Joint Verification and Refinement of Language Models for Safety-Constrained Planning
by: Yang, Yunhao, et al.
Published: (2024)
by: Yang, Yunhao, et al.
Published: (2024)
CSA: Data-efficient Mapping of Unimodal Features to Multimodal Features
by: Li, Po-han, et al.
Published: (2024)
by: Li, Po-han, et al.
Published: (2024)
VLM-KD: Knowledge Distillation from VLM for Long-Tail Visual Recognition
by: Zhang, Zaiwei, et al.
Published: (2024)
by: Zhang, Zaiwei, et al.
Published: (2024)
Evaluating Human Trust in LLM-Based Planners: A Preliminary Study
by: Chen, Shenghui, et al.
Published: (2025)
by: Chen, Shenghui, et al.
Published: (2025)
UNCAP: Uncertainty-Guided Neurosymbolic Planning Using Natural Language Communication for Cooperative Autonomous Vehicles
by: Bhatt, Neel P., et al.
Published: (2025)
by: Bhatt, Neel P., et al.
Published: (2025)
VIBE: Annotation-Free Video-to-Text Information Bottleneck Evaluation for TL;DR
by: Chen, Shenghui, et al.
Published: (2025)
by: Chen, Shenghui, et al.
Published: (2025)
ViSIL: Unified Evaluation of Information Loss in Multimodal Video Captioning
by: Li, Po-han, et al.
Published: (2026)
by: Li, Po-han, et al.
Published: (2026)
Toward Real-world BEV Perception: Depth Uncertainty Estimation via Gaussian Splatting
by: Lu, Shu-Wei, et al.
Published: (2025)
by: Lu, Shu-Wei, et al.
Published: (2025)
Unveiling the Underwater World: CLIP Perception Model-Guided Underwater Image Enhancement
by: Cao, Jiangzhong, et al.
Published: (2025)
by: Cao, Jiangzhong, et al.
Published: (2025)
EvoDriveVLA: Evolving Driving VLA Models via Collaborative Perception-Planning Distillation
by: Cao, Jiajun, et al.
Published: (2026)
by: Cao, Jiajun, et al.
Published: (2026)
VLMine: Long-Tail Data Mining with Vision Language Models
by: Ye, Mao, et al.
Published: (2024)
by: Ye, Mao, et al.
Published: (2024)
Language-Guided Visual Perception Disentanglement for Image Quality Assessment and Conditional Image Generation
by: Yang, Zhichao, et al.
Published: (2025)
by: Yang, Zhichao, et al.
Published: (2025)
DriveAgent-R1: Advancing VLM-based Autonomous Driving with Active Perception and Hybrid Thinking
by: Zheng, Weicheng, et al.
Published: (2025)
by: Zheng, Weicheng, et al.
Published: (2025)
Adversarial Observability and Performance Trade-offs in Optimal Control
by: Fotiadis, Filippos, et al.
Published: (2025)
by: Fotiadis, Filippos, et al.
Published: (2025)
Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities
by: Sathyam, Rajendramayavan, et al.
Published: (2025)
by: Sathyam, Rajendramayavan, et al.
Published: (2025)
Underwater Variable Zoom: Depth-Guided Perception Network for Underwater Image Enhancement
by: Huang, Zhixiong, et al.
Published: (2024)
by: Huang, Zhixiong, et al.
Published: (2024)
Collaborative Perception Datasets for Autonomous Driving: A Review
by: Wang, Naibang, et al.
Published: (2025)
by: Wang, Naibang, et al.
Published: (2025)
Zero-Shot Reinforcement Learning via Function Encoders
by: Ingebrand, Tyler, et al.
Published: (2024)
by: Ingebrand, Tyler, et al.
Published: (2024)
Online Foundation Model Selection in Robotics
by: Li, Po-han, et al.
Published: (2024)
by: Li, Po-han, et al.
Published: (2024)
UniDrive: Towards Universal Driving Perception Across Camera Configurations
by: Li, Ye, et al.
Published: (2024)
by: Li, Ye, et al.
Published: (2024)
MUSES: The Multi-Sensor Semantic Perception Dataset for Driving under Uncertainty
by: Brödermann, Tim, et al.
Published: (2024)
by: Brödermann, Tim, et al.
Published: (2024)
RepV: Safety-Separable Latent Spaces for Scalable Neurosymbolic Plan Verification
by: Yang, Yunhao, et al.
Published: (2025)
by: Yang, Yunhao, et al.
Published: (2025)
Repairing Catastrophic-Neglect in Text-to-Image Diffusion Models via Attention-Guided Feature Enhancement
by: Chang, Zhiyuan, et al.
Published: (2024)
by: Chang, Zhiyuan, et al.
Published: (2024)
Multi-Environment POMDPs: Discrete Model Uncertainty Under Partial Observability
by: Bovy, Eline M., et al.
Published: (2025)
by: Bovy, Eline M., et al.
Published: (2025)
Fine-Tuning Language Models Using Formal Methods Feedback
by: Yang, Yunhao, et al.
Published: (2023)
by: Yang, Yunhao, et al.
Published: (2023)
Towards Collaborative Autonomous Driving: Simulation Platform and End-to-End System
by: Liu, Genjia, et al.
Published: (2024)
by: Liu, Genjia, et al.
Published: (2024)
AD-SAM: Fine-Tuning the Segment Anything Vision Foundation Model for Autonomous Driving Perception
by: Camarena, Mario, et al.
Published: (2025)
by: Camarena, Mario, et al.
Published: (2025)
Human-Centric Foundation Models: Perception, Generation and Agentic Modeling
by: Tang, Shixiang, et al.
Published: (2025)
by: Tang, Shixiang, et al.
Published: (2025)
Hierarchical Graph Feature Enhancement with Adaptive Frequency Modulation for Visual Recognition
by: Zhao, Feiyue, et al.
Published: (2025)
by: Zhao, Feiyue, et al.
Published: (2025)
Free Lunch in Pathology Foundation Model: Task-specific Model Adaptation with Concept-Guided Feature Enhancement
by: Huang, Yanyan, et al.
Published: (2024)
by: Huang, Yanyan, et al.
Published: (2024)
Similar Items
-
Multimodal Pretrained Models for Verifiable Sequential Decision-Making: Planning, Grounding, and Perception
by: Yang, Yunhao, et al.
Published: (2023) -
Know Where You're Uncertain When Planning with Multimodal Foundation Models: A Formal Framework
by: Bhatt, Neel P., et al.
Published: (2024) -
Reasoning, Memorization, and Fine-Tuning Language Models for Non-Cooperative Games
by: Yang, Yunhao, et al.
Published: (2024) -
Any2Any: Incomplete Multimodal Retrieval with Conformal Prediction
by: Li, Po-han, et al.
Published: (2024) -
Real-Time Privacy Preservation for Robot Visual Perception
by: Choi, Minkyu, et al.
Published: (2025)