A Surveillance Based Interactive Robot
Fuente:
arXiv
Saved in:
| Main Authors: | Kavimandan, Kshitij, Mangal, Pooja, Mehta, Devanshi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RoboScript: Code Generation for Free-Form Manipulation Tasks across Real and Simulation
by: Chen, Junting, et al.
Published: (2024)
by: Chen, Junting, et al.
Published: (2024)
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
by: Li, Danyang, et al.
Published: (2025)
by: Li, Danyang, et al.
Published: (2025)
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
by: Lim, Shoon Kit, et al.
Published: (2025)
by: Lim, Shoon Kit, et al.
Published: (2025)
StratXplore: Strategic Novelty-seeking and Instruction-aligned Exploration for Vision and Language Navigation
by: Gopinathan, Muraleekrishna, et al.
Published: (2024)
by: Gopinathan, Muraleekrishna, et al.
Published: (2024)
Learning from Watching: Scalable Extraction of Manipulation Trajectories from Human Videos
by: Hu, X., et al.
Published: (2025)
by: Hu, X., et al.
Published: (2025)
Real-Time Localization Framework for Autonomous Basketball Robots
by: Medarametla, Naren, et al.
Published: (2026)
by: Medarametla, Naren, et al.
Published: (2026)
NavTopo: Leveraging Topological Maps For Autonomous Navigation Of a Mobile Robot
by: Muravyev, Kirill, et al.
Published: (2024)
by: Muravyev, Kirill, et al.
Published: (2024)
Autonomous Navigation and Collision Avoidance for Mobile Robots: Classification and Review
by: de Carvalho, Marcus Vinicius Leal, et al.
Published: (2024)
by: de Carvalho, Marcus Vinicius Leal, et al.
Published: (2024)
VLA Foundry: A Unified Framework for Training Vision-Language-Action Models
by: Mercat, Jean, et al.
Published: (2026)
by: Mercat, Jean, et al.
Published: (2026)
Look and Tell: A Dataset for Multimodal Grounding Across Egocentric and Exocentric Views
by: Deichler, Anna, et al.
Published: (2025)
by: Deichler, Anna, et al.
Published: (2025)
Unpacking Failure Modes of Generative Policies: Runtime Monitoring of Consistency and Progress
by: Agia, Christopher, et al.
Published: (2024)
by: Agia, Christopher, et al.
Published: (2024)
EMOS: Embodiment-aware Heterogeneous Multi-robot Operating System with LLM Agents
by: Chen, Junting, et al.
Published: (2024)
by: Chen, Junting, et al.
Published: (2024)
Large Language Models and 3D Vision for Intelligent Robotic Perception and Autonomy
by: Mehta, Vinit, et al.
Published: (2025)
by: Mehta, Vinit, et al.
Published: (2025)
A Survey of Spatial Memory Representations for Efficient Robot Navigation
by: Pangaliman, Ma. Madecheen S., et al.
Published: (2026)
by: Pangaliman, Ma. Madecheen S., et al.
Published: (2026)
MAP: End-to-End Autonomous Driving with Map-Assisted Planning
by: Yin, Huilin, et al.
Published: (2025)
by: Yin, Huilin, et al.
Published: (2025)
PALMS: Plane-based Accessible Indoor Localization Using Mobile Smartphones
by: Cheng, Yunqian, et al.
Published: (2024)
by: Cheng, Yunqian, et al.
Published: (2024)
Pointing-Guided Target Estimation via Transformer-Based Attention
by: Müller, Luca, et al.
Published: (2025)
by: Müller, Luca, et al.
Published: (2025)
Spatially-Aware Speaker for Vision-and-Language Navigation Instruction Generation
by: Gopinathan, Muraleekrishna, et al.
Published: (2024)
by: Gopinathan, Muraleekrishna, et al.
Published: (2024)
Deployment-Time Reliability of Learned Robot Policies
by: Agia, Christopher
Published: (2026)
by: Agia, Christopher
Published: (2026)
Spot-Compose: A Framework for Open-Vocabulary Object Retrieval and Drawer Manipulation in Point Clouds
by: Lemke, Oliver, et al.
Published: (2024)
by: Lemke, Oliver, et al.
Published: (2024)
VIN-NBV: A View Introspection Network for Next-Best-View Selection
by: Frahm, Noah, et al.
Published: (2025)
by: Frahm, Noah, et al.
Published: (2025)
CoEnv: Driving Embodied Multi-Agent Collaboration via Compositional Environment
by: Kang, Li, et al.
Published: (2026)
by: Kang, Li, et al.
Published: (2026)
RAE-NWM: Navigation World Model in Dense Visual Representation Space
by: Zhang, Mingkun, et al.
Published: (2026)
by: Zhang, Mingkun, et al.
Published: (2026)
PoseRefer: Pathway-Local Parameters for Semantically Grounded Reference Resolution
by: Deichler, Anna
Published: (2026)
by: Deichler, Anna
Published: (2026)
PRISM-TopoMap: Online Topological Mapping with Place Recognition and Scan Matching
by: Muravyev, Kirill, et al.
Published: (2024)
by: Muravyev, Kirill, et al.
Published: (2024)
Zero-Shot Object Goal Visual Navigation With Class-Independent Relationship Network
by: Li, Xinting, et al.
Published: (2023)
by: Li, Xinting, et al.
Published: (2023)
M3CAD: Towards Generic Cooperative Autonomous Driving Benchmark
by: Zhu, Morui, et al.
Published: (2025)
by: Zhu, Morui, et al.
Published: (2025)
A Segmented Robot Grasping Perception Neural Network for Edge AI
by: Bröcheler, Casper, et al.
Published: (2025)
by: Bröcheler, Casper, et al.
Published: (2025)
Industrial Robot Motion Planning with GPUs: Integration of cuRobo for Extended DOF Systems
by: Abuelsamen, Luai, et al.
Published: (2025)
by: Abuelsamen, Luai, et al.
Published: (2025)
HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos
by: Wang, Zhi, et al.
Published: (2026)
by: Wang, Zhi, et al.
Published: (2026)
Towards Localizing Structural Elements: Merging Geometrical Detection with Semantic Verification in RGB-D Data
by: Tourani, Ali, et al.
Published: (2024)
by: Tourani, Ali, et al.
Published: (2024)
CleanMAP: Distilling Multimodal LLMs for Confidence-Driven Crowdsourced HD Map Updates
by: Shaw, Ankit Kumar, et al.
Published: (2025)
by: Shaw, Ankit Kumar, et al.
Published: (2025)
Unveiling the Potential of iMarkers: Invisible Fiducial Markers for Advanced Robotics
by: Tourani, Ali, et al.
Published: (2025)
by: Tourani, Ali, et al.
Published: (2025)
Temporally Consistent Object 6D Pose Estimation for Robot Control
by: Zorina, Kateryna, et al.
Published: (2026)
by: Zorina, Kateryna, et al.
Published: (2026)
PerspAct: Enhancing LLM Situated Collaboration Skills through Perspective Taking and Active Vision
by: Patania, Sabrina, et al.
Published: (2025)
by: Patania, Sabrina, et al.
Published: (2025)
Multi-modal Loop Closure Detection with Foundation Models in Severely Unstructured Environments
by: Gonzalez, Laura Alejandra Encinar, et al.
Published: (2025)
by: Gonzalez, Laura Alejandra Encinar, et al.
Published: (2025)
Unsupervised Decomposition and Recombination with Discriminator-Driven Diffusion Models
by: Wang, Archer, et al.
Published: (2026)
by: Wang, Archer, et al.
Published: (2026)
SuperPoint-SLAM3: Augmenting ORB-SLAM3 with Deep Features, Adaptive NMS, and Learning-Based Loop Closure
by: Syed, Shahram Najam, et al.
Published: (2025)
by: Syed, Shahram Najam, et al.
Published: (2025)
DriveMRP: Enhancing Vision-Language Models with Synthetic Motion Data for Motion Risk Prediction
by: Hou, Zhiyi, et al.
Published: (2025)
by: Hou, Zhiyi, et al.
Published: (2025)
WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation
by: Niu, Yuwei, et al.
Published: (2025)
by: Niu, Yuwei, et al.
Published: (2025)
Similar Items
-
RoboScript: Code Generation for Free-Form Manipulation Tasks across Real and Simulation
by: Chen, Junting, et al.
Published: (2024) -
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
by: Li, Danyang, et al.
Published: (2025) -
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
by: Lim, Shoon Kit, et al.
Published: (2025) -
StratXplore: Strategic Novelty-seeking and Instruction-aligned Exploration for Vision and Language Navigation
by: Gopinathan, Muraleekrishna, et al.
Published: (2024) -
Learning from Watching: Scalable Extraction of Manipulation Trajectories from Human Videos
by: Hu, X., et al.
Published: (2025)