A Note on Semantic Diffusion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ryjov, Alexander P., Egorova, Alina A. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Measuring What Matters: Scenario-Driven Evaluation for Trajectory Predictors in Autonomous Driving
von: Da, Longchao, et al.
Veröffentlicht: (2025)
von: Da, Longchao, et al.
Veröffentlicht: (2025)
From Heuristics to Data: Quantifying Site Planning Layout Indicators with Deep Learning and Multi-Modal Data
von: Cao, Qian, et al.
Veröffentlicht: (2025)
von: Cao, Qian, et al.
Veröffentlicht: (2025)
Learning Accurate Whole-body Throwing with High-frequency Residual Policy and Pullback Tube Acceleration
von: Ma, Yuntao, et al.
Veröffentlicht: (2025)
von: Ma, Yuntao, et al.
Veröffentlicht: (2025)
RoboMoRe: LLM-based Robot Co-design via Joint Optimization of Morphology and Reward
von: Fang, Jiawei, et al.
Veröffentlicht: (2025)
von: Fang, Jiawei, et al.
Veröffentlicht: (2025)
Deep Probabilistic Traversability with Test-time Adaptation for Uncertainty-aware Planetary Rover Navigation
von: Endo, Masafumi, et al.
Veröffentlicht: (2024)
von: Endo, Masafumi, et al.
Veröffentlicht: (2024)
Agentic UAVs: LLM-Driven Autonomy with Integrated Tool-Calling and Cognitive Reasoning
von: Koubaa, Anis, et al.
Veröffentlicht: (2025)
von: Koubaa, Anis, et al.
Veröffentlicht: (2025)
Multimodal Generative AI for Story Point Estimation in Software Development
von: Islam, Mohammad Rubyet, et al.
Veröffentlicht: (2025)
von: Islam, Mohammad Rubyet, et al.
Veröffentlicht: (2025)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
Think, Act, Learn: A Framework for Autonomous Robotic Agents using Closed-Loop Large Language Models
von: Menon, Anjali R., et al.
Veröffentlicht: (2025)
von: Menon, Anjali R., et al.
Veröffentlicht: (2025)
Incremental Bootstrapping and Classification of Structured Scenes in a Fuzzy Ontology
von: Buoncompagni, Luca, et al.
Veröffentlicht: (2024)
von: Buoncompagni, Luca, et al.
Veröffentlicht: (2024)
TSPE-GS: Probabilistic Depth Extraction for Semi-Transparent Surface Reconstruction via 3D Gaussian Splatting
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2025)
Inducing Causal World Models in LLMs for Zero-Shot Physical Reasoning
von: Sharma, Aditya, et al.
Veröffentlicht: (2025)
von: Sharma, Aditya, et al.
Veröffentlicht: (2025)
Learning Abstract Visual Reasoning via Task Decomposition: A Case Study in Raven Progressive Matrices
von: Kwiatkowski, Jakub, et al.
Veröffentlicht: (2023)
von: Kwiatkowski, Jakub, et al.
Veröffentlicht: (2023)
From Demonstrations to Safe Deployment: Path-Consistent Safety Filtering for Diffusion Policies
von: Römer, Ralf, et al.
Veröffentlicht: (2025)
von: Römer, Ralf, et al.
Veröffentlicht: (2025)
Key-Scan-Based Mobile Robot Navigation: Integrated Mapping, Planning, and Control using Graphs of Scan Regions
von: Latha, Dharshan Bashkaran, et al.
Veröffentlicht: (2024)
von: Latha, Dharshan Bashkaran, et al.
Veröffentlicht: (2024)
Image-based Facial Rig Inversion
von: Yang, Tianxiang, et al.
Veröffentlicht: (2025)
von: Yang, Tianxiang, et al.
Veröffentlicht: (2025)
AI Model for Predicting Binding Affinity of Antidiabetic Compounds Targeting PPAR
von: Aman, La Ode, et al.
Veröffentlicht: (2024)
von: Aman, La Ode, et al.
Veröffentlicht: (2024)
ASkDAgger: Active Skill-level Data Aggregation for Interactive Imitation Learning
von: Luijkx, Jelle, et al.
Veröffentlicht: (2025)
von: Luijkx, Jelle, et al.
Veröffentlicht: (2025)
Proving Olympiad Algebraic Inequalities without Human Demonstrations
von: Wei, Chenrui, et al.
Veröffentlicht: (2024)
von: Wei, Chenrui, et al.
Veröffentlicht: (2024)
When Does Global Attention Help? A Unified Empirical Study on Atomistic Graph Learning
von: Chowdhury, Arindam, et al.
Veröffentlicht: (2025)
von: Chowdhury, Arindam, et al.
Veröffentlicht: (2025)
From Particles to Agents: Hallucination as a Metric for Cognitive Friction in Spatial Simulation
von: Sánchez-Vaquerizo, Javier Argota, et al.
Veröffentlicht: (2026)
von: Sánchez-Vaquerizo, Javier Argota, et al.
Veröffentlicht: (2026)
AgentCPM-GUI: Building Mobile-Use Agents with Reinforcement Fine-Tuning
von: Zhang, Zhong, et al.
Veröffentlicht: (2025)
von: Zhang, Zhong, et al.
Veröffentlicht: (2025)
Toward a Dialogue System Using a Large Language Model to Recognize User Emotions with a Camera
von: Tanioka, Hiroki, et al.
Veröffentlicht: (2024)
von: Tanioka, Hiroki, et al.
Veröffentlicht: (2024)
3DreamBooth: High-Fidelity 3D Subject-Driven Video Generation Model
von: Ko, Hyun-kyu, et al.
Veröffentlicht: (2026)
von: Ko, Hyun-kyu, et al.
Veröffentlicht: (2026)
SemanticFeels: Semantic Labeling during In-Hand Manipulation
von: Khalil, Anas Al Shikh, et al.
Veröffentlicht: (2026)
von: Khalil, Anas Al Shikh, et al.
Veröffentlicht: (2026)
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
von: Chahine, Makram, et al.
Veröffentlicht: (2024)
von: Chahine, Makram, et al.
Veröffentlicht: (2024)
MVTamperBench: Evaluating Robustness of Vision-Language Models
von: Agarwal, Amit, et al.
Veröffentlicht: (2024)
von: Agarwal, Amit, et al.
Veröffentlicht: (2024)
vailá: Versatile Anarcho Integrated Liberation Ánalysis in Multimodal Toolbox
von: Santiago, Paulo Roberto Pereira, et al.
Veröffentlicht: (2024)
von: Santiago, Paulo Roberto Pereira, et al.
Veröffentlicht: (2024)
Improving Efficiency of Sampling-based Motion Planning via Message-Passing Monte Carlo
von: Chahine, Makram, et al.
Veröffentlicht: (2024)
von: Chahine, Makram, et al.
Veröffentlicht: (2024)
S3Simulator: A benchmarking Side Scan Sonar Simulator dataset for Underwater Image Analysis
von: S, Kamal Basha, et al.
Veröffentlicht: (2024)
von: S, Kamal Basha, et al.
Veröffentlicht: (2024)
Immersive Robot Programming Interface for Human-Guided Automation and Randomized Path Planning
von: Malek, Kaveh, et al.
Veröffentlicht: (2024)
von: Malek, Kaveh, et al.
Veröffentlicht: (2024)
Experimental Evaluation of Road-Crossing Decisions by Autonomous Wheelchairs against Environmental Factors
von: Corradini, Franca, et al.
Veröffentlicht: (2024)
von: Corradini, Franca, et al.
Veröffentlicht: (2024)
Classifier Calibration at Scale: An Empirical Study of Model-Agnostic Post-Hoc Methods
von: Manokhin, Valery, et al.
Veröffentlicht: (2026)
von: Manokhin, Valery, et al.
Veröffentlicht: (2026)
COBRA-PPM: A Causal Bayesian Reasoning Architecture Using Probabilistic Programming for Robot Manipulation Under Uncertainty
von: Cannizzaro, Ricardo, et al.
Veröffentlicht: (2024)
von: Cannizzaro, Ricardo, et al.
Veröffentlicht: (2024)
SafeDMPs: Integrating Formal Safety with DMPs for Adaptive HRI
von: Nath, Soumyodipta, et al.
Veröffentlicht: (2026)
von: Nath, Soumyodipta, et al.
Veröffentlicht: (2026)
Evaluating Visual Mathematics in Multimodal LLMs: A Multilingual Benchmark Based on the Kangaroo Tests
von: Sáez, Arnau Igualde, et al.
Veröffentlicht: (2025)
von: Sáez, Arnau Igualde, et al.
Veröffentlicht: (2025)
OpenFusion++: An Open-vocabulary Real-time Scene Understanding System
von: Jin, Xiaofeng, et al.
Veröffentlicht: (2025)
von: Jin, Xiaofeng, et al.
Veröffentlicht: (2025)
Consistent Zero-shot 3D Texture Synthesis Using Geometry-aware Diffusion and Temporal Video Models
von: Kang, Donggoo, et al.
Veröffentlicht: (2025)
von: Kang, Donggoo, et al.
Veröffentlicht: (2025)
VisChainBench: A Benchmark for Multi-Turn, Multi-Image Visual Reasoning Beyond Language Priors
von: Lyu, Wenbo, et al.
Veröffentlicht: (2025)
von: Lyu, Wenbo, et al.
Veröffentlicht: (2025)
Failure Prediction at Runtime for Generative Robot Policies
von: Römer, Ralf, et al.
Veröffentlicht: (2025)
von: Römer, Ralf, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Measuring What Matters: Scenario-Driven Evaluation for Trajectory Predictors in Autonomous Driving
von: Da, Longchao, et al.
Veröffentlicht: (2025) -
From Heuristics to Data: Quantifying Site Planning Layout Indicators with Deep Learning and Multi-Modal Data
von: Cao, Qian, et al.
Veröffentlicht: (2025) -
Learning Accurate Whole-body Throwing with High-frequency Residual Policy and Pullback Tube Acceleration
von: Ma, Yuntao, et al.
Veröffentlicht: (2025) -
RoboMoRe: LLM-based Robot Co-design via Joint Optimization of Morphology and Reward
von: Fang, Jiawei, et al.
Veröffentlicht: (2025) -
Deep Probabilistic Traversability with Test-time Adaptation for Uncertainty-aware Planetary Rover Navigation
von: Endo, Masafumi, et al.
Veröffentlicht: (2024)