ESCA: Contextualizing Embodied Agents via Scene-Graph Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Huang, Jiani, Sethi, Amish, Kuo, Matthew, Keoliya, Mayank, Velingker, Neelay, Jung, JungHo, Lim, Ser-Nam, Li, Ziyang, Naik, Mayur |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LASER: A Neuro-Symbolic Framework for Learning Spatial-Temporal Scene Graphs with Weak Supervision
by: Huang, Jiani, et al.
Published: (2023)
by: Huang, Jiani, et al.
Published: (2023)
Delta Activations: A Representation for Finetuned Large Language Models
by: Xu, Zhiqiu, et al.
Published: (2025)
by: Xu, Zhiqiu, et al.
Published: (2025)
Once Upon an Input: Reasoning via Per-Instance Program Synthesis
by: Stein, Adam, et al.
Published: (2025)
by: Stein, Adam, et al.
Published: (2025)
DISCRET: Synthesizing Faithful Explanations For Treatment Effect Estimation
by: Wu, Yinjun, et al.
Published: (2024)
by: Wu, Yinjun, et al.
Published: (2024)
The Road to Generalizable Neuro-Symbolic Learning Should be Paved with Foundation Models
by: Stein, Adam, et al.
Published: (2025)
by: Stein, Adam, et al.
Published: (2025)
Relational Programming with Foundation Models
by: Li, Ziyang, et al.
Published: (2024)
by: Li, Ziyang, et al.
Published: (2024)
Data-Efficient Learning with Neural Programs
by: Solko-Breslin, Alaia, et al.
Published: (2024)
by: Solko-Breslin, Alaia, et al.
Published: (2024)
Stable Prediction of Adverse Events in Medical Time-Series Data
by: Keoliya, Mayank, et al.
Published: (2025)
by: Keoliya, Mayank, et al.
Published: (2025)
CAMEL: An ECG Language Model for Forecasting Cardiac Events
by: Velingker, Neelay, et al.
Published: (2026)
by: Velingker, Neelay, et al.
Published: (2026)
Optimal Multi-Task Learning at Regularization Horizon for Speech Translation Task
by: Jung, JungHo, et al.
Published: (2025)
by: Jung, JungHo, et al.
Published: (2025)
Summing divergent matrix series
by: Wang, Rongbiao, et al.
Published: (2024)
by: Wang, Rongbiao, et al.
Published: (2024)
Dolphin: A Programmable Framework for Scalable Neurosymbolic Learning
by: Naik, Aaditya, et al.
Published: (2024)
by: Naik, Aaditya, et al.
Published: (2024)
ML-assisted Randomization Tests for Detecting Treatment Effects in A/B Experiments
by: Guo, Wenxuan, et al.
Published: (2025)
by: Guo, Wenxuan, et al.
Published: (2025)
Learning Smooth Populations of Parameters with Trial Heterogeneity
by: Lee, JungHo, et al.
Published: (2025)
by: Lee, JungHo, et al.
Published: (2025)
Scene Co-pilot: Procedural Text to Video Generation with Human in the Loop
by: Qian, Zhaofang, et al.
Published: (2024)
by: Qian, Zhaofang, et al.
Published: (2024)
Towards Chunk-Wise Generation for Long Videos
by: Zhang, Siyang, et al.
Published: (2024)
by: Zhang, Siyang, et al.
Published: (2024)
Learning Treatment Effects during Resource Allocation via Priority-Queue Randomization
by: Lee, JungHo, et al.
Published: (2026)
by: Lee, JungHo, et al.
Published: (2026)
Black-Box Optimization with Implicit Constraints for Public Policy
by: Xing, Wenqian, et al.
Published: (2023)
by: Xing, Wenqian, et al.
Published: (2023)
What can Off-the-Shelves Large Multi-Modal Models do for Dynamic Scene Graph Generation?
by: Cui, Xuanming, et al.
Published: (2025)
by: Cui, Xuanming, et al.
Published: (2025)
IRIS: LLM-Assisted Static Analysis for Detecting Security Vulnerabilities
by: Li, Ziyang, et al.
Published: (2024)
by: Li, Ziyang, et al.
Published: (2024)
LARM: Large Auto-Regressive Model for Long-Horizon Embodied Intelligence
by: Li, Zhuoling, et al.
Published: (2024)
by: Li, Zhuoling, et al.
Published: (2024)
Caruca: Effective and Efficient Specification Mining for Opaque Software Components
by: Lamprou, Evangelos, et al.
Published: (2025)
by: Lamprou, Evangelos, et al.
Published: (2025)
VideoMerge: Towards Training-free Long Video Generation
by: Zhang, Siyang, et al.
Published: (2025)
by: Zhang, Siyang, et al.
Published: (2025)
Lobster: A GPU-Accelerated Framework for Neurosymbolic Programming
by: Biberstein, Paul, et al.
Published: (2025)
by: Biberstein, Paul, et al.
Published: (2025)
QLCoder: A Query Synthesizer For Static Analysis of Security Vulnerabilities
by: Wang, Claire, et al.
Published: (2025)
by: Wang, Claire, et al.
Published: (2025)
Beyond Generation: Unlocking Universal Editing via Self-Supervised Fine-Tuning
by: Chen, Harold Haodong, et al.
Published: (2024)
by: Chen, Harold Haodong, et al.
Published: (2024)
Nonparametric Estimation of Local Treatment Effects with Continuous Instruments
by: Zeng, Zhenghao, et al.
Published: (2025)
by: Zeng, Zhenghao, et al.
Published: (2025)
AirSketch: Generative Motion to Sketch
by: Lim, Hui Xian Grace, et al.
Published: (2024)
by: Lim, Hui Xian Grace, et al.
Published: (2024)
Program Structure Aware Precondition Generation
by: Dinella, Elizabeth, et al.
Published: (2023)
by: Dinella, Elizabeth, et al.
Published: (2023)
Video Decomposition Prior: A Methodology to Decompose Videos into Layers
by: Shrivastava, Gaurav, et al.
Published: (2024)
by: Shrivastava, Gaurav, et al.
Published: (2024)
Imaginative World Modeling with Scene Graphs for Embodied Agent Navigation
by: Hu, Yue, et al.
Published: (2025)
by: Hu, Yue, et al.
Published: (2025)
Niagara: Normal-Integrated Geometric Affine Field for Scene Reconstruction from a Single View
by: Wu, Xianzu, et al.
Published: (2025)
by: Wu, Xianzu, et al.
Published: (2025)
BOOKAGENT: Orchestrating Safety-Aware Visual Narratives via Multi-Agent Cognitive Calibration
by: Gao, Bo, et al.
Published: (2026)
by: Gao, Bo, et al.
Published: (2026)
An Exact Solution for the Kinetic Ising Model with Non-Reciprocity
by: Weiderpass, Gabriel, et al.
Published: (2024)
by: Weiderpass, Gabriel, et al.
Published: (2024)
DiReCT: Disentangled Regularization of Contrastive Trajectories for Physics-Refined Video Generation
by: Meyarian, Abolfazl, et al.
Published: (2026)
by: Meyarian, Abolfazl, et al.
Published: (2026)
Scene Graph-Guided Proactive Replanning for Failure-Resilient Embodied Agent
by: Yu, Che Rin, et al.
Published: (2025)
by: Yu, Che Rin, et al.
Published: (2025)
FSViewFusion: Few-Shots View Generation of Novel Objects
by: Hussain, Rukhshanda, et al.
Published: (2024)
by: Hussain, Rukhshanda, et al.
Published: (2024)
Compromising Embodied Agents with Contextual Backdoor Attacks
by: Liu, Aishan, et al.
Published: (2024)
by: Liu, Aishan, et al.
Published: (2024)
Do We Need Frontier Models to Verify Mathematical Proofs?
by: Naik, Aaditya, et al.
Published: (2026)
by: Naik, Aaditya, et al.
Published: (2026)
Environmental Understanding Vision-Language Model for Embodied Agent
by: Bang, Jinsik, et al.
Published: (2026)
by: Bang, Jinsik, et al.
Published: (2026)
Similar Items
-
LASER: A Neuro-Symbolic Framework for Learning Spatial-Temporal Scene Graphs with Weak Supervision
by: Huang, Jiani, et al.
Published: (2023) -
Delta Activations: A Representation for Finetuned Large Language Models
by: Xu, Zhiqiu, et al.
Published: (2025) -
Once Upon an Input: Reasoning via Per-Instance Program Synthesis
by: Stein, Adam, et al.
Published: (2025) -
DISCRET: Synthesizing Faithful Explanations For Treatment Effect Estimation
by: Wu, Yinjun, et al.
Published: (2024) -
The Road to Generalizable Neuro-Symbolic Learning Should be Paved with Foundation Models
by: Stein, Adam, et al.
Published: (2025)