Cross-Modal State-Space Graph Reasoning for Structured Summarization
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Hannah, Martinez, Sofia, Lee, Jason |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ReverBERT: A State Space Model for Efficient Text-Driven Speech Style Transfer
by: Brown, Michael, et al.
Published: (2025)
by: Brown, Michael, et al.
Published: (2025)
StyleMotif: Multi-Modal Motion Stylization using Style-Content Cross Fusion
by: Guo, Ziyu, et al.
Published: (2025)
by: Guo, Ziyu, et al.
Published: (2025)
State of the Art of Graph Visualization in non‐Euclidean Spaces
by: Jacob Miller, et al.
Published: (2024)
by: Jacob Miller, et al.
Published: (2024)
From Pixels to Policies: Reinforcing Spatial Reasoning in Language Models for Content-Aware Layout Design
by: Li, Sha, et al.
Published: (2026)
by: Li, Sha, et al.
Published: (2026)
FilmAgent: A Multi-Agent Framework for End-to-End Film Automation in Virtual 3D Spaces
by: Xu, Zhenran, et al.
Published: (2025)
by: Xu, Zhenran, et al.
Published: (2025)
DiffListener: Discrete Diffusion Model for Listener Generation
by: Jung, Siyeol, et al.
Published: (2025)
by: Jung, Siyeol, et al.
Published: (2025)
MMS Player: an open source software for parametric data-driven animation of Sign Language avatars
by: Nunnari, Fabrizio, et al.
Published: (2025)
by: Nunnari, Fabrizio, et al.
Published: (2025)
LayerFlow: Layer-wise Exploration of LLM Embeddings using Uncertainty-aware Interlinked Projections
by: Sevastjanova, Rita, et al.
Published: (2025)
by: Sevastjanova, Rita, et al.
Published: (2025)
Token Perturbation Guidance for Diffusion Models
by: Rajabi, Javad, et al.
Published: (2025)
by: Rajabi, Javad, et al.
Published: (2025)
Parametric type design in the era of variable and color fonts
by: Thottingal, Santhosh
Published: (2025)
by: Thottingal, Santhosh
Published: (2025)
Self-Improving CAD Generation Agents with Finite Element Analysis as Feedback
by: Son, Guijin, et al.
Published: (2026)
by: Son, Guijin, et al.
Published: (2026)
Visualizing Temporal Topic Embeddings with a Compass
by: Palamarchuk, Daniel, et al.
Published: (2024)
by: Palamarchuk, Daniel, et al.
Published: (2024)
The Effects of Embodiment and Personality Expression on Learning in LLM-based Educational Agents
by: Sonlu, Sinan, et al.
Published: (2024)
by: Sonlu, Sinan, et al.
Published: (2024)
Innovating China's Intangible Cultural Heritage with DeepSeek + MidJourney: The Case of Yangliuqing theme Woodblock Prints
by: Yang, RuiKun, et al.
Published: (2025)
by: Yang, RuiKun, et al.
Published: (2025)
CoolerSpace: A Language for Physically Correct and Computationally Efficient Color Programming
by: Chen, Ethan, et al.
Published: (2024)
by: Chen, Ethan, et al.
Published: (2024)
RotGS: Rotation‐Guided 3D Gaussian Splatting for Turntable Sequences without Structure‐from‐Motion
by: Kyumin Kim, et al.
Published: (2026)
by: Kyumin Kim, et al.
Published: (2026)
Harnessing Adaptive Topology Representations for Zero-Shot Graph Question Answering
by: Wei, Yanbin, et al.
Published: (2025)
by: Wei, Yanbin, et al.
Published: (2025)
DynamicGTR: Leveraging Graph Topology Representation Preferences to Boost VLM Capabilities on Graph QAs
by: Wei, Yanbin, et al.
Published: (2026)
by: Wei, Yanbin, et al.
Published: (2026)
DreamDPO: Aligning Text-to-3D Generation with Human Preferences via Direct Preference Optimization
by: Zhou, Zhenglin, et al.
Published: (2025)
by: Zhou, Zhenglin, et al.
Published: (2025)
From Words to Worlds: Transforming One-line Prompt into Immersive Multi-modal Digital Stories with Communicative LLM Agent
by: Sohn, Samuel S., et al.
Published: (2024)
by: Sohn, Samuel S., et al.
Published: (2024)
Cutscene Agent: An LLM Agent Framework for Automated 3D Cutscene Generation
by: He, Lanshan, et al.
Published: (2026)
by: He, Lanshan, et al.
Published: (2026)
3D-PreMise: Can Large Language Models Generate 3D Shapes with Sharp Features and Parametric Control?
by: Yuan, Zeqing, et al.
Published: (2024)
by: Yuan, Zeqing, et al.
Published: (2024)
YASPS: A Symbolic Framework for Extensible, High-Performance IPC Simulation
by: Tang, Xuan, et al.
Published: (2026)
by: Tang, Xuan, et al.
Published: (2026)
QQJ: Quantifying Qualitative Judgment for Scalable and Human-Aligned Evaluation of Generative AI
by: Veysi, Marjan, et al.
Published: (2026)
by: Veysi, Marjan, et al.
Published: (2026)
Finite-State Automaton To/From Regular Expression Visualization
by: Morazán, Marco T., et al.
Published: (2024)
by: Morazán, Marco T., et al.
Published: (2024)
Text-Driven Voice Conversion via Latent State-Space Modeling
by: Li, Wen, et al.
Published: (2025)
by: Li, Wen, et al.
Published: (2025)
An Evaluation-Centric Paradigm for Scientific Visualization Agents
by: Ai, Kuangshi, et al.
Published: (2025)
by: Ai, Kuangshi, et al.
Published: (2025)
Visual Guidance for User Placement in Avatar-Mediated Telepresence between Dissimilar Spaces
by: Yang, Dongseok, et al.
Published: (2022)
by: Yang, Dongseok, et al.
Published: (2022)
3DMambaComplete: Exploring Structured State Space Model for Point Cloud Completion
by: Li, Yixuan, et al.
Published: (2024)
by: Li, Yixuan, et al.
Published: (2024)
Gaussians on their Way: Wasserstein‐Constrained 4D Gaussian Splatting with State‐Space Modeling
by: J. Deng, et al.
Published: (2025)
by: J. Deng, et al.
Published: (2025)
GeoFusion-CAD: Structure-Aware Diffusion with Geometric State Space for Parametric 3D Design
by: Zhou, Xiaolei, et al.
Published: (2026)
by: Zhou, Xiaolei, et al.
Published: (2026)
RT-HDIST: Ray-Tracing Core-based Hausdorff Distance Computation
by: Kim, YoungWoo, et al.
Published: (2025)
by: Kim, YoungWoo, et al.
Published: (2025)
Depth for Multi‐Modal Contour Ensembles
by: N.F. Chaves‐de‐Plaza, et al.
Published: (2024)
by: N.F. Chaves‐de‐Plaza, et al.
Published: (2024)
Learning Human Motion from Monocular Videos via Cross-Modal Manifold Alignment
by: Hou, Shuaiying, et al.
Published: (2024)
by: Hou, Shuaiying, et al.
Published: (2024)
FIFA: Unified Faithfulness Evaluation Framework for Text-to-Video and Video-to-Text Generation
by: Jing, Liqiang, et al.
Published: (2025)
by: Jing, Liqiang, et al.
Published: (2025)
Is this chart lying to me? Automating the detection of misleading visualizations
by: Tonglet, Jonathan, et al.
Published: (2025)
by: Tonglet, Jonathan, et al.
Published: (2025)
A Study of the Framework and Real-World Applications of Language Embedding for 3D Scene Understanding
by: Zaouali, Mahmoud Chick, et al.
Published: (2025)
by: Zaouali, Mahmoud Chick, et al.
Published: (2025)
TexGS-VolVis: Expressive Scene Editing for Volume Visualization via Textured Gaussian Splatting
by: Tang, Kaiyuan, et al.
Published: (2025)
by: Tang, Kaiyuan, et al.
Published: (2025)
FlairGPT: Repurposing LLMs for Interior Designs
by: Littlefair, Gabrielle, et al.
Published: (2025)
by: Littlefair, Gabrielle, et al.
Published: (2025)
Co-Layout: LLM-driven Co-optimization for Interior Layout
by: Xiang, Chucheng, et al.
Published: (2025)
by: Xiang, Chucheng, et al.
Published: (2025)
Similar Items
-
ReverBERT: A State Space Model for Efficient Text-Driven Speech Style Transfer
by: Brown, Michael, et al.
Published: (2025) -
StyleMotif: Multi-Modal Motion Stylization using Style-Content Cross Fusion
by: Guo, Ziyu, et al.
Published: (2025) -
State of the Art of Graph Visualization in non‐Euclidean Spaces
by: Jacob Miller, et al.
Published: (2024) -
From Pixels to Policies: Reinforcing Spatial Reasoning in Language Models for Content-Aware Layout Design
by: Li, Sha, et al.
Published: (2026) -
FilmAgent: A Multi-Agent Framework for End-to-End Film Automation in Virtual 3D Spaces
by: Xu, Zhenran, et al.
Published: (2025)