TorchSpatial: A Location Encoding Framework and Benchmark for Spatial Representation Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Wu, Nemin, Cao, Qian, Wang, Zhangyu, Liu, Zeping, Qi, Yanlin, Zhang, Jielu, Ni, Joshua, Yao, Xiaobai, Ma, Hongxu, Mu, Lan, Ermon, Stefano, Ganu, Tanuja, Nambi, Akshay, Lao, Ni, Mai, Gengchen |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
GeoBS: Information-Theoretic Quantification of Geographic Bias in AI Models
por: Wang, Zhangyu, et al.
Publicado: (2025)
por: Wang, Zhangyu, et al.
Publicado: (2025)
LocDiff: Identifying Locations on Earth by Diffusing in the Hilbert Space
por: Wang, Zhangyu, et al.
Publicado: (2025)
por: Wang, Zhangyu, et al.
Publicado: (2025)
GAIR: Location-Aware Self-Supervised Contrastive Pre-Training with Geo-Aligned Implicit Representations
por: Liu, Zeping, et al.
Publicado: (2025)
por: Liu, Zeping, et al.
Publicado: (2025)
MC-GTA: Metric-Constrained Model-Based Clustering using Goodness-of-fit Tests with Autocorrelations
por: Wang, Zhangyu, et al.
Publicado: (2024)
por: Wang, Zhangyu, et al.
Publicado: (2024)
MMCTAgent: Multi-modal Critical Thinking Agent Framework for Complex Visual Reasoning
por: Kumar, Somnath, et al.
Publicado: (2024)
por: Kumar, Somnath, et al.
Publicado: (2024)
Img2Loc: Revisiting Image Geolocalization using Multi-modality Foundation Models and Image-based Retrieval-Augmented Generation
por: Zhou, Zhongliang, et al.
Publicado: (2024)
por: Zhou, Zhongliang, et al.
Publicado: (2024)
Probing the Information Theoretical Roots of Spatial Dependence Measures
por: Wang, Zhangyu, et al.
Publicado: (2024)
por: Wang, Zhangyu, et al.
Publicado: (2024)
PromptWizard: Task-Aware Prompt Optimization Framework
por: Agarwal, Eshaan, et al.
Publicado: (2024)
por: Agarwal, Eshaan, et al.
Publicado: (2024)
EnCortex: A General, Extensible and Scalable Framework for Decision Management in New-age Energy Systems
por: Roy, Millend, et al.
Publicado: (2025)
por: Roy, Millend, et al.
Publicado: (2025)
Do You See Me : A Multidimensional Benchmark for Evaluating Visual Perception in Multimodal LLMs
por: Kanade, Aditya, et al.
Publicado: (2025)
por: Kanade, Aditya, et al.
Publicado: (2025)
Text2Seg: Remote Sensing Image Semantic Segmentation via Text-Guided Visual Foundation Models
por: Zhang, Jielu, et al.
Publicado: (2023)
por: Zhang, Jielu, et al.
Publicado: (2023)
Bridging the Language Gap: Dynamic Learning Strategies for Improving Multilingual Performance in LLMs
por: Kumar, Somnath, et al.
Publicado: (2023)
por: Kumar, Somnath, et al.
Publicado: (2023)
Bridging the Gap: Dynamic Learning Strategies for Improving Multilingual Performance in LLMs
por: Kumar, Somnath, et al.
Publicado: (2024)
por: Kumar, Somnath, et al.
Publicado: (2024)
Chain-of-Thought Degrades Visual Spatial Reasoning Capabilities of Multimodal LLMs
por: Kancheti, Sai Srinivas, et al.
Publicado: (2026)
por: Kancheti, Sai Srinivas, et al.
Publicado: (2026)
SpatialMath: Spatial Comprehension-Infused Symbolic Reasoning for Mathematical Problem-Solving
por: Bajpai, Ashutosh, et al.
Publicado: (2026)
por: Bajpai, Ashutosh, et al.
Publicado: (2026)
Exposing Weak Links in Multi-Agent Systems under Adversarial Prompting
por: Arora, Nirmit, et al.
Publicado: (2025)
por: Arora, Nirmit, et al.
Publicado: (2025)
Multimodal Needle in a Haystack: Benchmarking Long-Context Capability of Multimodal Large Language Models
por: Wang, Hengyi, et al.
Publicado: (2024)
por: Wang, Hengyi, et al.
Publicado: (2024)
Faithful GRPO: Improving Visual Spatial Reasoning in Multimodal Language Models via Constrained Policy Optimization
por: Kancheti, Sai Srinivas, et al.
Publicado: (2026)
por: Kancheti, Sai Srinivas, et al.
Publicado: (2026)
RadPhi-3: Small Language Models for Radiology
por: Ranjit, Mercy, et al.
Publicado: (2024)
por: Ranjit, Mercy, et al.
Publicado: (2024)
Shiksha Copilot: Teacher-AI Collaboration for Curating and Customizing Lesson Plans in Low-Resource Schools
por: Dennison, Deepak Varuvel, et al.
Publicado: (2025)
por: Dennison, Deepak Varuvel, et al.
Publicado: (2025)
GeoLLM: Extracting Geospatial Knowledge from Large Language Models
por: Manvi, Rohin, et al.
Publicado: (2023)
por: Manvi, Rohin, et al.
Publicado: (2023)
Spatial-RAG: Spatial Retrieval Augmented Generation for Real-World Geospatial Reasoning Questions
por: Yu, Dazhou, et al.
Publicado: (2025)
por: Yu, Dazhou, et al.
Publicado: (2025)
TRAJGANR: Trajectory-Centric Urban Multimodal Learning via Geospatially Aligned Neural Representations
por: Siampou, Maria Despoina, et al.
Publicado: (2026)
por: Siampou, Maria Despoina, et al.
Publicado: (2026)
Designing Culturally Aligned AI Systems For Social Good in Non-Western Contexts
por: Dennison, Deepak Varuvel, et al.
Publicado: (2025)
por: Dennison, Deepak Varuvel, et al.
Publicado: (2025)
Spatial-Agent: Agentic Geo-spatial Reasoning with Scientific Core Concepts
por: Bao, Riyang, et al.
Publicado: (2026)
por: Bao, Riyang, et al.
Publicado: (2026)
TAMAS: Benchmarking Adversarial Risks in Multi-Agent LLM Systems
por: Kavathekar, Ishan, et al.
Publicado: (2025)
por: Kavathekar, Ishan, et al.
Publicado: (2025)
Evaluating LLMs' Mathematical Reasoning in Financial Document Question Answering
por: Srivastava, Pragya, et al.
Publicado: (2024)
por: Srivastava, Pragya, et al.
Publicado: (2024)
RAD-PHI2: Instruction Tuning PHI-2 for Radiology
por: Ranjit, Mercy, et al.
Publicado: (2024)
por: Ranjit, Mercy, et al.
Publicado: (2024)
Automated Image-Based Identification and Consistent Classification of Fire Patterns with Quantitative Shape Analysis and Spatial Location Identification
por: Liu, Pengkun, et al.
Publicado: (2024)
por: Liu, Pengkun, et al.
Publicado: (2024)
Geography According to ChatGPT -- How Generative AI Represents and Reasons about Geography
por: Janowicz, Krzysztof, et al.
Publicado: (2026)
por: Janowicz, Krzysztof, et al.
Publicado: (2026)
Self-Evolved Preference Optimization for Enhancing Mathematical Reasoning in Small Language Models
por: Singh, Joykirat, et al.
Publicado: (2025)
por: Singh, Joykirat, et al.
Publicado: (2025)
Exposing the Achilles' Heel: Evaluating LLMs Ability to Handle Mistakes in Mathematical Reasoning
por: Singh, Joykirat, et al.
Publicado: (2024)
por: Singh, Joykirat, et al.
Publicado: (2024)
Mind's Eye: A Benchmark of Visual Abstraction, Transformation and Composition for Multimodal LLMs
por: Sinha, Rohit, et al.
Publicado: (2026)
por: Sinha, Rohit, et al.
Publicado: (2026)
LAID: Lightweight AI-Generated Image Detection in Spatial and Spectral Domains
por: Chivaran, Nicholas, et al.
Publicado: (2025)
por: Chivaran, Nicholas, et al.
Publicado: (2025)
An Empirical Study on the Impact of Positional Encoding in Transformer-based Monaural Speech Enhancement
por: Zhang, Qiquan, et al.
Publicado: (2024)
por: Zhang, Qiquan, et al.
Publicado: (2024)
Neural Radiance Fields with Torch Units
por: Ni, Bingnan, et al.
Publicado: (2024)
por: Ni, Bingnan, et al.
Publicado: (2024)
Spatial-Regularization-Aware Dual-Branch Collaborative Inference for Training-Free OVSS in Remote Sensing Imagery
por: Wang, Jianzheng, et al.
Publicado: (2026)
por: Wang, Jianzheng, et al.
Publicado: (2026)
Locate-then-Merge: Neuron-Level Parameter Fusion for Mitigating Catastrophic Forgetting in Multimodal LLMs
por: Yu, Zeping, et al.
Publicado: (2025)
por: Yu, Zeping, et al.
Publicado: (2025)
Spatially Varying Gene Regulatory Networks via Bayesian Nonparametric Covariate-Dependent Directed Cyclic Graphical Models
por: Dawn, Trisha, et al.
Publicado: (2025)
por: Dawn, Trisha, et al.
Publicado: (2025)
Multiple Spatial‐Spectral Encoding Empowered High Fidelity Snapshot Spectral Imaging
por: Xuechan Lang, et al.
Publicado: (2025)
por: Xuechan Lang, et al.
Publicado: (2025)
Ejemplares similares
-
GeoBS: Information-Theoretic Quantification of Geographic Bias in AI Models
por: Wang, Zhangyu, et al.
Publicado: (2025) -
LocDiff: Identifying Locations on Earth by Diffusing in the Hilbert Space
por: Wang, Zhangyu, et al.
Publicado: (2025) -
GAIR: Location-Aware Self-Supervised Contrastive Pre-Training with Geo-Aligned Implicit Representations
por: Liu, Zeping, et al.
Publicado: (2025) -
MC-GTA: Metric-Constrained Model-Based Clustering using Goodness-of-fit Tests with Autocorrelations
por: Wang, Zhangyu, et al.
Publicado: (2024) -
MMCTAgent: Multi-modal Critical Thinking Agent Framework for Complex Visual Reasoning
por: Kumar, Somnath, et al.
Publicado: (2024)