GeoLLM-Engine: A Realistic Environment for Building Geospatial Copilots
Fuente:
arXiv
Guardado en:
| Autores principales: | Singh, Simranjit, Fore, Michael, Stamoulis, Dimitrios |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Evaluating Tool-Augmented Agents in Remote Sensing Platforms
por: Singh, Simranjit, et al.
Publicado: (2024)
por: Singh, Simranjit, et al.
Publicado: (2024)
GeckOpt: LLM System Efficiency via Intent-Based Tool Selection
por: Fore, Michael, et al.
Publicado: (2024)
por: Fore, Michael, et al.
Publicado: (2024)
An LLM-Tool Compiler for Fused Parallel Function Calling
por: Singh, Simranjit, et al.
Publicado: (2024)
por: Singh, Simranjit, et al.
Publicado: (2024)
GeoLLM: Extracting Geospatial Knowledge from Large Language Models
por: Manvi, Rohin, et al.
Publicado: (2023)
por: Manvi, Rohin, et al.
Publicado: (2023)
WebArena: A Realistic Web Environment for Building Autonomous Agents
por: Zhou, Shuyan, et al.
Publicado: (2023)
por: Zhou, Shuyan, et al.
Publicado: (2023)
Evaluating Zero-Shot GPT-4V Performance on 3D Visual Question Answering Benchmarks
por: Singh, Simranjit, et al.
Publicado: (2024)
por: Singh, Simranjit, et al.
Publicado: (2024)
Unlearning Climate Misinformation in Large Language Models
por: Fore, Michael, et al.
Publicado: (2024)
por: Fore, Michael, et al.
Publicado: (2024)
Multi-Agent Geospatial Copilots for Remote Sensing Workflows
por: Lee, Chaehong, et al.
Publicado: (2025)
por: Lee, Chaehong, et al.
Publicado: (2025)
Transformer Copilot: Learning from The Mistake Log in LLM Fine-tuning
por: Zou, Jiaru, et al.
Publicado: (2025)
por: Zou, Jiaru, et al.
Publicado: (2025)
GeoFlow: Agentic Workflow Automation for Geospatial Tasks
por: Bhattaram, Amulya, et al.
Publicado: (2025)
por: Bhattaram, Amulya, et al.
Publicado: (2025)
Geo-OLM: Enabling Sustainable Earth Observation Studies with Cost-Efficient Open Language Models & State-Driven Workflows
por: Stamoulis, Dimitrios, et al.
Publicado: (2025)
por: Stamoulis, Dimitrios, et al.
Publicado: (2025)
Robustly Improving LLM Fairness in Realistic Settings via Interpretability
por: Karvonen, Adam, et al.
Publicado: (2025)
por: Karvonen, Adam, et al.
Publicado: (2025)
ED-Copilot: Reduce Emergency Department Wait Time with Language Model Diagnostic Assistance
por: Sun, Liwen, et al.
Publicado: (2024)
por: Sun, Liwen, et al.
Publicado: (2024)
MLR-Copilot: Autonomous Machine Learning Research based on Large Language Models Agents
por: Li, Ruochen, et al.
Publicado: (2024)
por: Li, Ruochen, et al.
Publicado: (2024)
REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations
por: Liang, Buyun, et al.
Publicado: (2026)
por: Liang, Buyun, et al.
Publicado: (2026)
Mitigating Geospatial Knowledge Hallucination in Large Language Models: Benchmarking and Dynamic Factuality Aligning
por: Wang, Shengyuan, et al.
Publicado: (2025)
por: Wang, Shengyuan, et al.
Publicado: (2025)
Dive into the Agent Matrix: A Realistic Evaluation of Self-Replication Risk in LLM Agents
por: Zhang, Boxuan, et al.
Publicado: (2025)
por: Zhang, Boxuan, et al.
Publicado: (2025)
Evaluating Language-Model Agents on Realistic Autonomous Tasks
por: Kinniment, Megan, et al.
Publicado: (2023)
por: Kinniment, Megan, et al.
Publicado: (2023)
LLM-dCache: Improving Tool-Augmented LLMs with GPT-Driven Localized Data Caching
por: Singh, Simranjit, et al.
Publicado: (2024)
por: Singh, Simranjit, et al.
Publicado: (2024)
Uncertainty Quantification for Language Models: A Suite of Black-Box, White-Box, LLM Judge, and Ensemble Scorers
por: Bouchard, Dylan, et al.
Publicado: (2025)
por: Bouchard, Dylan, et al.
Publicado: (2025)
CodeRefine: A Pipeline for Enhancing LLM-Generated Code Implementations of Research Papers
por: Trofimova, Ekaterina, et al.
Publicado: (2024)
por: Trofimova, Ekaterina, et al.
Publicado: (2024)
EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis
por: Song, Xiaoshuai, et al.
Publicado: (2026)
por: Song, Xiaoshuai, et al.
Publicado: (2026)
Multi-LLM QA with Embodied Exploration
por: Patel, Bhrij, et al.
Publicado: (2024)
por: Patel, Bhrij, et al.
Publicado: (2024)
Learning to Trust the Crowd: A Multi-Model Consensus Reasoning Engine for Large Language Models
por: Kallem, Pranav
Publicado: (2026)
por: Kallem, Pranav
Publicado: (2026)
ORPO-Distill: Mixed-Policy Preference Optimization for Cross-Architecture LLM Distillation
por: Singh, Aasheesh, et al.
Publicado: (2025)
por: Singh, Aasheesh, et al.
Publicado: (2025)
Multimodal Generative Engine Optimization: Rank Manipulation for Vision-Language Model Rankers
por: Du, Yixuan, et al.
Publicado: (2026)
por: Du, Yixuan, et al.
Publicado: (2026)
WebChoreArena: Evaluating Web Browsing Agents on Realistic Tedious Web Tasks
por: Miyai, Atsuyuki, et al.
Publicado: (2025)
por: Miyai, Atsuyuki, et al.
Publicado: (2025)
Code Comprehension then Auditing for Unsupervised LLM Evaluation
por: Patel, Bhrij, et al.
Publicado: (2024)
por: Patel, Bhrij, et al.
Publicado: (2024)
Disambiguation-Centric Finetuning Makes Enterprise Tool-Calling LLMs More Realistic and Less Risky
por: Hathidara, Ashutosh, et al.
Publicado: (2025)
por: Hathidara, Ashutosh, et al.
Publicado: (2025)
Repo2Run: Automated Building Executable Environment for Code Repository at Scale
por: Hu, Ruida, et al.
Publicado: (2025)
por: Hu, Ruida, et al.
Publicado: (2025)
CaRT: Teaching LLM Agents to Know When They Know Enough
por: Liu, Grace, et al.
Publicado: (2025)
por: Liu, Grace, et al.
Publicado: (2025)
SQuARE: Sequential Question Answering Reasoning Engine for Enhanced Chain-of-Thought in Large Language Models
por: Fleischer, Daniel, et al.
Publicado: (2025)
por: Fleischer, Daniel, et al.
Publicado: (2025)
Functional Entropy: Predicting Functional Correctness in LLM-Generated Code with Uncertainty Quantification
por: Bouchard, Dylan, et al.
Publicado: (2026)
por: Bouchard, Dylan, et al.
Publicado: (2026)
Building Production-Ready Probes For Gemini
por: Kramár, János, et al.
Publicado: (2026)
por: Kramár, János, et al.
Publicado: (2026)
Set-LLM: A Permutation-Invariant LLM
por: Egressy, Beni, et al.
Publicado: (2025)
por: Egressy, Beni, et al.
Publicado: (2025)
AutoEnv: Automated Environments for Measuring Cross-Environment Agent Learning
por: Zhang, Jiayi, et al.
Publicado: (2025)
por: Zhang, Jiayi, et al.
Publicado: (2025)
ClawGym: A Scalable Framework for Building Effective Claw Agents
por: Bai, Fei, et al.
Publicado: (2026)
por: Bai, Fei, et al.
Publicado: (2026)
SQLBarber: A System Leveraging Large Language Models to Generate Customized and Realistic SQL Workloads
por: Lao, Jiale, et al.
Publicado: (2025)
por: Lao, Jiale, et al.
Publicado: (2025)
Prakriti200: A Questionnaire-Based Dataset of 200 Ayurvedic Prakriti Assessments
por: Singh, Aryan Kumar, et al.
Publicado: (2025)
por: Singh, Aryan Kumar, et al.
Publicado: (2025)
End-to-end Text-to-SQL Generation within an Analytics Insight Engine
por: Maamari, Karime, et al.
Publicado: (2024)
por: Maamari, Karime, et al.
Publicado: (2024)
Ejemplares similares
-
Evaluating Tool-Augmented Agents in Remote Sensing Platforms
por: Singh, Simranjit, et al.
Publicado: (2024) -
GeckOpt: LLM System Efficiency via Intent-Based Tool Selection
por: Fore, Michael, et al.
Publicado: (2024) -
An LLM-Tool Compiler for Fused Parallel Function Calling
por: Singh, Simranjit, et al.
Publicado: (2024) -
GeoLLM: Extracting Geospatial Knowledge from Large Language Models
por: Manvi, Rohin, et al.
Publicado: (2023) -
WebArena: A Realistic Web Environment for Building Autonomous Agents
por: Zhou, Shuyan, et al.
Publicado: (2023)