DSBC : Data Science task Benchmarking with Context engineering
Fuente:
arXiv
Saved in:
| Main Authors: | Kadiyala, Ram Mohan Rao, Gupta, Siddhant, Purbey, Jebish, Martini, Giulio, Shafique, Ali, Debnath, Suman, Farooq, Hamza |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Multilingual Capabilities with Cultural and Local Knowledge in Large Language Models While Enhancing Native Performance
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025)
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025)
Uncovering Cultural Representation Disparities in Vision-Language Models
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025)
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025)
1-800-SHARED-TASKS at RegNLP: Lexical Reranking of Semantic Retrieval (LeSeR) for Regulatory Question Answering
by: Purbey, Jebish, et al.
Published: (2024)
by: Purbey, Jebish, et al.
Published: (2024)
Protecting Context and Prompts: Deterministic Security for Non-Deterministic AI
by: Rajagopalan, Mohan, et al.
Published: (2026)
by: Rajagopalan, Mohan, et al.
Published: (2026)
Geode: A Zero-shot Geospatial Question-Answering Agent with Explicit Reasoning and Precise Spatio-Temporal Retrieval
by: Gupta, Devashish Vikas, et al.
Published: (2024)
by: Gupta, Devashish Vikas, et al.
Published: (2024)
SeQwen at the Financial Misinformation Detection Challenge Task: Sequential Learning for Claim Verification and Explanation Generation in Financial Domains
by: Purbey, Jebish, et al.
Published: (2024)
by: Purbey, Jebish, et al.
Published: (2024)
Fairness Driven Multi-Agent Path Finding Problem
by: Anand, Aditi, et al.
Published: (2026)
by: Anand, Aditi, et al.
Published: (2026)
SWE-WebDevBench: Evaluating Coding Agent Application Platforms as Virtual Software Agencies
by: Saxena, Siddhant, et al.
Published: (2026)
by: Saxena, Siddhant, et al.
Published: (2026)
Minimizing Regret in Billboard Advertisement under Zonal Influence Constraint
by: Ali, Dildar, et al.
Published: (2024)
by: Ali, Dildar, et al.
Published: (2024)
Satellites swarm cooperation for pursuit-attachment tasks with transformer-based reinforcement learning
by: Li, yonghao
Published: (2024)
by: Li, yonghao
Published: (2024)
Multi-agent systems for chemical engineering: A review and perspective
by: Rupprecht, Sophia, et al.
Published: (2025)
by: Rupprecht, Sophia, et al.
Published: (2025)
Modeling Prejudice and Its Effect on Societal Prosperity
by: Mohan, Deep Inder, et al.
Published: (2021)
by: Mohan, Deep Inder, et al.
Published: (2021)
Distributed Online Task Assignment via Inexact ADMM for unplanned online tasks and its Applications to Security
by: Yang, Ziqi, et al.
Published: (2025)
by: Yang, Ziqi, et al.
Published: (2025)
The price of decentralization in managing engineering systems through multi-agent reinforcement learning
by: Bhustali, Prateek, et al.
Published: (2026)
by: Bhustali, Prateek, et al.
Published: (2026)
Who is Helping Whom? Analyzing Inter-dependencies to Evaluate Cooperation in Human-AI Teaming
by: Biswas, Upasana, et al.
Published: (2025)
by: Biswas, Upasana, et al.
Published: (2025)
Argo: Efficient Importance Labeling for Enterprise Email Systems
by: Ray, Siddhant, et al.
Published: (2026)
by: Ray, Siddhant, et al.
Published: (2026)
Group Trip Planning Query Problem with Multimodal Journey
by: Ali, Dildar, et al.
Published: (2025)
by: Ali, Dildar, et al.
Published: (2025)
Scalable Submodular Policy Optimization via Pruned Submodularity Graph
by: Anand, Aditi, et al.
Published: (2025)
by: Anand, Aditi, et al.
Published: (2025)
Distributed Networked Multi-task Learning
by: Hong, Lingzhou, et al.
Published: (2024)
by: Hong, Lingzhou, et al.
Published: (2024)
Social Norm Reasoning in Multimodal Language Models: An Evaluation
by: Chowdhury, Oishik, et al.
Published: (2026)
by: Chowdhury, Oishik, et al.
Published: (2026)
Bridging Finite and Infinite-Horizon Nash Equilibria in Linear Quadratic Games
by: Salizzoni, Giulio, et al.
Published: (2025)
by: Salizzoni, Giulio, et al.
Published: (2025)
LLMDR: Large language model driven framework for missing data recovery in mixed data under low resource regime
by: Keshav, Durga, et al.
Published: (2026)
by: Keshav, Durga, et al.
Published: (2026)
Game-theoretic Occlusion-Aware Motion Planning: an Efficient Hybrid-Information Approach
by: Gupta, Kushagra, et al.
Published: (2023)
by: Gupta, Kushagra, et al.
Published: (2023)
Socialized Learning and Emergent Behaviors in Multi-Agent Systems based on Multimodal Large Language Models
by: Akin, Sureyya, et al.
Published: (2025)
by: Akin, Sureyya, et al.
Published: (2025)
Task Allocation of UAVs for Monitoring Missions via Hardware-in-the-Loop Simulation and Experimental Validation
by: Chakraa, Hamza, et al.
Published: (2025)
by: Chakraa, Hamza, et al.
Published: (2025)
GenGrid: A Generalised Distributed Experimental Environmental Grid for Swarm Robotics
by: Kedia, Pranav, et al.
Published: (2025)
by: Kedia, Pranav, et al.
Published: (2025)
A segment anchoring-based balancing algorithm for agricultural multi-robot task allocation with energy constraints
by: Chen, Peng, et al.
Published: (2025)
by: Chen, Peng, et al.
Published: (2025)
A Multi-LLM Orchestration Engine for Personalized, Context-Rich Assistance
by: Rasal, Sumedh
Published: (2024)
by: Rasal, Sumedh
Published: (2024)
Evaluating LLM Alignment With Human Trust Models
by: Debnath, Anushka, et al.
Published: (2026)
by: Debnath, Anushka, et al.
Published: (2026)
Towards Regret Free Slot Allocation in Billboard Advertisement
by: Ali, Dildar, et al.
Published: (2024)
by: Ali, Dildar, et al.
Published: (2024)
Robust and Fine-Grained Detection of AI Generated Texts
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025)
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025)
Authenticated Workflows: A Systems Approach to Protecting Agentic AI
by: Rajagopalan, Mohan, et al.
Published: (2026)
by: Rajagopalan, Mohan, et al.
Published: (2026)
Joint Optimization of Autonomous Electric Vehicle Fleet Operations and Charging Station Siting
by: Luke, Justin, et al.
Published: (2021)
by: Luke, Justin, et al.
Published: (2021)
Context-Aware Agent-based Model for Smart Long Distance Transport System
by: Raees, Muhammad, et al.
Published: (2024)
by: Raees, Muhammad, et al.
Published: (2024)
EASE Configuration Facilitates A Reproducible Science of LLM Social Simulations
by: Sarangi, Sneheel, et al.
Published: (2026)
by: Sarangi, Sneheel, et al.
Published: (2026)
DataCross: A Unified Benchmark and Agent Framework for Cross-Modal Heterogeneous Data Analysis
by: Qi, Ruyi, et al.
Published: (2026)
by: Qi, Ruyi, et al.
Published: (2026)
AgentWebBench: Benchmarking Multi-Agent Coordination in Agentic Web
by: Zhong, Shanshan, et al.
Published: (2026)
by: Zhong, Shanshan, et al.
Published: (2026)
Heterogeneous Multi-Agent Task-Assignment with Uncertain Execution Times and Preferences
by: Wei, Qinshuang, et al.
Published: (2025)
by: Wei, Qinshuang, et al.
Published: (2025)
Hybrid Training for Enhanced Multi-task Generalization in Multi-agent Reinforcement Learning
by: Zhang, Mingliang, et al.
Published: (2024)
by: Zhang, Mingliang, et al.
Published: (2024)
Learning Communication Skills in Multi-task Multi-agent Deep Reinforcement Learning
by: Zhu, Changxi, et al.
Published: (2025)
by: Zhu, Changxi, et al.
Published: (2025)
Similar Items
-
Improving Multilingual Capabilities with Cultural and Local Knowledge in Large Language Models While Enhancing Native Performance
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025) -
Uncovering Cultural Representation Disparities in Vision-Language Models
by: Kadiyala, Ram Mohan Rao, et al.
Published: (2025) -
1-800-SHARED-TASKS at RegNLP: Lexical Reranking of Semantic Retrieval (LeSeR) for Regulatory Question Answering
by: Purbey, Jebish, et al.
Published: (2024) -
Protecting Context and Prompts: Deterministic Security for Non-Deterministic AI
by: Rajagopalan, Mohan, et al.
Published: (2026) -
Geode: A Zero-shot Geospatial Question-Answering Agent with Explicit Reasoning and Precise Spatio-Temporal Retrieval
by: Gupta, Devashish Vikas, et al.
Published: (2024)