Dataforge: Agentic Platform for Autonomous Data Engineering
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Xinyuan, Cao, Hongyu, Liu, Kunpeng, Fu, Yanjie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mitigating Shortcut Reasoning in Language Models: A Gradient-Aware Training Approach
by: Cao, Hongyu, et al.
Published: (2026)
by: Cao, Hongyu, et al.
Published: (2026)
Evolving Demonstration Optimization for Chain-of-Thought Feature Transformation
by: Wang, Xinyuan, et al.
Published: (2026)
by: Wang, Xinyuan, et al.
Published: (2026)
Autonomous Data Agents: A New Opportunity for Smart Data
by: Fu, Yanjie, et al.
Published: (2025)
by: Fu, Yanjie, et al.
Published: (2025)
Causally-Guided Diffusion for Stable Feature Selection
by: Malarkkan, Arun Vignesh, et al.
Published: (2026)
by: Malarkkan, Arun Vignesh, et al.
Published: (2026)
Topology-aware Reinforcement Feature Space Reconstruction for Graph Data
by: Ying, Wangyang, et al.
Published: (2024)
by: Ying, Wangyang, et al.
Published: (2024)
Data-Efficient Symbolic Regression via Foundation Model Distillation
by: Ying, Wangyang, et al.
Published: (2025)
by: Ying, Wangyang, et al.
Published: (2025)
Sim2Act: Robust Simulation-to-Decision Learning via Adversarial Calibration and Group-Relative Perturbation
by: Cao, Hongyu, et al.
Published: (2026)
by: Cao, Hongyu, et al.
Published: (2026)
Agentic Feature Augmentation: Unifying Selection and Generation with Teaming, Planning, and Memories
by: Gong, Nanxu, et al.
Published: (2025)
by: Gong, Nanxu, et al.
Published: (2025)
Efficient Post-Training Refinement of Latent Reasoning in Large Language Models
by: Wang, Xinyuan, et al.
Published: (2025)
by: Wang, Xinyuan, et al.
Published: (2025)
AgentOS: From Application Silos to a Natural Language-Driven Data Ecosystem
by: Liu, Rui, et al.
Published: (2026)
by: Liu, Rui, et al.
Published: (2026)
City Editing: Hierarchical Agentic Execution for Dependency-Aware Urban Geospatial Modification
by: Liu, Rui, et al.
Published: (2026)
by: Liu, Rui, et al.
Published: (2026)
LLM-Enhanced User-Item Interactions: Leveraging Edge Information for Optimized Recommendations
by: Wang, Xinyuan, et al.
Published: (2024)
by: Wang, Xinyuan, et al.
Published: (2024)
Towards Data-Centric AI: A Comprehensive Survey of Traditional, Reinforcement, and Generative Approaches for Tabular Data Transformation
by: Wang, Dongjie, et al.
Published: (2025)
by: Wang, Dongjie, et al.
Published: (2025)
A Comprehensive Survey on Data Augmentation
by: Wang, Zaitian, et al.
Published: (2024)
by: Wang, Zaitian, et al.
Published: (2024)
Towards Urban Planing AI Agent in the Age of Agentic AI
by: Liu, Rui, et al.
Published: (2025)
by: Liu, Rui, et al.
Published: (2025)
Cognitive Platform Engineering for Autonomous Cloud Operations
by: Punniyamoorthy, Vinoth, et al.
Published: (2026)
by: Punniyamoorthy, Vinoth, et al.
Published: (2026)
DRNet: A Decision-Making Method for Autonomous Lane Changingwith Deep Reinforcement Learning
by: Xu, Kunpeng, et al.
Published: (2023)
by: Xu, Kunpeng, et al.
Published: (2023)
Supply Chain Optimization via Generative Simulation and Iterative Decision Policies
by: Bai, Haoyue, et al.
Published: (2025)
by: Bai, Haoyue, et al.
Published: (2025)
Knockoff-Guided Feature Selection via A Single Pre-trained Reinforced Agent
by: Wang, Xinyuan, et al.
Published: (2024)
by: Wang, Xinyuan, et al.
Published: (2024)
AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models
by: Wang, Yixu, et al.
Published: (2025)
by: Wang, Yixu, et al.
Published: (2025)
StaRPO: Stability-Augmented Reinforcement Policy Optimization
by: Zhang, Jinghan, et al.
Published: (2026)
by: Zhang, Jinghan, et al.
Published: (2026)
Causally-Guided Automated Feature Engineering with Multi-Agent Reinforcement Learning
by: Malarkkan, Arun Vignesh, et al.
Published: (2026)
by: Malarkkan, Arun Vignesh, et al.
Published: (2026)
Bridging the Domain Gap in Equation Distillation with Reinforcement Feedback
by: Ying, Wangyang, et al.
Published: (2025)
by: Ying, Wangyang, et al.
Published: (2025)
LLM-ML Teaming: Integrated Symbolic Decoding and Gradient Search for Valid and Stable Generative Feature Transformation
by: Wang, Xinyuan, et al.
Published: (2025)
by: Wang, Xinyuan, et al.
Published: (2025)
Generative AI Meets Future Cities: Towards an Era of Autonomous Urban Intelligence
by: Wang, Dongjie, et al.
Published: (2023)
by: Wang, Dongjie, et al.
Published: (2023)
Demystifying the Lifecycle of Failures in Platform-Orchestrated Agentic Workflows
by: Ma, Xuyan, et al.
Published: (2025)
by: Ma, Xuyan, et al.
Published: (2025)
The Auton Agentic AI Framework
by: Cao, Sheng, et al.
Published: (2026)
by: Cao, Sheng, et al.
Published: (2026)
A Survey on Data-Centric AI: Tabular Learning from Reinforcement Learning and Generative AI Perspective
by: Ying, Wangyang, et al.
Published: (2025)
by: Ying, Wangyang, et al.
Published: (2025)
DeepTravel: An End-to-End Agentic Reinforcement Learning Framework for Autonomous Travel Planning Agents
by: Ning, Yansong, et al.
Published: (2025)
by: Ning, Yansong, et al.
Published: (2025)
Agentic AI Scientists Are Not Built For Autonomous Scientific Discovery
by: Bisht, Harshit, et al.
Published: (2026)
by: Bisht, Harshit, et al.
Published: (2026)
MetAdv: A Unified and Interactive Adversarial Testing Platform for Autonomous Driving
by: Liu, Aishan, et al.
Published: (2025)
by: Liu, Aishan, et al.
Published: (2025)
DeepAnalyze: Agentic Large Language Models for Autonomous Data Science
by: Zhang, Shaolei, et al.
Published: (2025)
by: Zhang, Shaolei, et al.
Published: (2025)
Haystack Engineering: Context Engineering for Heterogeneous and Agentic Long-Context Evaluation
by: Li, Mufei, et al.
Published: (2025)
by: Li, Mufei, et al.
Published: (2025)
AgenticData: An Agentic Data Analytics System for Heterogeneous Data
by: Sun, Ji, et al.
Published: (2025)
by: Sun, Ji, et al.
Published: (2025)
An Agentic Framework for Autonomous Materials Computation
by: Xia, Zeyu, et al.
Published: (2025)
by: Xia, Zeyu, et al.
Published: (2025)
CEDAR: Context Engineering for Agentic Data Science
by: Roy, Rishiraj Saha, et al.
Published: (2026)
by: Roy, Rishiraj Saha, et al.
Published: (2026)
Beyond Retrieval: Modeling Confidence Decay and Deterministic Agentic Platforms in Generative Engine Optimization
by: Zhao, XinYu, et al.
Published: (2026)
by: Zhao, XinYu, et al.
Published: (2026)
Kinematics-Aware Latent World Models for Data-Efficient Autonomous Driving
by: Li, Jiazhuo, et al.
Published: (2026)
by: Li, Jiazhuo, et al.
Published: (2026)
InternAgent-1.5: A Unified Agentic Framework for Long-Horizon Autonomous Scientific Discovery
by: Feng, Shiyang, et al.
Published: (2026)
by: Feng, Shiyang, et al.
Published: (2026)
Provable Long-Range Benefits of Next-Token Prediction
by: Cao, Xinyuan, et al.
Published: (2025)
by: Cao, Xinyuan, et al.
Published: (2025)
Similar Items
-
Mitigating Shortcut Reasoning in Language Models: A Gradient-Aware Training Approach
by: Cao, Hongyu, et al.
Published: (2026) -
Evolving Demonstration Optimization for Chain-of-Thought Feature Transformation
by: Wang, Xinyuan, et al.
Published: (2026) -
Autonomous Data Agents: A New Opportunity for Smart Data
by: Fu, Yanjie, et al.
Published: (2025) -
Causally-Guided Diffusion for Stable Feature Selection
by: Malarkkan, Arun Vignesh, et al.
Published: (2026) -
Topology-aware Reinforcement Feature Space Reconstruction for Graph Data
by: Ying, Wangyang, et al.
Published: (2024)