A Tale of LLMs and Induced Small Proxies: Scalable Agents for Knowledge Mining
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Sipeng, Yun, Longfei, Wang, Zilong, Shang, Jingbo, Peng, Letian |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Price of Format: Diversity Collapse in LLMs
by: Yun, Longfei, et al.
Published: (2025)
by: Yun, Longfei, et al.
Published: (2025)
Memorize or Generalize? Evaluating LLM Code Generation with Code Rewriting
by: Zhang, Lizhe, et al.
Published: (2025)
by: Zhang, Lizhe, et al.
Published: (2025)
UltraGen: Extremely Fine-grained Controllable Generation via Attribute Reconstruction and Global Preference Optimization
by: Yun, Longfei, et al.
Published: (2025)
by: Yun, Longfei, et al.
Published: (2025)
Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step
by: Zhong, Li, et al.
Published: (2024)
by: Zhong, Li, et al.
Published: (2024)
Learn from Failure: Fine-Tuning LLMs with Trial-and-Error Data for Intuitionistic Propositional Logic Proving
by: An, Chenyang, et al.
Published: (2024)
by: An, Chenyang, et al.
Published: (2024)
RRO: LLM Agent Optimization Through Rising Reward Trajectories
by: Wang, Zilong, et al.
Published: (2025)
by: Wang, Zilong, et al.
Published: (2025)
Toward Scalable Verifiable Reward: Proxy State-Based Evaluation for Multi-turn Tool-Calling LLM Agents
by: Chuang, Yun-Shiuan, et al.
Published: (2026)
by: Chuang, Yun-Shiuan, et al.
Published: (2026)
Predicting LLM Reasoning Performance with Small Proxy Model
by: Koh, Woosung, et al.
Published: (2025)
by: Koh, Woosung, et al.
Published: (2025)
Toward Student-Oriented Teacher Network Training For Knowledge Distillation
by: Dong, Chengyu, et al.
Published: (2022)
by: Dong, Chengyu, et al.
Published: (2022)
Controllable Data Augmentation for Few-Shot Text Mining with Chain-of-Thought Attribute Manipulation
by: Peng, Letian, et al.
Published: (2023)
by: Peng, Letian, et al.
Published: (2023)
Translation as a Scalable Proxy for Multilingual Evaluation
by: Issaka, Sheriff, et al.
Published: (2026)
by: Issaka, Sheriff, et al.
Published: (2026)
QuadrupedGPT: Towards a Versatile Quadruped Agent in Open-ended Worlds
by: Mei, Yuting, et al.
Published: (2024)
by: Mei, Yuting, et al.
Published: (2024)
Optimization and Scalability of Collaborative Filtering Algorithms in Large Language Models
by: Yang, Haowei, et al.
Published: (2024)
by: Yang, Haowei, et al.
Published: (2024)
MineEvolve: Self-Evolution with Accumulated Knowledge for Long-Horizon Embodied Minecraft Agents
by: Xie, Zhengwei, et al.
Published: (2026)
by: Xie, Zhengwei, et al.
Published: (2026)
GUI-explorer: Autonomous Exploration and Mining of Transition-aware Knowledge for GUI Agent
by: Xie, Bin, et al.
Published: (2025)
by: Xie, Bin, et al.
Published: (2025)
ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection
by: Li, Jiaqi, et al.
Published: (2025)
by: Li, Jiaqi, et al.
Published: (2025)
Cuckoo: An IE Free Rider Hatched by Massive Nutrition in LLM's Nest
by: Peng, Letian, et al.
Published: (2025)
by: Peng, Letian, et al.
Published: (2025)
Deriving Character Logic from Storyline as Codified Decision Trees
by: Peng, Letian, et al.
Published: (2026)
by: Peng, Letian, et al.
Published: (2026)
Codified Foreshadowing-Payoff Text Generation
by: Yun, Longfei, et al.
Published: (2026)
by: Yun, Longfei, et al.
Published: (2026)
Can RL Teach Long-Horizon Reasoning to LLMs? Expressiveness Is Key
by: Wang, Tianle, et al.
Published: (2026)
by: Wang, Tianle, et al.
Published: (2026)
AI-native Memory: A Pathway from LLMs Towards AGI
by: Shang, Jingbo, et al.
Published: (2024)
by: Shang, Jingbo, et al.
Published: (2024)
Scalable Exact Verification of Optimization Proxies for Large-Scale Optimal Power Flow
by: Nellikkath, Rahul, et al.
Published: (2024)
by: Nellikkath, Rahul, et al.
Published: (2024)
FedProxy: Federated Fine-Tuning of LLMs via Proxy SLMs and Heterogeneity-Aware Fusion
by: Fan, Tao, et al.
Published: (2026)
by: Fan, Tao, et al.
Published: (2026)
ArchPilot: A Proxy-Guided Multi-Agent Approach for Machine Learning Engineering
by: Yuan, Zhuowen, et al.
Published: (2025)
by: Yuan, Zhuowen, et al.
Published: (2025)
Small LLMs Are Weak Tool Learners: A Multi-LLM Agent
by: Shen, Weizhou, et al.
Published: (2024)
by: Shen, Weizhou, et al.
Published: (2024)
W-PCA Based Gradient-Free Proxy for Efficient Search of Lightweight Language Models
by: Wang, Shang
Published: (2025)
by: Wang, Shang
Published: (2025)
Does RAG Know When Retrieval Is Wrong? Diagnosing Context Compliance under Knowledge Conflict
by: Chen, Yihang, et al.
Published: (2026)
by: Chen, Yihang, et al.
Published: (2026)
AgentEval: Generative Agents as Reliable Proxies for Human Evaluation of AI-Generated Content
by: Vu, Thanh, et al.
Published: (2025)
by: Vu, Thanh, et al.
Published: (2025)
From Unstructured Communication to Intelligent RAG: Multi-Agent Automation for Supply Chain Knowledge Bases
by: Zhang, Yao, et al.
Published: (2025)
by: Zhang, Yao, et al.
Published: (2025)
IntPro: A Proxy Agent for Context-Aware Intent Understanding via Retrieval-conditioned Inference
by: Liu, Guanming, et al.
Published: (2026)
by: Liu, Guanming, et al.
Published: (2026)
Revitalizing Black-Box Interpretability: Actionable Interpretability for LLMs via Proxy Models
by: Liu, Junhao, et al.
Published: (2025)
by: Liu, Junhao, et al.
Published: (2025)
Fast and Scalable Analytical Diffusion
by: Shang, Xinyi, et al.
Published: (2026)
by: Shang, Xinyi, et al.
Published: (2026)
Optimal-Agent-Selection: State-Aware Routing Framework for Efficient Multi-Agent Collaboration
by: Wang, Jingbo, et al.
Published: (2025)
by: Wang, Jingbo, et al.
Published: (2025)
Bidirectional Curriculum Generation: A Multi-Agent Framework for Data-Efficient Mathematical Reasoning
by: Hu, Boren, et al.
Published: (2026)
by: Hu, Boren, et al.
Published: (2026)
M2-PALE: A Framework for Explaining Multi-Agent MCTS--Minimax Hybrids via Process Mining and LLMs
by: Qian, Yiyu, et al.
Published: (2026)
by: Qian, Yiyu, et al.
Published: (2026)
In-Context Examples Suppress Scientific Knowledge Recall in LLMs
by: Jang, Chaemin, et al.
Published: (2026)
by: Jang, Chaemin, et al.
Published: (2026)
Balancing Knowledge Delivery and Emotional Comfort in Healthcare Conversational Systems
by: Tsai, Shang-Chi, et al.
Published: (2025)
by: Tsai, Shang-Chi, et al.
Published: (2025)
Codifying Character Logic in Role-Playing
by: Peng, Letian, et al.
Published: (2025)
by: Peng, Letian, et al.
Published: (2025)
Tracing the Roots: A Multi-Agent Framework for Uncovering Data Lineage in Post-Training LLMs
by: Li, Yu, et al.
Published: (2026)
by: Li, Yu, et al.
Published: (2026)
Incubating Text Classifiers Following User Instruction with Nothing but LLM
by: Peng, Letian, et al.
Published: (2024)
by: Peng, Letian, et al.
Published: (2024)
Similar Items
-
The Price of Format: Diversity Collapse in LLMs
by: Yun, Longfei, et al.
Published: (2025) -
Memorize or Generalize? Evaluating LLM Code Generation with Code Rewriting
by: Zhang, Lizhe, et al.
Published: (2025) -
UltraGen: Extremely Fine-grained Controllable Generation via Attribute Reconstruction and Global Preference Optimization
by: Yun, Longfei, et al.
Published: (2025) -
Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step
by: Zhong, Li, et al.
Published: (2024) -
Learn from Failure: Fine-Tuning LLMs with Trial-and-Error Data for Intuitionistic Propositional Logic Proving
by: An, Chenyang, et al.
Published: (2024)