Logic Distillation: Learning from Code Function by Function for Decision-making Tasks
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Dong, Zhang, Shilin, Gao, Fei, Zhuang, Yueting, Tang, Siliang, Liu, Qidong, Xu, Mingliang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Improving Large Models with Small models: Lower Costs and Better Performance
di: Chen, Dong, et al.
Pubblicazione: (2024)
di: Chen, Dong, et al.
Pubblicazione: (2024)
KKA: Improving Vision Anomaly Detection through Anomaly-related Knowledge from Large Language Models
di: Chen, Dong, et al.
Pubblicazione: (2025)
di: Chen, Dong, et al.
Pubblicazione: (2025)
T2S-GPT: Dynamic Vector Quantization for Autoregressive Sign Language Production from Text
di: Yin, Aoxiong, et al.
Pubblicazione: (2024)
di: Yin, Aoxiong, et al.
Pubblicazione: (2024)
IDEAL: Leveraging Infinite and Dynamic Characterizations of Large Language Models for Query-focused Summarization
di: Cao, Jie, et al.
Pubblicazione: (2024)
di: Cao, Jie, et al.
Pubblicazione: (2024)
Bridging Local Details and Global Context in Text-Attributed Graphs
di: Wang, Yaoke, et al.
Pubblicazione: (2024)
di: Wang, Yaoke, et al.
Pubblicazione: (2024)
DuetRAG: Collaborative Retrieval-Augmented Generation
di: Jiao, Dian, et al.
Pubblicazione: (2024)
di: Jiao, Dian, et al.
Pubblicazione: (2024)
InstructVid2Vid: Controllable Video Editing with Natural Language Instructions
di: Qin, Bosheng, et al.
Pubblicazione: (2023)
di: Qin, Bosheng, et al.
Pubblicazione: (2023)
WorldGPT: Empowering LLM as Multimodal World Model
di: Ge, Zhiqi, et al.
Pubblicazione: (2024)
di: Ge, Zhiqi, et al.
Pubblicazione: (2024)
CrossView Suite: Harnessing Cross-view Spatial Intelligence of MLLMs with Dataset, Model and Benchmark
di: Wang, Wei, et al.
Pubblicazione: (2026)
di: Wang, Wei, et al.
Pubblicazione: (2026)
Meta-Reflection: A Feedback-Free Reflection Learning Framework
di: Wang, Yaoke, et al.
Pubblicazione: (2024)
di: Wang, Yaoke, et al.
Pubblicazione: (2024)
Iris: Breaking GUI Complexity with Adaptive Focus and Self-Refining
di: Ge, Zhiqi, et al.
Pubblicazione: (2024)
di: Ge, Zhiqi, et al.
Pubblicazione: (2024)
Chart-HQA: A Benchmark for Hypothetical Question Answering in Charts
di: Chen, Xiangnan, et al.
Pubblicazione: (2025)
di: Chen, Xiangnan, et al.
Pubblicazione: (2025)
Towards Meta-Cognitive Knowledge Editing for Multimodal LLMs
di: Fan, Zhaoyu, et al.
Pubblicazione: (2025)
di: Fan, Zhaoyu, et al.
Pubblicazione: (2025)
Text-to-Decision Agent: Offline Meta-Reinforcement Learning from Natural Language Supervision
di: Zhang, Shilin, et al.
Pubblicazione: (2025)
di: Zhang, Shilin, et al.
Pubblicazione: (2025)
Align$^2$LLaVA: Cascaded Human and Large Language Model Preference Alignment for Multi-modal Instruction Curation
di: Huang, Hongzhe, et al.
Pubblicazione: (2024)
di: Huang, Hongzhe, et al.
Pubblicazione: (2024)
Semantic Encryption: Secure and Effective Interaction with Cloud-based Large Language Models via Semantic Transformation
di: Chen, Dong, et al.
Pubblicazione: (2025)
di: Chen, Dong, et al.
Pubblicazione: (2025)
An Empirical Study of Knowledge Distillation for Code Understanding Tasks
di: Wang, Ruiqi, et al.
Pubblicazione: (2025)
di: Wang, Ruiqi, et al.
Pubblicazione: (2025)
MoA: Heterogeneous Mixture of Adapters for Parameter-Efficient Fine-Tuning of Large Language Models
di: Cao, Jie, et al.
Pubblicazione: (2025)
di: Cao, Jie, et al.
Pubblicazione: (2025)
TaskWeaver: A Code-First Agent Framework
di: Qiao, Bo, et al.
Pubblicazione: (2023)
di: Qiao, Bo, et al.
Pubblicazione: (2023)
TaskBench: Benchmarking Large Language Models for Task Automation
di: Shen, Yongliang, et al.
Pubblicazione: (2023)
di: Shen, Yongliang, et al.
Pubblicazione: (2023)
MathFimer: Enhancing Mathematical Reasoning by Expanding Reasoning Steps through Fill-in-the-Middle Task
di: Yan, Yuchen, et al.
Pubblicazione: (2025)
di: Yan, Yuchen, et al.
Pubblicazione: (2025)
Functional Matching of Logic Subgraphs: Beyond Structural Isomorphism
di: Zheng, Ziyang, et al.
Pubblicazione: (2025)
di: Zheng, Ziyang, et al.
Pubblicazione: (2025)
FAME: Adaptive Functional Attention with Expert Routing for Function-on-Function Regression
di: Gao, Yifei, et al.
Pubblicazione: (2025)
di: Gao, Yifei, et al.
Pubblicazione: (2025)
HalluciDoctor: Mitigating Hallucinatory Toxicity in Visual Instruction Data
di: Yu, Qifan, et al.
Pubblicazione: (2023)
di: Yu, Qifan, et al.
Pubblicazione: (2023)
Distilling Vision-Language Foundation Models: A Data-Free Approach via Prompt Diversification
di: Xuan, Yunyi, et al.
Pubblicazione: (2024)
di: Xuan, Yunyi, et al.
Pubblicazione: (2024)
UI-S1: Advancing GUI Automation via Semi-online Reinforcement Learning
di: Lu, Zhengxi, et al.
Pubblicazione: (2025)
di: Lu, Zhengxi, et al.
Pubblicazione: (2025)
EvoCurr: Self-evolving Curriculum with Behavior Code Generation for Complex Decision-making
di: Cheng, Yang, et al.
Pubblicazione: (2025)
di: Cheng, Yang, et al.
Pubblicazione: (2025)
TeamLoRA: Boosting Low-Rank Adaptation with Expert Collaboration and Competition
di: Lin, Tianwei, et al.
Pubblicazione: (2024)
di: Lin, Tianwei, et al.
Pubblicazione: (2024)
Relation Modeling and Distillation for Learning with Noisy Labels
di: Che, Xiaming, et al.
Pubblicazione: (2024)
di: Che, Xiaming, et al.
Pubblicazione: (2024)
Online Reinforcement Learning-Based Dynamic Adaptive Evaluation Function for Real-Time Strategy Tasks
di: Yang, Weilong, et al.
Pubblicazione: (2025)
di: Yang, Weilong, et al.
Pubblicazione: (2025)
Self-Distilled Agentic Reinforcement Learning
di: Lu, Zhengxi, et al.
Pubblicazione: (2026)
di: Lu, Zhengxi, et al.
Pubblicazione: (2026)
In-Context Curiosity: Distilling Exploration for Decision-Pretrained Transformers on Bandit Tasks
di: Yang, Huitao, et al.
Pubblicazione: (2025)
di: Yang, Huitao, et al.
Pubblicazione: (2025)
Decision-making with Speculative Opponent Models
di: Sun, Jing, et al.
Pubblicazione: (2022)
di: Sun, Jing, et al.
Pubblicazione: (2022)
VisualThink-VLA: Visual Intermediate Reasoning for Effective and Low-Latency Vision-Language-Action Policies
di: Gao, Mingjian, et al.
Pubblicazione: (2026)
di: Gao, Mingjian, et al.
Pubblicazione: (2026)
\textit{FocaLogic}: Logic-Based Interpretation of Visual Model Decisions
di: Zhao, Chenchen, et al.
Pubblicazione: (2026)
di: Zhao, Chenchen, et al.
Pubblicazione: (2026)
TCPO: Thought-Centric Preference Optimization for Effective Embodied Decision-making
di: Jiao, Kechen, et al.
Pubblicazione: (2025)
di: Jiao, Kechen, et al.
Pubblicazione: (2025)
Structural Estimation of Markov Decision Processes in High-Dimensional State Space with Finite-Time Guarantees
di: Zeng, Siliang, et al.
Pubblicazione: (2022)
di: Zeng, Siliang, et al.
Pubblicazione: (2022)
A Note on Knowledge Distillation Loss Function for Object Classification
di: Chen, Defang
Pubblicazione: (2021)
di: Chen, Defang
Pubblicazione: (2021)
GMFVAD: Using Grained Multi-modal Feature to Improve Video Anomaly Detection
di: Dai, Guangyu, et al.
Pubblicazione: (2025)
di: Dai, Guangyu, et al.
Pubblicazione: (2025)
From Easy to Hard: Learning Curricular Shape-aware Features for Robust Panoptic Scene Graph Generation
di: Shi, Hanrong, et al.
Pubblicazione: (2024)
di: Shi, Hanrong, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Improving Large Models with Small models: Lower Costs and Better Performance
di: Chen, Dong, et al.
Pubblicazione: (2024) -
KKA: Improving Vision Anomaly Detection through Anomaly-related Knowledge from Large Language Models
di: Chen, Dong, et al.
Pubblicazione: (2025) -
T2S-GPT: Dynamic Vector Quantization for Autoregressive Sign Language Production from Text
di: Yin, Aoxiong, et al.
Pubblicazione: (2024) -
IDEAL: Leveraging Infinite and Dynamic Characterizations of Large Language Models for Query-focused Summarization
di: Cao, Jie, et al.
Pubblicazione: (2024) -
Bridging Local Details and Global Context in Text-Attributed Graphs
di: Wang, Yaoke, et al.
Pubblicazione: (2024)