SCOPE: Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Ren, Tianfei, Yan, Zhipeng, Zhao, Yiming, Fang, Zhen, Zeng, Yu, Zhang, Guohui, Xu, Hang, Ma, Xiaoxiao, Huang, Shiting, Xu, Ke, Huang, Wenxuan, Wang, Lionel Z., Chen, Lin, Chen, Zehui, Huang, Jie, Zhao, Feng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SkillFlow:Benchmarking Lifelong Skill Discovery and Evolution for Autonomous Agents
by: Zhang, Ziao, et al.
Published: (2026)
by: Zhang, Ziao, et al.
Published: (2026)
ADORA: Training Reasoning Models with Dynamic Advantage Estimation on Reinforcement Learning
by: Ren, Qingnan, et al.
Published: (2026)
by: Ren, Qingnan, et al.
Published: (2026)
Beyond Accuracy: Unveiling Inefficiency Patterns in Tool-Integrated Reasoning
by: Su, Qisheng, et al.
Published: (2026)
by: Su, Qisheng, et al.
Published: (2026)
Agentic Jigsaw Interaction Learning for Enhancing Visual Perception and Reasoning in Vision-Language Models
by: Zeng, Yu, et al.
Published: (2025)
by: Zeng, Yu, et al.
Published: (2025)
ACC: Compiling Agent Trajectories for Long-Context Training
by: Su, Qisheng, et al.
Published: (2026)
by: Su, Qisheng, et al.
Published: (2026)
Internalizing Meta-Experience into Memory for Guided Reinforcement Learning in Large Language Models
by: Huang, Shiting, et al.
Published: (2026)
by: Huang, Shiting, et al.
Published: (2026)
CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scenarios
by: Huang, Shiting, et al.
Published: (2025)
by: Huang, Shiting, et al.
Published: (2025)
Vision-DeepResearch Benchmark: Rethinking Visual and Textual Search for Multimodal Large Language Models
by: Zeng, Yu, et al.
Published: (2026)
by: Zeng, Yu, et al.
Published: (2026)
Flow-OPD: On-Policy Distillation for Flow Matching Models
by: Fang, Zhen, et al.
Published: (2026)
by: Fang, Zhen, et al.
Published: (2026)
DualVLA: Building a Generalizable Embodied Agent via Partial Decoupling of Reasoning and Action
by: Fang, Zhen, et al.
Published: (2025)
by: Fang, Zhen, et al.
Published: (2025)
VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning
by: Qi, Yukun, et al.
Published: (2025)
by: Qi, Yukun, et al.
Published: (2025)
SaaSBench: Exploring the Boundaries of Coding Agents in Long-Horizon Enterprise SaaS Engineering
by: Ren, Qingnan, et al.
Published: (2026)
by: Ren, Qingnan, et al.
Published: (2026)
SCOPE: Sign Language Contextual Processing with Embedding from LLMs
by: Liu, Yuqi, et al.
Published: (2024)
by: Liu, Yuqi, et al.
Published: (2024)
AsyncTool: Evaluating the Asynchronous Function Calling Capability under Multi-Task Scenarios
by: Shi, Kou, et al.
Published: (2026)
by: Shi, Kou, et al.
Published: (2026)
Affordance Agent Harness: Verification-Gated Skill Orchestration
by: Huang, Haojian, et al.
Published: (2026)
by: Huang, Haojian, et al.
Published: (2026)
MaskFocus: Focusing Policy Optimization on Critical Steps for Masked Image Generation
by: Zhang, Guohui, et al.
Published: (2025)
by: Zhang, Guohui, et al.
Published: (2025)
VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation
by: Zhao, Yiming, et al.
Published: (2026)
by: Zhao, Yiming, et al.
Published: (2026)
MAR-GRPO: Stabilized GRPO for AR-diffusion Hybrid Image Generation
by: Ma, Xiaoxiao, et al.
Published: (2026)
by: Ma, Xiaoxiao, et al.
Published: (2026)
Skill-Conditioned Gated Self-Distillation for LLM Reasoning
by: Huang, Jiazhen, et al.
Published: (2026)
by: Huang, Jiazhen, et al.
Published: (2026)
Confident RAG: Enhancing the Performance of LLMs for Mathematics Question Answering through Multi-Embedding and Confidence Scoring
by: Chen, Shiting, et al.
Published: (2025)
by: Chen, Shiting, et al.
Published: (2025)
Diffusion Model for Manifold Data: Score Decomposition, Curvature, and Statistical Complexity
by: Zhang, Zixuan, et al.
Published: (2026)
by: Zhang, Zixuan, et al.
Published: (2026)
ModSkill: Physical Character Skill Modularization
by: Huang, Yiming, et al.
Published: (2025)
by: Huang, Yiming, et al.
Published: (2025)
Highly Efficient Test-Time Scaling for T2I Diffusion Models with Text Embedding Perturbation
by: Xu, Hang, et al.
Published: (2025)
by: Xu, Hang, et al.
Published: (2025)
FR-TTS: Test-Time Scaling for NTP-based Image Generation with Effective Filling-based Reward Signal
by: Xu, Hang, et al.
Published: (2025)
by: Xu, Hang, et al.
Published: (2025)
UniCorn: Towards Self-Improving Unified Multimodal Models through Self-Generated Supervision
by: Han, Ruiyan, et al.
Published: (2026)
by: Han, Ruiyan, et al.
Published: (2026)
Air‐Stable Hydrogen‐Substituted Graphdiyne/2D Halide Perovskite Heterojunction for Self‐Powered Neuromorphic Vision
by: Yiming Yuan, et al.
Published: (2025)
by: Yiming Yuan, et al.
Published: (2025)
UniCAIM: A Unified CAM/CIM Architecture with Static-Dynamic KV Cache Pruning for Efficient Long-Context LLM Inference
by: Xu, Weikai, et al.
Published: (2025)
by: Xu, Weikai, et al.
Published: (2025)
SCOPE: Performance Testing for Serverless Computing
by: Wen, Jinfeng, et al.
Published: (2023)
by: Wen, Jinfeng, et al.
Published: (2023)
SCOPE: Scene-Contextualized Incremental Few-Shot 3D Segmentation
by: Thengane, Vishal, et al.
Published: (2026)
by: Thengane, Vishal, et al.
Published: (2026)
SCOPE: Spectral Concentration by Distributionally Robust Joint Covariance-Precision Estimation
by: Chen, Renjie, et al.
Published: (2025)
by: Chen, Renjie, et al.
Published: (2025)
On a class of logarithmic Schrödinger equations via perturbation method
by: Huang, Chen, et al.
Published: (2026)
by: Huang, Chen, et al.
Published: (2026)
Organizing, Orchestrating, and Benchmarking Agent Skills at Ecosystem Scale
by: Li, Hao, et al.
Published: (2026)
by: Li, Hao, et al.
Published: (2026)
Chatting about Conditional Trajectory Prediction
by: Zhao, Yuxiang, et al.
Published: (2026)
by: Zhao, Yuxiang, et al.
Published: (2026)
TimeKAN: KAN-based Frequency Decomposition Learning Architecture for Long-term Time Series Forecasting
by: Huang, Songtao, et al.
Published: (2025)
by: Huang, Songtao, et al.
Published: (2025)
Fast-ARDiff: An Entropy-informed Acceleration Framework for Continuous Space Autoregressive Generation
by: Zou, Zhen, et al.
Published: (2025)
by: Zou, Zhen, et al.
Published: (2025)
Experimental study on the wave pressure of liquefied silty soil.
by: Huang, Zhe, et al.
Published: (2016)
by: Huang, Zhe, et al.
Published: (2016)
SI-Bench: Benchmarking Social Intelligence of Large Language Models in Human-to-Human Conversations
by: Huang, Shuai, et al.
Published: (2025)
by: Huang, Shuai, et al.
Published: (2025)
Reassessing the predictive power of bedfinder: insights into machine learning for subglacial bedform detection – Comments on ‘Automatic identification of streamlined subglacial bedforms using machine learning: an open‐source Python approach’
by: Ming Li, et al.
Published: (2025)
by: Ming Li, et al.
Published: (2025)
Drift-AR: Single-Step Visual Autoregressive Generation via Anti-Symmetric Drifting
by: Zou, Zhen, et al.
Published: (2026)
by: Zou, Zhen, et al.
Published: (2026)
Specific Emitter Identification Based on Joint Variational Mode Decomposition
by: Chen, Xiaofang, et al.
Published: (2024)
by: Chen, Xiaofang, et al.
Published: (2024)
Similar Items
-
SkillFlow:Benchmarking Lifelong Skill Discovery and Evolution for Autonomous Agents
by: Zhang, Ziao, et al.
Published: (2026) -
ADORA: Training Reasoning Models with Dynamic Advantage Estimation on Reinforcement Learning
by: Ren, Qingnan, et al.
Published: (2026) -
Beyond Accuracy: Unveiling Inefficiency Patterns in Tool-Integrated Reasoning
by: Su, Qisheng, et al.
Published: (2026) -
Agentic Jigsaw Interaction Learning for Enhancing Visual Perception and Reasoning in Vision-Language Models
by: Zeng, Yu, et al.
Published: (2025) -
ACC: Compiling Agent Trajectories for Long-Context Training
by: Su, Qisheng, et al.
Published: (2026)