LossAgent: Towards Any Optimization Objectives for Image Processing with LLM Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Bingchen, Li, Xin, Lu, Yiting, Chen, Zhibo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Hybrid Agents for Image Restoration
von: Li, Bingchen, et al.
Veröffentlicht: (2025)
von: Li, Bingchen, et al.
Veröffentlicht: (2025)
Comp-X: On Defining an Interactive Learned Image Compression Paradigm With Expert-driven LLM Agent
von: Gao, Yixin, et al.
Veröffentlicht: (2025)
von: Gao, Yixin, et al.
Veröffentlicht: (2025)
PromptCIR: Blind Compressed Image Restoration with Prompt Learning
von: Li, Bingchen, et al.
Veröffentlicht: (2024)
von: Li, Bingchen, et al.
Veröffentlicht: (2024)
Q-Adapt: Adapting LMM for Visual Quality Assessment with Progressive Instruction Tuning
von: Lu, Yiting, et al.
Veröffentlicht: (2025)
von: Lu, Yiting, et al.
Veröffentlicht: (2025)
LiftVSR: Lifting Image Diffusion to Video Super-Resolution via Hybrid Temporal Modeling with Only 4$\times$RTX 4090s
von: Wang, Xijun, et al.
Veröffentlicht: (2025)
von: Wang, Xijun, et al.
Veröffentlicht: (2025)
IQA-Spider: Unifying Multi-Granularity Image Quality Assessment with Reasoning, Grounding and Referring
von: Peng, Xinge, et al.
Veröffentlicht: (2026)
von: Peng, Xinge, et al.
Veröffentlicht: (2026)
Test-Time Preference Optimization for Image Restoration
von: Li, Bingchen, et al.
Veröffentlicht: (2025)
von: Li, Bingchen, et al.
Veröffentlicht: (2025)
MambaCSR: Dual-Interleaved Scanning for Compressed Image Super-Resolution With SSMs
von: Ren, Yulin, et al.
Veröffentlicht: (2024)
von: Ren, Yulin, et al.
Veröffentlicht: (2024)
Is Vanilla MLP in Neural Radiance Field Enough for Few-shot View Synthesis?
von: Zhu, Hanxin, et al.
Veröffentlicht: (2024)
von: Zhu, Hanxin, et al.
Veröffentlicht: (2024)
UCIP: A Universal Framework for Compressed Image Super-Resolution using Dynamic Prompt
von: Li, Xin, et al.
Veröffentlicht: (2024)
von: Li, Xin, et al.
Veröffentlicht: (2024)
QMamba: On First Exploration of Vision Mamba for Image Quality Assessment
von: Guan, Fengbin, et al.
Veröffentlicht: (2024)
von: Guan, Fengbin, et al.
Veröffentlicht: (2024)
UniMIC: Towards Universal Multi-modality Perceptual Image Compression
von: Gao, Yixin, et al.
Veröffentlicht: (2024)
von: Gao, Yixin, et al.
Veröffentlicht: (2024)
SeD: Semantic-Aware Discriminator for Image Super-Resolution
von: Li, Bingchen, et al.
Veröffentlicht: (2024)
von: Li, Bingchen, et al.
Veröffentlicht: (2024)
Priorformer: A UGC-VQA Method with content and distortion priors
von: Pei, Yajing, et al.
Veröffentlicht: (2024)
von: Pei, Yajing, et al.
Veröffentlicht: (2024)
ColorFLUX: A Structure-Color Decoupling Framework for Old Photo Colorization
von: Li, Bingchen, et al.
Veröffentlicht: (2026)
von: Li, Bingchen, et al.
Veröffentlicht: (2026)
Spider: Any-to-Many Multimodal LLM
von: Lai, Jinxiang, et al.
Veröffentlicht: (2024)
von: Lai, Jinxiang, et al.
Veröffentlicht: (2024)
MoE-DiffIR: Task-customized Diffusion Priors for Universal Compressed Image Restoration
von: Ren, Yulin, et al.
Veröffentlicht: (2024)
von: Ren, Yulin, et al.
Veröffentlicht: (2024)
PaAgent: Portrait-Aware Image Restoration Agent via Subjective-Objective Reinforcement Learning
von: Wang, Yijian, et al.
Veröffentlicht: (2026)
von: Wang, Yijian, et al.
Veröffentlicht: (2026)
TIV-Diffusion: Towards Object-Centric Movement for Text-driven Image to Video Generation
von: Wang, Xingrui, et al.
Veröffentlicht: (2024)
von: Wang, Xingrui, et al.
Veröffentlicht: (2024)
GenArtist: Multimodal LLM as an Agent for Unified Image Generation and Editing
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
Towards Accurate One-Stage Object Detection with AP-Loss
von: Chen, Kean, et al.
Veröffentlicht: (2019)
von: Chen, Kean, et al.
Veröffentlicht: (2019)
AgentFoX: LLM Agent-Guided Fusion with eXplainability for AI-Generated Image Detection
von: Yu, Yangxin, et al.
Veröffentlicht: (2026)
von: Yu, Yangxin, et al.
Veröffentlicht: (2026)
Video Quality Assessment Based on Swin TransformerV2 and Coarse to Fine Strategy
von: Yu, Zihao, et al.
Veröffentlicht: (2024)
von: Yu, Zihao, et al.
Veröffentlicht: (2024)
InternVQA: Advancing Compressed Video Quality Assessment with Distilling Large Foundation Model
von: Guan, Fengbin, et al.
Veröffentlicht: (2025)
von: Guan, Fengbin, et al.
Veröffentlicht: (2025)
Why Compress What You Can Generate? When GPT-4o Generation Ushers in Image Compression Fields
von: Gao, Yixin, et al.
Veröffentlicht: (2025)
von: Gao, Yixin, et al.
Veröffentlicht: (2025)
Derain-Agent: A Plug-and-Play Agent Framework for Rainy Image Restoration
von: Yu, Zhaocheng, et al.
Veröffentlicht: (2026)
von: Yu, Zhaocheng, et al.
Veröffentlicht: (2026)
CoNo: Consistency Noise Injection for Tuning-free Long Video Diffusion
von: Wang, Xingrui, et al.
Veröffentlicht: (2024)
von: Wang, Xingrui, et al.
Veröffentlicht: (2024)
SpongeBob: Sync-Aware Harmonious Audio-Visual Generative Editing
von: Liang, Sen, et al.
Veröffentlicht: (2026)
von: Liang, Sen, et al.
Veröffentlicht: (2026)
What Limits Virtual Agent Application? OmniBench: A Scalable Multi-Dimensional Benchmark for Essential Virtual Agent Capabilities
von: Bu, Wendong, et al.
Veröffentlicht: (2025)
von: Bu, Wendong, et al.
Veröffentlicht: (2025)
AnyI2V: Animating Any Conditional Image with Motion Control
von: Li, Ziye, et al.
Veröffentlicht: (2025)
von: Li, Ziye, et al.
Veröffentlicht: (2025)
NTIRE 2025 Challenge on Short-form UGC Video Quality Assessment and Enhancement: KwaiSR Dataset and Study
von: Li, Xin, et al.
Veröffentlicht: (2025)
von: Li, Xin, et al.
Veröffentlicht: (2025)
Learning to Credit the Right Steps: Objective-aware Process Optimization for Visual Generation
von: Li, Rui, et al.
Veröffentlicht: (2026)
von: Li, Rui, et al.
Veröffentlicht: (2026)
AniMaker: Multi-Agent Animated Storytelling with MCTS-Driven Clip Generation
von: Shi, Haoyuan, et al.
Veröffentlicht: (2025)
von: Shi, Haoyuan, et al.
Veröffentlicht: (2025)
SAGE: Training Smart Any-Horizon Agents for Long Video Reasoning with Reinforcement Learning
von: Jain, Jitesh, et al.
Veröffentlicht: (2025)
von: Jain, Jitesh, et al.
Veröffentlicht: (2025)
CCA: Collaborative Competitive Agents for Image Editing
von: Hang, Tiankai, et al.
Veröffentlicht: (2024)
von: Hang, Tiankai, et al.
Veröffentlicht: (2024)
HiconAgent: History Context-aware Policy Optimization for GUI Agents
von: Zhou, Xurui, et al.
Veröffentlicht: (2025)
von: Zhou, Xurui, et al.
Veröffentlicht: (2025)
4DWorldBench: A Comprehensive Evaluation Framework for 3D/4D World Generation Models
von: Lu, Yiting, et al.
Veröffentlicht: (2025)
von: Lu, Yiting, et al.
Veröffentlicht: (2025)
Q&C: When Quantization Meets Cache in Efficient Image Generation
von: Ding, Xin, et al.
Veröffentlicht: (2025)
von: Ding, Xin, et al.
Veröffentlicht: (2025)
JarvisEvo: Towards a Self-Evolving Photo Editing Agent with Synergistic Editor-Evaluator Optimization
von: Lin, Yunlong, et al.
Veröffentlicht: (2025)
von: Lin, Yunlong, et al.
Veröffentlicht: (2025)
Interpretable Text-Guided Image Clustering via Iterative Search
von: Zhao, Bingchen, et al.
Veröffentlicht: (2025)
von: Zhao, Bingchen, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Hybrid Agents for Image Restoration
von: Li, Bingchen, et al.
Veröffentlicht: (2025) -
Comp-X: On Defining an Interactive Learned Image Compression Paradigm With Expert-driven LLM Agent
von: Gao, Yixin, et al.
Veröffentlicht: (2025) -
PromptCIR: Blind Compressed Image Restoration with Prompt Learning
von: Li, Bingchen, et al.
Veröffentlicht: (2024) -
Q-Adapt: Adapting LMM for Visual Quality Assessment with Progressive Instruction Tuning
von: Lu, Yiting, et al.
Veröffentlicht: (2025) -
LiftVSR: Lifting Image Diffusion to Video Super-Resolution via Hybrid Temporal Modeling with Only 4$\times$RTX 4090s
von: Wang, Xijun, et al.
Veröffentlicht: (2025)