VideoAgent: Self-Improving Video Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Soni, Achint, Venkataraman, Sreyas, Chandra, Abhranil, Fischmeister, Sebastian, Liang, Percy, Dai, Bo, Yang, Sherry |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DiffClone: Enhanced Behaviour Cloning in Robotics with Diffusion-Driven Policy Learning
by: Mani, Sabariswaran, et al.
Published: (2024)
by: Mani, Sabariswaran, et al.
Published: (2024)
Shape of Thought: When Distribution Matters More than Correctness in Reasoning Tasks
by: Chandra, Abhranil, et al.
Published: (2025)
by: Chandra, Abhranil, et al.
Published: (2025)
Reinforcement Learning for Machine Learning Engineering Agents
by: Yang, Sherry, et al.
Published: (2025)
by: Yang, Sherry, et al.
Published: (2025)
Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback
by: Furuta, Hiroki, et al.
Published: (2024)
by: Furuta, Hiroki, et al.
Published: (2024)
VideoAgent: Personalized Synthesis of Scientific Videos
by: Liang, Xiao, et al.
Published: (2025)
by: Liang, Xiao, et al.
Published: (2025)
MLE-Smith: Scaling MLE Tasks with Automated Multi-Agent Pipeline
by: Qiang, Rushi, et al.
Published: (2025)
by: Qiang, Rushi, et al.
Published: (2025)
Real-World Offline Reinforcement Learning from Vision Language Model Feedback
by: Venkataraman, Sreyas, et al.
Published: (2024)
by: Venkataraman, Sreyas, et al.
Published: (2024)
Evaluating Self-Supervised Learning via Risk Decomposition
by: Dubois, Yann, et al.
Published: (2023)
by: Dubois, Yann, et al.
Published: (2023)
VideoAgent: Long-form Video Understanding with Large Language Model as Agent
by: Wang, Xiaohan, et al.
Published: (2024)
by: Wang, Xiaohan, et al.
Published: (2024)
MLAgentBench: Evaluating Language Agents on Machine Learning Experimentation
by: Huang, Qian, et al.
Published: (2023)
by: Huang, Qian, et al.
Published: (2023)
LOCATEdit: Graph Laplacian Optimized Cross Attention for Localized Text-Guided Image Editing
by: Soni, Achint, et al.
Published: (2025)
by: Soni, Achint, et al.
Published: (2025)
Improving Video Generation with Human Feedback
by: Liu, Jie, et al.
Published: (2025)
by: Liu, Jie, et al.
Published: (2025)
Compressed Video Aggregator: Content-driven Module for Efficient Micro-Video Recommendation
by: Xiao, Yang, et al.
Published: (2026)
by: Xiao, Yang, et al.
Published: (2026)
Self-Improving AI Agents through Self-Play
by: Chojecki, Przemyslaw
Published: (2025)
by: Chojecki, Przemyslaw
Published: (2025)
On The Statistical Limits of Self-Improving Agents
by: Wang, Charles L., et al.
Published: (2025)
by: Wang, Charles L., et al.
Published: (2025)
Value-Incentivized Preference Optimization: A Unified Approach to Online and Offline RLHF
by: Cen, Shicong, et al.
Published: (2024)
by: Cen, Shicong, et al.
Published: (2024)
VideoAgentTrek: Computer Use Pretraining from Unlabeled Videos
by: Lu, Dunjie, et al.
Published: (2025)
by: Lu, Dunjie, et al.
Published: (2025)
VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation
by: He, Xuan, et al.
Published: (2024)
by: He, Xuan, et al.
Published: (2024)
KV Cache Quantization for Self-Forcing Video Generation: A 33-Method Empirical Study
by: Ranganath, Suraj, et al.
Published: (2026)
by: Ranganath, Suraj, et al.
Published: (2026)
AgentOCR: Reimagining Agent History via Optical Self-Compression
by: Feng, Lang, et al.
Published: (2026)
by: Feng, Lang, et al.
Published: (2026)
Experiential Reflective Learning for Self-Improving LLM Agents
by: Allard, Marc-Antoine, et al.
Published: (2026)
by: Allard, Marc-Antoine, et al.
Published: (2026)
Re-ENACT: Reinforcement Learning for Emotional Speech Generation using Actor-Critic Strategy
by: Shankar, Ravi, et al.
Published: (2024)
by: Shankar, Ravi, et al.
Published: (2024)
LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals
by: Park, Joon Sung, et al.
Published: (2024)
by: Park, Joon Sung, et al.
Published: (2024)
Enabling Self-Improving Agents to Learn at Test Time With Human-In-The-Loop Guidance
by: He, Yufei, et al.
Published: (2025)
by: He, Yufei, et al.
Published: (2025)
Learning Representations in Video Game Agents with Supervised Contrastive Imitation Learning
by: Celemin, Carlos, et al.
Published: (2025)
by: Celemin, Carlos, et al.
Published: (2025)
Continual Harness: Online Adaptation for Self-Improving Foundation Agents
by: Karten, Seth, et al.
Published: (2026)
by: Karten, Seth, et al.
Published: (2026)
Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation
by: Jeong, Hyeonho, et al.
Published: (2024)
by: Jeong, Hyeonho, et al.
Published: (2024)
Object-centric 3D Motion Field for Robot Learning from Human Videos
by: Yin, Zhao-Heng, et al.
Published: (2025)
by: Yin, Zhao-Heng, et al.
Published: (2025)
Fantastic Pretraining Optimizers and Where to Find Them
by: Wen, Kaiyue, et al.
Published: (2025)
by: Wen, Kaiyue, et al.
Published: (2025)
Combee: Scaling Prompt Learning for Self-Improving Language Model Agents
by: Li, Hanchen, et al.
Published: (2026)
by: Li, Hanchen, et al.
Published: (2026)
Laugh, Relate, Engage: Stylized Comment Generation for Short Videos
by: Ouyang, Xuan, et al.
Published: (2025)
by: Ouyang, Xuan, et al.
Published: (2025)
FrameBridge: Improving Image-to-Video Generation with Bridge Models
by: Wang, Yuji, et al.
Published: (2024)
by: Wang, Yuji, et al.
Published: (2024)
Contrastive and Variational Approaches in Self-Supervised Learning for Complex Data Mining
by: Liang, Yingbin, et al.
Published: (2025)
by: Liang, Yingbin, et al.
Published: (2025)
Personalized Student Knowledge Modeling for Future Learning Resource Prediction
by: Hashemifar, Soroush, et al.
Published: (2025)
by: Hashemifar, Soroush, et al.
Published: (2025)
Self-Improving LLM Agents at Test-Time
by: Acikgoz, Emre Can, et al.
Published: (2025)
by: Acikgoz, Emre Can, et al.
Published: (2025)
Video2Policy: Scaling up Manipulation Tasks in Simulation through Internet Videos
by: Ye, Weirui, et al.
Published: (2025)
by: Ye, Weirui, et al.
Published: (2025)
LongVideoAgent: Multi-Agent Reasoning with Long Videos
by: Liu, Runtao, et al.
Published: (2025)
by: Liu, Runtao, et al.
Published: (2025)
Neural Graph Matching for Video Retrieval in Large-Scale Video-driven E-commerce
by: Ji, Houye, et al.
Published: (2024)
by: Ji, Houye, et al.
Published: (2024)
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs
by: Bian, Song, et al.
Published: (2025)
by: Bian, Song, et al.
Published: (2025)
On the Entropy Calibration of Language Models
by: Cao, Steven, et al.
Published: (2025)
by: Cao, Steven, et al.
Published: (2025)
Similar Items
-
DiffClone: Enhanced Behaviour Cloning in Robotics with Diffusion-Driven Policy Learning
by: Mani, Sabariswaran, et al.
Published: (2024) -
Shape of Thought: When Distribution Matters More than Correctness in Reasoning Tasks
by: Chandra, Abhranil, et al.
Published: (2025) -
Reinforcement Learning for Machine Learning Engineering Agents
by: Yang, Sherry, et al.
Published: (2025) -
Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback
by: Furuta, Hiroki, et al.
Published: (2024) -
VideoAgent: Personalized Synthesis of Scientific Videos
by: Liang, Xiao, et al.
Published: (2025)