Allegro: Open the Black Box of Commercial-Level Video Generation Model
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Yuan, Wang, Qiuyue, Cai, Yuxuan, Yang, Huan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Multi-Subject Consistency in Open-Domain Image Generation with Isolation and Reposition Attention
by: He, Huiguo, et al.
Published: (2024)
by: He, Huiguo, et al.
Published: (2024)
Long Video Diffusion Generation with Segmented Cross-Attention and Content-Rich Video Data Curation
by: Yan, Xin, et al.
Published: (2024)
by: Yan, Xin, et al.
Published: (2024)
Fleximo: Towards Flexible Text-to-Human Motion Video Generation
by: Zhang, Yuhang, et al.
Published: (2024)
by: Zhang, Yuhang, et al.
Published: (2024)
DiffExplainer: Unveiling Black Box Models Via Counterfactual Generation
by: Fang, Yingying, et al.
Published: (2024)
by: Fang, Yingying, et al.
Published: (2024)
Query-Efficient Hard-Label Black-Box Attack against Vision Transformers
by: Zhou, Chao, et al.
Published: (2024)
by: Zhou, Chao, et al.
Published: (2024)
Chameleon: Benchmarking Detection and Backtracking on Commercial-Grade AI-Generated Videos
by: Liao, Xingming, et al.
Published: (2025)
by: Liao, Xingming, et al.
Published: (2025)
Lumen: Consistent Video Relighting and Harmonious Background Replacement with Video Generative Models
by: Zeng, Jianshu, et al.
Published: (2025)
by: Zeng, Jianshu, et al.
Published: (2025)
Box-Level Class-Balanced Sampling for Active Object Detection
by: Liao, Jingyi, et al.
Published: (2025)
by: Liao, Jingyi, et al.
Published: (2025)
RNA-FM: Flow-Matching Generative Model for Genome-wide RNA-Seq Prediction
by: Song, Yaxuan, et al.
Published: (2026)
by: Song, Yaxuan, et al.
Published: (2026)
Open-Sora Plan: Open-Source Large Video Generation Model
by: Lin, Bin, et al.
Published: (2024)
by: Lin, Bin, et al.
Published: (2024)
Wan: Open and Advanced Large-Scale Video Generative Models
by: Wan, Team, et al.
Published: (2025)
by: Wan, Team, et al.
Published: (2025)
CoCoVideo: The High-Quality Commercial-Model-Based Contrastive Benchmark for AI-Generated Video Detection
by: Feng, Huidong, et al.
Published: (2026)
by: Feng, Huidong, et al.
Published: (2026)
DreamStory: Open-Domain Story Visualization by LLM-Guided Multi-Subject Consistent Diffusion
by: He, Huiguo, et al.
Published: (2024)
by: He, Huiguo, et al.
Published: (2024)
Open Sesame! Universal Black Box Jailbreaking of Large Language Models
by: Lapid, Raz, et al.
Published: (2023)
by: Lapid, Raz, et al.
Published: (2023)
OpenBox: Annotate Any Bounding Boxes in 3D
by: Lee, In-Jae, et al.
Published: (2025)
by: Lee, In-Jae, et al.
Published: (2025)
YingVideo-MV: Music-Driven Multi-Stage Video Generation
by: Chen, Jiahui, et al.
Published: (2025)
by: Chen, Jiahui, et al.
Published: (2025)
Frame-Level Captions for Long Video Generation with Complex Multi Scenes
by: Zheng, Guangcong, et al.
Published: (2025)
by: Zheng, Guangcong, et al.
Published: (2025)
Improving Black-Box Generative Attacks via Generator Semantic Consistency
by: Jeong, Jongoh, et al.
Published: (2025)
by: Jeong, Jongoh, et al.
Published: (2025)
Test-Time Hinting for Black-Box Vision-Language Models
by: Hou, Kaihua, et al.
Published: (2026)
by: Hou, Kaihua, et al.
Published: (2026)
E.T. Bench: Towards Open-Ended Event-Level Video-Language Understanding
by: Liu, Ye, et al.
Published: (2024)
by: Liu, Ye, et al.
Published: (2024)
Arbitrary Generative Video Interpolation
by: Zhang, Guozhen, et al.
Published: (2025)
by: Zhang, Guozhen, et al.
Published: (2025)
PM-VIS: High-Performance Box-Supervised Video Instance Segmentation
by: Yang, Zhangjing, et al.
Published: (2024)
by: Yang, Zhangjing, et al.
Published: (2024)
BizGenEval: A Systematic Benchmark for Commercial Visual Content Generation
by: Li, Yan, et al.
Published: (2026)
by: Li, Yan, et al.
Published: (2026)
Adapting in the Dark: Efficient and Stable Test-Time Adaptation for Black-Box Models
by: Zhang, Yunbei, et al.
Published: (2026)
by: Zhang, Yunbei, et al.
Published: (2026)
How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
by: Chen, Zhe, et al.
Published: (2024)
by: Chen, Zhe, et al.
Published: (2024)
Opening the Black Box: Preliminary Insights into Affective Modeling in Multimodal Foundation Models
by: Zhang, Zhen, et al.
Published: (2026)
by: Zhang, Zhen, et al.
Published: (2026)
Prime Once, then Reprogram Locally: An Efficient Alternative to Black-Box Service Model Adaptation
by: Zhang, Yunbei, et al.
Published: (2026)
by: Zhang, Yunbei, et al.
Published: (2026)
VideoGLUE: Video General Understanding Evaluation of Foundation Models
by: Yuan, Liangzhe, et al.
Published: (2023)
by: Yuan, Liangzhe, et al.
Published: (2023)
Box-Free Model Watermarks Are Prone to Black-Box Removal Attacks
by: An, Haonan, et al.
Published: (2024)
by: An, Haonan, et al.
Published: (2024)
VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models
by: Wang, Yuxuan, et al.
Published: (2024)
by: Wang, Yuxuan, et al.
Published: (2024)
Eyes Wide Open: Ego Proactive Video-LLM for Streaming Video
by: Zhang, Yulin, et al.
Published: (2025)
by: Zhang, Yulin, et al.
Published: (2025)
A Simple Low-bit Quantization Framework for Video Snapshot Compressive Imaging
by: Cao, Miao, et al.
Published: (2024)
by: Cao, Miao, et al.
Published: (2024)
Helios: Real Real-Time Long Video Generation Model
by: Yuan, Shenghai, et al.
Published: (2026)
by: Yuan, Shenghai, et al.
Published: (2026)
Hard-Label Black-Box Attacks on 3D Point Clouds
by: Liu, Daizong, et al.
Published: (2024)
by: Liu, Daizong, et al.
Published: (2024)
Federated Black-Box Adaptation for Semantic Segmentation
by: Paranjape, Jay N., et al.
Published: (2024)
by: Paranjape, Jay N., et al.
Published: (2024)
Black-Box Adversarial Attack on Vision Language Models for Autonomous Driving
by: Wang, Lu, et al.
Published: (2025)
by: Wang, Lu, et al.
Published: (2025)
Generalizing Deepfake Video Detection with Plug-and-Play: Video-Level Blending and Spatiotemporal Adapter Tuning
by: Yan, Zhiyuan, et al.
Published: (2024)
by: Yan, Zhiyuan, et al.
Published: (2024)
UniVA: Universal Video Agent towards Open-Source Next-Generation Video Generalist
by: Liang, Zhengyang, et al.
Published: (2025)
by: Liang, Zhengyang, et al.
Published: (2025)
OpenSubject: Leveraging Video-Derived Identity and Diversity Priors for Subject-driven Image Generation and Manipulation
by: Liu, Yexin, et al.
Published: (2025)
by: Liu, Yexin, et al.
Published: (2025)
Towards Real-time Video Compressive Sensing on Mobile Devices
by: Cao, Miao, et al.
Published: (2024)
by: Cao, Miao, et al.
Published: (2024)
Similar Items
-
Improving Multi-Subject Consistency in Open-Domain Image Generation with Isolation and Reposition Attention
by: He, Huiguo, et al.
Published: (2024) -
Long Video Diffusion Generation with Segmented Cross-Attention and Content-Rich Video Data Curation
by: Yan, Xin, et al.
Published: (2024) -
Fleximo: Towards Flexible Text-to-Human Motion Video Generation
by: Zhang, Yuhang, et al.
Published: (2024) -
DiffExplainer: Unveiling Black Box Models Via Counterfactual Generation
by: Fang, Yingying, et al.
Published: (2024) -
Query-Efficient Hard-Label Black-Box Attack against Vision Transformers
by: Zhou, Chao, et al.
Published: (2024)