Verifier Threshold: An Efficient Test-Time Scaling Approach for Image Generation
Fuente:
arXiv
Salvato in:
| Autori principali: | Sundaresha, Vignesh, Haridas, Akash, Appia, Vikram, Varshney, Lav R. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Designing Parameter and Compute Efficient Diffusion Transformers using Distillation
di: Sundaresha, Vignesh
Pubblicazione: (2025)
di: Sundaresha, Vignesh
Pubblicazione: (2025)
Growing Efficient Accurate and Robust Neural Networks on the Edge
di: Sundaresha, Vignesh, et al.
Pubblicazione: (2024)
di: Sundaresha, Vignesh, et al.
Pubblicazione: (2024)
DC-DiT: Adaptive Compute and Elastic Inference for Visual Generation via Dynamic Chunking
di: Haridas, Akash, et al.
Pubblicazione: (2026)
di: Haridas, Akash, et al.
Pubblicazione: (2026)
High-Fidelity Human Avatars from Laptop Webcams using Edge Compute
di: Haridas, Akash, et al.
Pubblicazione: (2025)
di: Haridas, Akash, et al.
Pubblicazione: (2025)
CASCADE: Context-Aware Relaxation for Speculative Image Decoding
di: Yildirim, Selin, et al.
Pubblicazione: (2026)
di: Yildirim, Selin, et al.
Pubblicazione: (2026)
Semantically Grounded QFormer for Efficient Vision Language Understanding
di: Choraria, Moulik, et al.
Pubblicazione: (2023)
di: Choraria, Moulik, et al.
Pubblicazione: (2023)
dMLLM-TTS: Self-Verified and Efficient Test-Time Scaling for Diffusion Multi-Modal Large Language Models
di: Xin, Yi, et al.
Pubblicazione: (2025)
di: Xin, Yi, et al.
Pubblicazione: (2025)
Learning from One and Only One Shot
di: Yu, Haizi, et al.
Pubblicazione: (2022)
di: Yu, Haizi, et al.
Pubblicazione: (2022)
DeepInsert: Early Layer Bypass for Efficient and Performant Multimodal Understanding
di: Choraria, Moulik, et al.
Pubblicazione: (2025)
di: Choraria, Moulik, et al.
Pubblicazione: (2025)
Progress by Pieces: Test-Time Scaling for Autoregressive Image Generation
di: Park, Joonhyung, et al.
Pubblicazione: (2025)
di: Park, Joonhyung, et al.
Pubblicazione: (2025)
Reflect-DiT: Inference-Time Scaling for Text-to-Image Diffusion Transformers via In-Context Reflection
di: Li, Shufan, et al.
Pubblicazione: (2025)
di: Li, Shufan, et al.
Pubblicazione: (2025)
Rethinking Test Time Scaling for Flow-Matching Generative Models
di: Yu, Qingtao, et al.
Pubblicazione: (2025)
di: Yu, Qingtao, et al.
Pubblicazione: (2025)
Stream-T1: Test-Time Scaling for Streaming Video Generation
di: Tu, Yijing, et al.
Pubblicazione: (2026)
di: Tu, Yijing, et al.
Pubblicazione: (2026)
Test-Time Modality Generalization for Medical Image Segmentation
di: Nam, Ju-Hyeon, et al.
Pubblicazione: (2025)
di: Nam, Ju-Hyeon, et al.
Pubblicazione: (2025)
TreeCUA: Efficiently Scaling GUI Automation with Tree-Structured Verifiable Evolution
di: Jiang, Deyang, et al.
Pubblicazione: (2026)
di: Jiang, Deyang, et al.
Pubblicazione: (2026)
EdgeDiT: Hardware-Aware Diffusion Transformers for Efficient On-Device Image Generation
di: Kodavanti, Sravanth, et al.
Pubblicazione: (2026)
di: Kodavanti, Sravanth, et al.
Pubblicazione: (2026)
Context Diffusion: In-Context Aware Image Generation
di: Najdenkoska, Ivona, et al.
Pubblicazione: (2023)
di: Najdenkoska, Ivona, et al.
Pubblicazione: (2023)
A Large Scale Benchmark for Test Time Adaptation Methods in Medical Image Segmentation
di: Yu, Wenjing, et al.
Pubblicazione: (2025)
di: Yu, Wenjing, et al.
Pubblicazione: (2025)
S$^3$-TTA: Scale-Style Selection for Test-Time Augmentation in Biomedical Image Segmentation
di: Xie, Kangxian, et al.
Pubblicazione: (2023)
di: Xie, Kangxian, et al.
Pubblicazione: (2023)
Tiny Inference-Time Scaling with Latent Verifiers
di: Bucciarelli, Davide, et al.
Pubblicazione: (2026)
di: Bucciarelli, Davide, et al.
Pubblicazione: (2026)
Skip-It? Theoretical Conditions for Layer Skipping in Vision-Language Models
di: Hartman, Max, et al.
Pubblicazione: (2025)
di: Hartman, Max, et al.
Pubblicazione: (2025)
Efficient Test-Time Scaling for Small Vision-Language Models
di: Kaya, Mehmet Onurcan, et al.
Pubblicazione: (2025)
di: Kaya, Mehmet Onurcan, et al.
Pubblicazione: (2025)
Efficient Open Set Single Image Test Time Adaptation of Vision Language Models
di: Sreenivas, Manogna, et al.
Pubblicazione: (2024)
di: Sreenivas, Manogna, et al.
Pubblicazione: (2024)
TTS-VAR: A Test-Time Scaling Framework for Visual Auto-Regressive Generation
di: Chen, Zhekai, et al.
Pubblicazione: (2025)
di: Chen, Zhekai, et al.
Pubblicazione: (2025)
Video-T1: Test-Time Scaling for Video Generation
di: Liu, Fangfu, et al.
Pubblicazione: (2025)
di: Liu, Fangfu, et al.
Pubblicazione: (2025)
No Concept Left Behind: Test-Time Optimization for Compositional Text-to-Image Generation
di: Sameti, Mohammad Hossein, et al.
Pubblicazione: (2025)
di: Sameti, Mohammad Hossein, et al.
Pubblicazione: (2025)
Test-Time Dynamic Image Fusion
di: Cao, Bing, et al.
Pubblicazione: (2024)
di: Cao, Bing, et al.
Pubblicazione: (2024)
OrienText: Surface Oriented Textual Image Generation
di: Paliwal, Shubham Singh, et al.
Pubblicazione: (2025)
di: Paliwal, Shubham Singh, et al.
Pubblicazione: (2025)
Scaling Image and Video Generation via Test-Time Evolutionary Search
di: He, Haoran, et al.
Pubblicazione: (2025)
di: He, Haoran, et al.
Pubblicazione: (2025)
Spectral Evolution Search: Efficient Inference-Time Scaling for Reward-Aligned Image Generation
di: Ye, Jinyan, et al.
Pubblicazione: (2026)
di: Ye, Jinyan, et al.
Pubblicazione: (2026)
Rethinking Skip Connections: Additive U-Net for Robust and Interpretable Denoising
di: Lakkavalli, Vikram R
Pubblicazione: (2026)
di: Lakkavalli, Vikram R
Pubblicazione: (2026)
Edge-Efficient Image Restoration: Transformer Distillation into State-Space Models
di: Miriyala, Srinivas Soumitri, et al.
Pubblicazione: (2026)
di: Miriyala, Srinivas Soumitri, et al.
Pubblicazione: (2026)
CHATS: Combining Human-Aligned Optimization and Test-Time Sampling for Text-to-Image Generation
di: Fu, Minghao, et al.
Pubblicazione: (2025)
di: Fu, Minghao, et al.
Pubblicazione: (2025)
Tuning Real-World Image Restoration at Inference: A Test-Time Scaling Paradigm for Flow Matching Models
di: Bai, Purui, et al.
Pubblicazione: (2026)
di: Bai, Purui, et al.
Pubblicazione: (2026)
Test-Time Preference Optimization for Image Restoration
di: Li, Bingchen, et al.
Pubblicazione: (2025)
di: Li, Bingchen, et al.
Pubblicazione: (2025)
Single Image Test-Time Adaptation for Segmentation
di: Janouskova, Klara, et al.
Pubblicazione: (2023)
di: Janouskova, Klara, et al.
Pubblicazione: (2023)
Gameplay Highlights Generation
di: Edithal, Vignesh, et al.
Pubblicazione: (2025)
di: Edithal, Vignesh, et al.
Pubblicazione: (2025)
Can Test-Time Scaling Improve World Foundation Model?
di: Cong, Wenyan, et al.
Pubblicazione: (2025)
di: Cong, Wenyan, et al.
Pubblicazione: (2025)
Rethinking Dense Optical Flow without Test-Time Scaling
di: Chanda, Praroop, et al.
Pubblicazione: (2026)
di: Chanda, Praroop, et al.
Pubblicazione: (2026)
Guided Trajectory Optimization with Sparse Scaling for Test-Time Diffusion
di: Dai, Gang, et al.
Pubblicazione: (2026)
di: Dai, Gang, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Designing Parameter and Compute Efficient Diffusion Transformers using Distillation
di: Sundaresha, Vignesh
Pubblicazione: (2025) -
Growing Efficient Accurate and Robust Neural Networks on the Edge
di: Sundaresha, Vignesh, et al.
Pubblicazione: (2024) -
DC-DiT: Adaptive Compute and Elastic Inference for Visual Generation via Dynamic Chunking
di: Haridas, Akash, et al.
Pubblicazione: (2026) -
High-Fidelity Human Avatars from Laptop Webcams using Edge Compute
di: Haridas, Akash, et al.
Pubblicazione: (2025) -
CASCADE: Context-Aware Relaxation for Speculative Image Decoding
di: Yildirim, Selin, et al.
Pubblicazione: (2026)