ScalingNoise: Scaling Inference-Time Search for Generating Infinite Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Haolin, Tang, Feilong, Hu, Ming, Yin, Qingyu, Li, Yulong, Liu, Yexin, Peng, Zelin, Gao, Peng, He, Junjun, Ge, Zongyuan, Razzak, Imran |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Dynamic 3D Reconstruction of Hand-Instrument Interaction in Ophthalmic Surgery
by: Hu, Ming, et al.
Published: (2025)
by: Hu, Ming, et al.
Published: (2025)
MMRC: A Large-Scale Benchmark for Understanding Multimodal Large Language Model in Real-World Conversation
by: Xue, Haochen, et al.
Published: (2025)
by: Xue, Haochen, et al.
Published: (2025)
StreamAgent: Towards Anticipatory Agents for Streaming Video Understanding
by: Yang, Haolin, et al.
Published: (2025)
by: Yang, Haolin, et al.
Published: (2025)
PG-SAM: Prior-Guided SAM with Medical for Multi-organ Segmentation
by: Zhong, Yiheng, et al.
Published: (2025)
by: Zhong, Yiheng, et al.
Published: (2025)
DeLo: Dual Decomposed Low-Rank Experts Collaboration for Continual Missing Modality Learning
by: Liu, Xiwei, et al.
Published: (2026)
by: Liu, Xiwei, et al.
Published: (2026)
Towards Efficient Medical Reasoning with Minimal Fine-Tuning Data
by: Zhuang, Xinlin, et al.
Published: (2025)
by: Zhuang, Xinlin, et al.
Published: (2025)
Robust Multimodal Learning for Ophthalmic Disease Grading via Disentangled Representation
by: Wang, Xinkun, et al.
Published: (2025)
by: Wang, Xinkun, et al.
Published: (2025)
Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding
by: Tang, Feilong, et al.
Published: (2025)
by: Tang, Feilong, et al.
Published: (2025)
TAGS: A Test-Time Generalist-Specialist Framework with Retrieval-Augmented Reasoning and Verification
by: Wu, Jianghao, et al.
Published: (2025)
by: Wu, Jianghao, et al.
Published: (2025)
Discriminating retinal microvascular and neuronal differences related to migraines: Deep Learning based Crossectional Study
by: Tang, Feilong, et al.
Published: (2024)
by: Tang, Feilong, et al.
Published: (2024)
MSWAL: 3D Multi-class Segmentation of Whole Abdominal Lesions Dataset
by: Wu, Zhaodong, et al.
Published: (2025)
by: Wu, Zhaodong, et al.
Published: (2025)
Diffusion Model Driven Test-Time Image Adaptation for Robust Skin Lesion Classification
by: Hu, Ming, et al.
Published: (2024)
by: Hu, Ming, et al.
Published: (2024)
OphNet: A Large-Scale Video Benchmark for Ophthalmic Surgical Workflow Understanding
by: Hu, Ming, et al.
Published: (2024)
by: Hu, Ming, et al.
Published: (2024)
Ophora: A Large-Scale Data-Driven Text-Guided Ophthalmic Surgical Video Generation Model
by: Li, Wei, et al.
Published: (2025)
by: Li, Wei, et al.
Published: (2025)
Phenome-Wide Multi-Omics Integration Uncovers Distinct Archetypes of Human Aging
by: Li, Huifa, et al.
Published: (2025)
by: Li, Huifa, et al.
Published: (2025)
LATA: Laplacian-Assisted Transductive Adaptation for Conformal Uncertainty in Medical VLMs
by: Bozorgtabar, Behzad, et al.
Published: (2026)
by: Bozorgtabar, Behzad, et al.
Published: (2026)
Towards Robust Visual Continual Learning with Multi-Prototype Supervision
by: Liu, Xiwei, et al.
Published: (2025)
by: Liu, Xiwei, et al.
Published: (2025)
A Machine Learning Approach to Predict Biological Age and its Longitudinal Drivers
by: Dunbayeva, Nazira, et al.
Published: (2025)
by: Dunbayeva, Nazira, et al.
Published: (2025)
OphCLIP: Hierarchical Retrieval-Augmented Learning for Ophthalmic Surgical Video-Language Pretraining
by: Hu, Ming, et al.
Published: (2024)
by: Hu, Ming, et al.
Published: (2024)
Neighbor Does Matter: Density-Aware Contrastive Learning for Medical Semi-supervised Segmentation
by: Tang, Feilong, et al.
Published: (2024)
by: Tang, Feilong, et al.
Published: (2024)
Generalizing to Unseen Domains in Diabetic Retinopathy with Disentangled Representations
by: Xia, Peng, et al.
Published: (2024)
by: Xia, Peng, et al.
Published: (2024)
Dual-Stage Value-Guided Inference with Margin-Based Reward Adjustment for Fast and Faithful VLM Captioning
by: Deria, Ankan, et al.
Published: (2025)
by: Deria, Ankan, et al.
Published: (2025)
LatSearch: Latent Reward-Guided Search for Faster Inference-Time Scaling in Video Diffusion
by: Zhao, Zengqun, et al.
Published: (2026)
by: Zhao, Zengqun, et al.
Published: (2026)
Rhythm of Opinion: A Hawkes-Graph Framework for Dynamic Propagation Analysis
by: Li, Yulong, et al.
Published: (2025)
by: Li, Yulong, et al.
Published: (2025)
ClinCoT: Clinical-Aware Visual Chain-of-Thought for Medical Vision Language Models
by: Liu, Xiwei, et al.
Published: (2026)
by: Liu, Xiwei, et al.
Published: (2026)
Sample, Scrutinize and Scale: Effective Inference-Time Search by Scaling Verification
by: Zhao, Eric, et al.
Published: (2025)
by: Zhao, Eric, et al.
Published: (2025)
Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps
by: Ma, Nanye, et al.
Published: (2025)
by: Ma, Nanye, et al.
Published: (2025)
Inference-Time Scaling for Visual AutoRegressive modeling by Searching Representative Samples
by: Tang, Weidong, et al.
Published: (2026)
by: Tang, Weidong, et al.
Published: (2026)
Hunting Attributes: Context Prototype-Aware Learning for Weakly Supervised Semantic Segmentation
by: Tang, Feilong, et al.
Published: (2024)
by: Tang, Feilong, et al.
Published: (2024)
Inference-Time Alignment of Diffusion Models with Direct Noise Optimization
by: Tang, Zhiwei, et al.
Published: (2024)
by: Tang, Zhiwei, et al.
Published: (2024)
ETS: Efficient Tree Search for Inference-Time Scaling
by: Hooper, Coleman, et al.
Published: (2025)
by: Hooper, Coleman, et al.
Published: (2025)
SAM-aware Test-time Adaptation for Universal Medical Image Segmentation
by: Wu, Jianghao, et al.
Published: (2025)
by: Wu, Jianghao, et al.
Published: (2025)
SAM-DCE: Addressing Token Uniformity and Semantic Over-Smoothing in Medical Segmentation
by: Hu, Yingzhen, et al.
Published: (2025)
by: Hu, Yingzhen, et al.
Published: (2025)
CoME-VL: Scaling Complementary Multi-Encoder Vision-Language Learning
by: Deria, Ankan, et al.
Published: (2026)
by: Deria, Ankan, et al.
Published: (2026)
Inference-Time Scaling for Generalist Reward Modeling
by: Liu, Zijun, et al.
Published: (2025)
by: Liu, Zijun, et al.
Published: (2025)
Scaling Image and Video Generation via Test-Time Evolutionary Search
by: He, Haoran, et al.
Published: (2025)
by: He, Haoran, et al.
Published: (2025)
DermAgent: A Self-Reflective Agentic System for Dermatological Image Analysis with Multi-Tool Reasoning and Traceable Decision-Making
by: Liu, Yize, et al.
Published: (2026)
by: Liu, Yize, et al.
Published: (2026)
Test-Time Scaling of Diffusion Models via Noise Trajectory Search
by: Ramesh, Vignav, et al.
Published: (2025)
by: Ramesh, Vignav, et al.
Published: (2025)
LMPT: Prompt Tuning with Class-Specific Embedding Loss for Long-tailed Multi-Label Visual Recognition
by: Xia, Peng, et al.
Published: (2023)
by: Xia, Peng, et al.
Published: (2023)
Adaptive Inference-Time Scaling via Cyclic Diffusion Search
by: Lee, Gyubin, et al.
Published: (2025)
by: Lee, Gyubin, et al.
Published: (2025)
Similar Items
-
Towards Dynamic 3D Reconstruction of Hand-Instrument Interaction in Ophthalmic Surgery
by: Hu, Ming, et al.
Published: (2025) -
MMRC: A Large-Scale Benchmark for Understanding Multimodal Large Language Model in Real-World Conversation
by: Xue, Haochen, et al.
Published: (2025) -
StreamAgent: Towards Anticipatory Agents for Streaming Video Understanding
by: Yang, Haolin, et al.
Published: (2025) -
PG-SAM: Prior-Guided SAM with Medical for Multi-organ Segmentation
by: Zhong, Yiheng, et al.
Published: (2025) -
DeLo: Dual Decomposed Low-Rank Experts Collaboration for Continual Missing Modality Learning
by: Liu, Xiwei, et al.
Published: (2026)