Policy-based Foveated Imaging and Perception
Fuente:
arXiv
Salvato in:
| Autori principali: | Xiao, Howard, Ackermann, Jan, Deng, Boyang, Wetzstein, Gordon |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Foveated Diffusion: Efficient Spatially Adaptive Image and Video Generation
di: Chao, Brian, et al.
Pubblicazione: (2026)
di: Chao, Brian, et al.
Pubblicazione: (2026)
GeoFlow: Enforcing Implicit Geometric Consistency in Video Generation
di: Ackermann, Jan, et al.
Pubblicazione: (2026)
di: Ackermann, Jan, et al.
Pubblicazione: (2026)
Spectral Progressive Diffusion for Efficient Image and Video Generation
di: Xiao, Howard, et al.
Pubblicazione: (2026)
di: Xiao, Howard, et al.
Pubblicazione: (2026)
Asymmetric Flow Models
di: Chen, Hansheng, et al.
Pubblicazione: (2026)
di: Chen, Hansheng, et al.
Pubblicazione: (2026)
Towards Vision-Language-Garment Models for Web Knowledge Garment Understanding and Generation
di: Ackermann, Jan, et al.
Pubblicazione: (2025)
di: Ackermann, Jan, et al.
Pubblicazione: (2025)
Streetscapes: Large-scale Consistent Street View Generation Using Autoregressive Video Diffusion
di: Deng, Boyang, et al.
Pubblicazione: (2024)
di: Deng, Boyang, et al.
Pubblicazione: (2024)
Visual Chronicles: Using Multimodal LLMs to Analyze Massive Collections of Images
di: Deng, Boyang, et al.
Pubblicazione: (2025)
di: Deng, Boyang, et al.
Pubblicazione: (2025)
Image2Garment: Simulation-ready Garment Generation from a Single Image
di: Can, Selim Emir, et al.
Pubblicazione: (2026)
di: Can, Selim Emir, et al.
Pubblicazione: (2026)
CL-Splats: Continual Learning of Gaussian Splatting with Local Optimization
di: Ackermann, Jan, et al.
Pubblicazione: (2025)
di: Ackermann, Jan, et al.
Pubblicazione: (2025)
Towards Understanding Depth Perception in Foveated Rendering
di: Kergaßner, Sophie, et al.
Pubblicazione: (2025)
di: Kergaßner, Sophie, et al.
Pubblicazione: (2025)
BulletTime: Decoupled Control of Time and Camera Pose for Video Generation
di: Wang, Yiming, et al.
Pubblicazione: (2025)
di: Wang, Yiming, et al.
Pubblicazione: (2025)
Foveated Instance Segmentation
di: Zeng, Hongyi, et al.
Pubblicazione: (2025)
di: Zeng, Hongyi, et al.
Pubblicazione: (2025)
AIpparel: A Multimodal Foundation Model for Digital Garments
di: Nakayama, Kiyohiro, et al.
Pubblicazione: (2024)
di: Nakayama, Kiyohiro, et al.
Pubblicazione: (2024)
Foveated Reasoning: Stateful, Action-based Visual Focusing for Vision-Language Models
di: Min, Juhong, et al.
Pubblicazione: (2026)
di: Min, Juhong, et al.
Pubblicazione: (2026)
Robust Symmetry Detection via Riemannian Langevin Dynamics
di: Je, Jihyeon, et al.
Pubblicazione: (2024)
di: Je, Jihyeon, et al.
Pubblicazione: (2024)
Stable Diffusion for Data Augmentation in COCO and Weed Datasets
di: Deng, Boyang
Pubblicazione: (2023)
di: Deng, Boyang
Pubblicazione: (2023)
Reading in the Dark with Foveated Event Vision
di: Brander, Carl, et al.
Pubblicazione: (2025)
di: Brander, Carl, et al.
Pubblicazione: (2025)
Orthogonal Adaptation for Modular Customization of Diffusion Models
di: Po, Ryan, et al.
Pubblicazione: (2023)
di: Po, Ryan, et al.
Pubblicazione: (2023)
Towards Two-Stream Foveation-based Active Vision Learning
di: Ibrayev, Timur, et al.
Pubblicazione: (2024)
di: Ibrayev, Timur, et al.
Pubblicazione: (2024)
ScanGAN360: A Generative Model of Realistic Scanpaths for 360$^{\circ}$ Images
di: Martin, Daniel, et al.
Pubblicazione: (2021)
di: Martin, Daniel, et al.
Pubblicazione: (2021)
Geometric Algebra Planes: Convex Implicit Neural Volumes
di: Sivgin, Irmak, et al.
Pubblicazione: (2024)
di: Sivgin, Irmak, et al.
Pubblicazione: (2024)
GazeProphet: Software-Only Gaze Prediction for VR Foveated Rendering
di: Ebadulla, Farhaan, et al.
Pubblicazione: (2025)
di: Ebadulla, Farhaan, et al.
Pubblicazione: (2025)
Predicting Reaction Time to Comprehend Scenes with Foveated Scene Understanding Maps
di: Wen, Ziqi, et al.
Pubblicazione: (2025)
di: Wen, Ziqi, et al.
Pubblicazione: (2025)
Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models
di: Zhang, Lvmin, et al.
Pubblicazione: (2025)
di: Zhang, Lvmin, et al.
Pubblicazione: (2025)
Hybrid Foveated Path Tracing with Peripheral Gaussians for Immersive Anatomy
di: Kleinbeck, Constantin, et al.
Pubblicazione: (2026)
di: Kleinbeck, Constantin, et al.
Pubblicazione: (2026)
Visual Acuity Consistent Foveated Rendering towards Retinal Resolution
di: Zhang, Zhi, et al.
Pubblicazione: (2025)
di: Zhang, Zhi, et al.
Pubblicazione: (2025)
Semi-Supervised Weed Detection in Vegetable Fields: In-domain and Cross-domain Experiments
di: Deng, Boyang, et al.
Pubblicazione: (2025)
di: Deng, Boyang, et al.
Pubblicazione: (2025)
Neural Ganglion Sensors: Learning Task-specific Event Cameras Inspired by the Neural Circuit of the Human Retina
di: So, Haley M., et al.
Pubblicazione: (2025)
di: So, Haley M., et al.
Pubblicazione: (2025)
GazeFusion: Saliency-Guided Image Generation
di: Zhang, Yunxiang, et al.
Pubblicazione: (2024)
di: Zhang, Yunxiang, et al.
Pubblicazione: (2024)
GaussFusion: Improving 3D Reconstruction in the Wild with A Geometry-Informed Video Generator
di: Zhu, Liyuan, et al.
Pubblicazione: (2026)
di: Zhu, Liyuan, et al.
Pubblicazione: (2026)
ThermalNeRF: Thermal Radiance Fields
di: Lin, Yvette Y., et al.
Pubblicazione: (2024)
di: Lin, Yvette Y., et al.
Pubblicazione: (2024)
Self-Calibrating Gaussian Splatting for Large Field of View Reconstruction
di: Deng, Youming, et al.
Pubblicazione: (2025)
di: Deng, Youming, et al.
Pubblicazione: (2025)
Caption-Driven Explorations: Aligning Image and Text Embeddings through Human-Inspired Foveated Vision
di: Zanca, Dario, et al.
Pubblicazione: (2024)
di: Zanca, Dario, et al.
Pubblicazione: (2024)
Taming Flow-based I2V Models for Creative Video Editing
di: Kong, Xianghao, et al.
Pubblicazione: (2025)
di: Kong, Xianghao, et al.
Pubblicazione: (2025)
FiVA: Fine-grained Visual Attribute Dataset for Text-to-Image Diffusion Models
di: Wu, Tong, et al.
Pubblicazione: (2024)
di: Wu, Tong, et al.
Pubblicazione: (2024)
GenDoP: Auto-regressive Camera Trajectory Generation as a Director of Photography
di: Zhang, Mengchen, et al.
Pubblicazione: (2025)
di: Zhang, Mengchen, et al.
Pubblicazione: (2025)
pi-Flow: Policy-Based Few-Step Generation via Imitation Distillation
di: Chen, Hansheng, et al.
Pubblicazione: (2025)
di: Chen, Hansheng, et al.
Pubblicazione: (2025)
Foveated Retinotopy Improves Classification and Localization in Convolutional Neural Networks
di: Jérémie, Jean-Nicolas, et al.
Pubblicazione: (2024)
di: Jérémie, Jean-Nicolas, et al.
Pubblicazione: (2024)
BAgger: Backwards Aggregation for Mitigating Drift in Autoregressive Video Diffusion Models
di: Po, Ryan, et al.
Pubblicazione: (2025)
di: Po, Ryan, et al.
Pubblicazione: (2025)
Generated Reality: Human-centric World Simulation using Interactive Video Generation with Hand and Camera Control
di: Xie, Linxi, et al.
Pubblicazione: (2026)
di: Xie, Linxi, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Foveated Diffusion: Efficient Spatially Adaptive Image and Video Generation
di: Chao, Brian, et al.
Pubblicazione: (2026) -
GeoFlow: Enforcing Implicit Geometric Consistency in Video Generation
di: Ackermann, Jan, et al.
Pubblicazione: (2026) -
Spectral Progressive Diffusion for Efficient Image and Video Generation
di: Xiao, Howard, et al.
Pubblicazione: (2026) -
Asymmetric Flow Models
di: Chen, Hansheng, et al.
Pubblicazione: (2026) -
Towards Vision-Language-Garment Models for Web Knowledge Garment Understanding and Generation
di: Ackermann, Jan, et al.
Pubblicazione: (2025)