NARAIM: Native Aspect Ratio Autoregressive Image Models
Fuente:
arXiv
Saved in:
| Main Authors: | Fernández, Daniel Gallo, van der Klis, Robert, Matişan, Răzvan-Andrei, Partyka, Janusz, Gavves, Efstratios, Papa, Samuele, Lippe, Phillip |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reproducibility Study of "ITI-GEN: Inclusive Text-to-Image Generation"
by: Fernández, Daniel Gallo, et al.
Published: (2024)
by: Fernández, Daniel Gallo, et al.
Published: (2024)
How to Train Neural Field Representations: A Comprehensive Study and Benchmark
by: Papa, Samuele, et al.
Published: (2023)
by: Papa, Samuele, et al.
Published: (2023)
DuoDiff: Accelerating Diffusion Models with a Dual-Backbone Approach
by: Fernández, Daniel Gallo, et al.
Published: (2024)
by: Fernández, Daniel Gallo, et al.
Published: (2024)
From MLP to NeoMLP: Leveraging Self-Attention for Neural Fields
by: Kofinas, Miltiadis, et al.
Published: (2024)
by: Kofinas, Miltiadis, et al.
Published: (2024)
Dream to Manipulate: Compositional World Models Empowering Robot Imitation Learning with Imagination
by: Barcellona, Leonardo, et al.
Published: (2024)
by: Barcellona, Leonardo, et al.
Published: (2024)
Language Agents Meet Causality -- Bridging LLMs and Causal World Models
by: Gkountouras, John, et al.
Published: (2024)
by: Gkountouras, John, et al.
Published: (2024)
Grounding Continuous Representations in Geometry: Equivariant Neural Fields
by: Wessels, David R, et al.
Published: (2024)
by: Wessels, David R, et al.
Published: (2024)
Any-Resolution AI-Generated Image Detection by Spectral Learning
by: Karageorgiou, Dimitrios, et al.
Published: (2024)
by: Karageorgiou, Dimitrios, et al.
Published: (2024)
Reproducibility study of "LICO: Explainable Models with Language-Image Consistency"
by: Fletcher, Luan, et al.
Published: (2024)
by: Fletcher, Luan, et al.
Published: (2024)
Amortized Equation Discovery in Hybrid Dynamical Systems
by: Liu, Yongtuo, et al.
Published: (2024)
by: Liu, Yongtuo, et al.
Published: (2024)
Dual Guidance Semi-Supervised Action Detection
by: Singh, Ankit, et al.
Published: (2025)
by: Singh, Ankit, et al.
Published: (2025)
LoTUS: Large-Scale Machine Unlearning with a Taste of Uncertainty
by: Spartalis, Christoforos N., et al.
Published: (2025)
by: Spartalis, Christoforos N., et al.
Published: (2025)
Unleashing Uncertainty: Efficient Machine Unlearning for Generative AI
by: Spartalis, Christoforos N., et al.
Published: (2025)
by: Spartalis, Christoforos N., et al.
Published: (2025)
Purrception: Variational Flow Matching for Vector-Quantized Image Generation
by: Matişan, Răzvan-Andrei, et al.
Published: (2025)
by: Matişan, Răzvan-Andrei, et al.
Published: (2025)
Mechanistic Interpretability for AI Safety -- A Review
by: Bereska, Leonard, et al.
Published: (2024)
by: Bereska, Leonard, et al.
Published: (2024)
Probabilistic Prototype Calibration of Vision-Language Models for Generalized Few-shot Semantic Segmentation
by: Liu, Jie, et al.
Published: (2025)
by: Liu, Jie, et al.
Published: (2025)
From Explainable to Explained AI: Ideas for Falsifying and Quantifying Explanations
by: Schirris, Yoni, et al.
Published: (2025)
by: Schirris, Yoni, et al.
Published: (2025)
Redefining Normal: A Novel Object-Level Approach for Multi-Object Novelty Detection
by: Salehi, Mohammadreza, et al.
Published: (2024)
by: Salehi, Mohammadreza, et al.
Published: (2024)
Dynamic Prototype Adaptation with Distillation for Few-shot Point Cloud Segmentation
by: Liu, Jie, et al.
Published: (2024)
by: Liu, Jie, et al.
Published: (2024)
UltraFlux: Data-Model Co-Design for High-quality Native 4K Text-to-Image Generation across Diverse Aspect Ratios
by: Ye, Tian, et al.
Published: (2025)
by: Ye, Tian, et al.
Published: (2025)
CTRL-O: Language-Controllable Object-Centric Visual Representation Learning
by: Didolkar, Aniket, et al.
Published: (2025)
by: Didolkar, Aniket, et al.
Published: (2025)
Mining Open Semantics from CLIP: A Relation Transition Perspective for Few-Shot Learning
by: Yan, Cilin, et al.
Published: (2024)
by: Yan, Cilin, et al.
Published: (2024)
PRADA: Probability-Ratio-Based Attribution and Detection of Autoregressive-Generated Images
by: Damm, Simon, et al.
Published: (2025)
by: Damm, Simon, et al.
Published: (2025)
Space-Time Continuous PDE Forecasting using Equivariant Neural Fields
by: Knigge, David M., et al.
Published: (2024)
by: Knigge, David M., et al.
Published: (2024)
GeneralAD: Anomaly Detection Across Domains by Attending to Distorted Features
by: Sträter, Luc P. J., et al.
Published: (2024)
by: Sträter, Luc P. J., et al.
Published: (2024)
Probabilistic Interactive 3D Segmentation with Hierarchical Neural Processes
by: Liu, Jie, et al.
Published: (2025)
by: Liu, Jie, et al.
Published: (2025)
MoSiC: Optimal-Transport Motion Trajectory for Dense Self-Supervised Learning
by: Salehi, Mohammadreza, et al.
Published: (2025)
by: Salehi, Mohammadreza, et al.
Published: (2025)
ContRail: A Framework for Realistic Railway Image Synthesis using ControlNet
by: Alexandrescu, Andrei-Robert, et al.
Published: (2024)
by: Alexandrescu, Andrei-Robert, et al.
Published: (2024)
Ai-Sampler: Adversarial Learning of Markov kernels with involutive maps
by: Egorov, Evgenii, et al.
Published: (2024)
by: Egorov, Evgenii, et al.
Published: (2024)
Mechanistic Neural Networks for Scientific Machine Learning
by: Pervez, Adeel, et al.
Published: (2024)
by: Pervez, Adeel, et al.
Published: (2024)
Mechanistic PDE Networks for Discovery of Governing Equations
by: Pervez, Adeel, et al.
Published: (2025)
by: Pervez, Adeel, et al.
Published: (2025)
VISA: Reasoning Video Object Segmentation via Large Language Models
by: Yan, Cilin, et al.
Published: (2024)
by: Yan, Cilin, et al.
Published: (2024)
SIGMA: Sinkhorn-Guided Masked Video Modeling
by: Salehi, Mohammadreza, et al.
Published: (2024)
by: Salehi, Mohammadreza, et al.
Published: (2024)
PiClick: Picking the desired mask from multiple candidates in click-based interactive segmentation
by: Yan, Cilin, et al.
Published: (2023)
by: Yan, Cilin, et al.
Published: (2023)
Equivariant Multiscale Learned Invertible Reconstruction for Cone Beam CT: From Simulated to Real Data
by: Moriakov, Nikita, et al.
Published: (2025)
by: Moriakov, Nikita, et al.
Published: (2025)
Self-supervised visual learning in the low-data regime: a comparative evaluation
by: Konstantakos, Sotirios, et al.
Published: (2024)
by: Konstantakos, Sotirios, et al.
Published: (2024)
QuadGPT: Native Quadrilateral Mesh Generation with Autoregressive Models
by: Liu, Jian, et al.
Published: (2025)
by: Liu, Jian, et al.
Published: (2025)
ANOLE: An Open, Autoregressive, Native Large Multimodal Models for Interleaved Image-Text Generation
by: Chern, Ethan, et al.
Published: (2024)
by: Chern, Ethan, et al.
Published: (2024)
LLaVA-UHD: an LMM Perceiving Any Aspect Ratio and High-Resolution Images
by: Xu, Ruyi, et al.
Published: (2024)
by: Xu, Ruyi, et al.
Published: (2024)
Mesh Neural Networks for SE(3)-Equivariant Hemodynamics Estimation on the Artery Wall
by: Suk, Julian, et al.
Published: (2022)
by: Suk, Julian, et al.
Published: (2022)
Similar Items
-
Reproducibility Study of "ITI-GEN: Inclusive Text-to-Image Generation"
by: Fernández, Daniel Gallo, et al.
Published: (2024) -
How to Train Neural Field Representations: A Comprehensive Study and Benchmark
by: Papa, Samuele, et al.
Published: (2023) -
DuoDiff: Accelerating Diffusion Models with a Dual-Backbone Approach
by: Fernández, Daniel Gallo, et al.
Published: (2024) -
From MLP to NeoMLP: Leveraging Self-Attention for Neural Fields
by: Kofinas, Miltiadis, et al.
Published: (2024) -
Dream to Manipulate: Compositional World Models Empowering Robot Imitation Learning with Imagination
by: Barcellona, Leonardo, et al.
Published: (2024)