Controllable Image Generation with Composed Parallel Token Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Stirling, Jamie, Al-Moubayed, Noura, Willcocks, Chris G., Shum, Hubert P. H. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Controllable Image Generation with Composed Parallel Token Prediction
von: Stirling, Jamie, et al.
Veröffentlicht: (2024)
von: Stirling, Jamie, et al.
Veröffentlicht: (2024)
Investigating Permutation-Invariant Discrete Representation Learning for Spatially Aligned Images
von: Stirling, Jamie S. J., et al.
Veröffentlicht: (2026)
von: Stirling, Jamie S. J., et al.
Veröffentlicht: (2026)
Inference-Time Decomposition of Activations (ITDA): A Scalable Approach to Interpreting Large Language Models
von: Leask, Patrick, et al.
Veröffentlicht: (2025)
von: Leask, Patrick, et al.
Veröffentlicht: (2025)
Video Prediction of Dynamic Physical Simulations With Pixel-Space Spatiotemporal Transformers
von: Slack, Dean L, et al.
Veröffentlicht: (2025)
von: Slack, Dean L, et al.
Veröffentlicht: (2025)
Disentangling Racial Phenotypes: Fine-Grained Control of Race-related Facial Phenotype Characteristics
von: Yucer, Seyma, et al.
Veröffentlicht: (2024)
von: Yucer, Seyma, et al.
Veröffentlicht: (2024)
Breaking Down Financial News Impact: A Novel AI Approach with Geometric Hypergraphs
von: Harit, Anoushka, et al.
Veröffentlicht: (2024)
von: Harit, Anoushka, et al.
Veröffentlicht: (2024)
Everything is a Video: Unifying Modalities through Next-Frame Prediction
von: Hudson, G. Thomas, et al.
Veröffentlicht: (2024)
von: Hudson, G. Thomas, et al.
Veröffentlicht: (2024)
Repeat and Concatenate: 2D to 3D Image Translation with 3D to 3D Generative Modeling
von: Corona-Figueroa, Abril, et al.
Veröffentlicht: (2024)
von: Corona-Figueroa, Abril, et al.
Veröffentlicht: (2024)
$\infty$-Diff: Infinite Resolution Diffusion with Subsampled Mollified States
von: Bond-Taylor, Sam, et al.
Veröffentlicht: (2023)
von: Bond-Taylor, Sam, et al.
Veröffentlicht: (2023)
Data Augmentation via Mixed Class Interpolation using Cycle-Consistent Generative Adversarial Networks Applied to Cross-Domain Imagery
von: Sasaki, Hiroshi, et al.
Veröffentlicht: (2020)
von: Sasaki, Hiroshi, et al.
Veröffentlicht: (2020)
The Power of Next-Frame Prediction for Learning Physical Laws
von: Winterbottom, Thomas, et al.
Veröffentlicht: (2024)
von: Winterbottom, Thomas, et al.
Veröffentlicht: (2024)
Textual Localization: Decomposing Multi-concept Images for Subject-Driven Text-to-Image Generation
von: Shentu, Junjie, et al.
Veröffentlicht: (2024)
von: Shentu, Junjie, et al.
Veröffentlicht: (2024)
Sparse Autoencoders Do Not Find Canonical Units of Analysis
von: Leask, Patrick, et al.
Veröffentlicht: (2025)
von: Leask, Patrick, et al.
Veröffentlicht: (2025)
Parallel Token Prediction for Language Models
von: Draxler, Felix, et al.
Veröffentlicht: (2025)
von: Draxler, Felix, et al.
Veröffentlicht: (2025)
RAPiD-Seg: Range-Aware Pointwise Distance Distribution Networks for 3D LiDAR Segmentation
von: Li, Li, et al.
Veröffentlicht: (2024)
von: Li, Li, et al.
Veröffentlicht: (2024)
Neuro-Symbolic Contrastive Learning for Cross-domain Inference
von: Liu, Mingyue, et al.
Veröffentlicht: (2025)
von: Liu, Mingyue, et al.
Veröffentlicht: (2025)
LLMs for LLMs: A Structured Prompting Methodology for Long Legal Documents
von: Klem, Strahinja, et al.
Veröffentlicht: (2025)
von: Klem, Strahinja, et al.
Veröffentlicht: (2025)
MuLD: The Multitask Long Document Benchmark
von: Hudson, G Thomas, et al.
Veröffentlicht: (2022)
von: Hudson, G Thomas, et al.
Veröffentlicht: (2022)
One-Index Vector Quantization Based Adversarial Attack on Image Classification
von: Fan, Haiju, et al.
Veröffentlicht: (2024)
von: Fan, Haiju, et al.
Veröffentlicht: (2024)
AttenCraft: Attention-guided Disentanglement of Multiple Concepts for Text-to-Image Customization
von: Shentu, Junjie, et al.
Veröffentlicht: (2024)
von: Shentu, Junjie, et al.
Veröffentlicht: (2024)
BOOST: Out-of-Distribution-Informed Adaptive Sampling for Bias Mitigation in Stylistic Convolutional Neural Networks
von: Vijendran, Mridula, et al.
Veröffentlicht: (2025)
von: Vijendran, Mridula, et al.
Veröffentlicht: (2025)
TraIL-Det: Transformation-Invariant Local Feature Networks for 3D LiDAR Object Detection with Unsupervised Pre-Training
von: Li, Li, et al.
Veröffentlicht: (2024)
von: Li, Li, et al.
Veröffentlicht: (2024)
Improving Token-Based World Models with Parallel Observation Prediction
von: Cohen, Lior, et al.
Veröffentlicht: (2024)
von: Cohen, Lior, et al.
Veröffentlicht: (2024)
Early Detection and Reduction of Memorisation for Domain Adaptation and Instruction Tuning
von: Slack, Dean L., et al.
Veröffentlicht: (2025)
von: Slack, Dean L., et al.
Veröffentlicht: (2025)
MIEB: Massive Image Embedding Benchmark
von: Xiao, Chenghao, et al.
Veröffentlicht: (2025)
von: Xiao, Chenghao, et al.
Veröffentlicht: (2025)
Dirichlet-Based Monte Carlo Dropout for Uncertainty Estimation in Neural Networks
von: Hoblos, Rouaa, et al.
Veröffentlicht: (2026)
von: Hoblos, Rouaa, et al.
Veröffentlicht: (2026)
On the Design Fundamentals of Diffusion Models: A Survey
von: Chang, Ziyi, et al.
Veröffentlicht: (2023)
von: Chang, Ziyi, et al.
Veröffentlicht: (2023)
RAR-b: Reasoning as Retrieval Benchmark
von: Xiao, Chenghao, et al.
Veröffentlicht: (2024)
von: Xiao, Chenghao, et al.
Veröffentlicht: (2024)
Orthrus: Memory-Efficient Parallel Token Generation via Dual-View Diffusion
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2026)
von: Van Nguyen, Chien, et al.
Veröffentlicht: (2026)
ProPD: Dynamic Token Tree Pruning and Generation for LLM Parallel Decoding
von: Zhong, Shuzhang, et al.
Veröffentlicht: (2024)
von: Zhong, Shuzhang, et al.
Veröffentlicht: (2024)
Generative Verifiers: Reward Modeling as Next-Token Prediction
von: Zhang, Lunjun, et al.
Veröffentlicht: (2024)
von: Zhang, Lunjun, et al.
Veröffentlicht: (2024)
Composable Crystals: Controllable Materials Discovery via Concept Learning
von: Liu, Nian, et al.
Veröffentlicht: (2026)
von: Liu, Nian, et al.
Veröffentlicht: (2026)
Proof of Concept: Multi-Target Wildfire Risk Prediction and Large Language Model Synthesis
von: Caron, Nicolas, et al.
Veröffentlicht: (2026)
von: Caron, Nicolas, et al.
Veröffentlicht: (2026)
Localized Forest Fire Risk Prediction: A Department-Aware Approach for Operational Decision Support
von: Caron, Nicolas, et al.
Veröffentlicht: (2025)
von: Caron, Nicolas, et al.
Veröffentlicht: (2025)
Multi-Token Residual Prediction
von: Xu, Yufeng, et al.
Veröffentlicht: (2026)
von: Xu, Yufeng, et al.
Veröffentlicht: (2026)
A Deep Learning Approach Towards Student Performance Prediction in Online Courses: Challenges Based on a Global Perspective
von: Moubayed, Abdallah, et al.
Veröffentlicht: (2024)
von: Moubayed, Abdallah, et al.
Veröffentlicht: (2024)
Triplet Synthesis For Enhancing Composed Image Retrieval via Counterfactual Image Generation
von: Uesugi, Kenta, et al.
Veröffentlicht: (2025)
von: Uesugi, Kenta, et al.
Veröffentlicht: (2025)
Separable Expert Architecture: Toward Privacy-Preserving LLM Personalization via Composable Adapters and Deletable User Proxies
von: Schneider, Chris, et al.
Veröffentlicht: (2026)
von: Schneider, Chris, et al.
Veröffentlicht: (2026)
GeoToken: Hierarchical Geolocalization of Images via Next Token Prediction
von: Ghasemi, Narges, et al.
Veröffentlicht: (2025)
von: Ghasemi, Narges, et al.
Veröffentlicht: (2025)
FACap: A Large-scale Fashion Dataset for Fine-grained Composed Image Retrieval
von: Gardères, François, et al.
Veröffentlicht: (2025)
von: Gardères, François, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Controllable Image Generation with Composed Parallel Token Prediction
von: Stirling, Jamie, et al.
Veröffentlicht: (2024) -
Investigating Permutation-Invariant Discrete Representation Learning for Spatially Aligned Images
von: Stirling, Jamie S. J., et al.
Veröffentlicht: (2026) -
Inference-Time Decomposition of Activations (ITDA): A Scalable Approach to Interpreting Large Language Models
von: Leask, Patrick, et al.
Veröffentlicht: (2025) -
Video Prediction of Dynamic Physical Simulations With Pixel-Space Spatiotemporal Transformers
von: Slack, Dean L, et al.
Veröffentlicht: (2025) -
Disentangling Racial Phenotypes: Fine-Grained Control of Race-related Facial Phenotype Characteristics
von: Yucer, Seyma, et al.
Veröffentlicht: (2024)