Controlla: Learning Controllability via Graph-Constrained Latent Geometry
Fuente:
arXiv
Saved in:
| Main Authors: | Murthy, Jamuna S., Monsefi, Amin Karimi, Ramnath, Rajiv |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Frequency-Guided Masking for Enhanced Vision Self-Supervised Learning
by: Monsefi, Amin Karimi, et al.
Published: (2024)
by: Monsefi, Amin Karimi, et al.
Published: (2024)
DiReCT: Disentangled Regularization of Contrastive Trajectories for Physics-Refined Video Generation
by: Meyarian, Abolfazl, et al.
Published: (2026)
by: Meyarian, Abolfazl, et al.
Published: (2026)
DetailCLIP: Detail-Oriented CLIP for Fine-Grained Tasks
by: Monsefi, Amin Karimi, et al.
Published: (2024)
by: Monsefi, Amin Karimi, et al.
Published: (2024)
KnobGen: Controlling the Sophistication of Artwork in Sketch-Based Diffusion Models
by: Navard, Pouyan, et al.
Published: (2024)
by: Navard, Pouyan, et al.
Published: (2024)
TaxaDiffusion: Progressively Trained Diffusion Model for Fine-Grained Species Generation
by: Monsefi, Amin Karimi, et al.
Published: (2025)
by: Monsefi, Amin Karimi, et al.
Published: (2025)
Masked LoGoNet: Fast and Accurate 3D Image Analysis for Medical Domain
by: Monsefi, Amin Karimi, et al.
Published: (2024)
by: Monsefi, Amin Karimi, et al.
Published: (2024)
SeamCam: Quantifying Seamless Camouflage via Multi-Cue Visual Detectability
by: Monsefi, Amin Karimi, et al.
Published: (2026)
by: Monsefi, Amin Karimi, et al.
Published: (2026)
TaxaAdapter: Vision Taxonomy Models are Key to Fine-grained Image Generation over the Tree of Life
by: Khurana, Mridul, et al.
Published: (2026)
by: Khurana, Mridul, et al.
Published: (2026)
MGA-VQA: Secure and Interpretable Graph-Augmented Visual Question Answering with Memory-Guided Protection Against Unauthorized Knowledge Use
by: Mohammadshirazi, Ahmad, et al.
Published: (2025)
by: Mohammadshirazi, Ahmad, et al.
Published: (2025)
DocParseNet: Advanced Semantic Segmentation and OCR Embeddings for Efficient Scanned Document Annotation
by: Mohammadshirazi, Ahmad, et al.
Published: (2024)
by: Mohammadshirazi, Ahmad, et al.
Published: (2024)
ARIAL: An Agentic Framework for Document VQA with Precise Answer Localization
by: Mohammadshirazi, Ahmad, et al.
Published: (2025)
by: Mohammadshirazi, Ahmad, et al.
Published: (2025)
DLaVA: Document Language and Vision Assistant for Answer Localization with Enhanced Interpretability and Trustworthiness
by: Mohammadshirazi, Ahmad, et al.
Published: (2024)
by: Mohammadshirazi, Ahmad, et al.
Published: (2024)
DSV-LFS: Unifying LLM-Driven Semantic Cues with Visual Features for Robust Few-Shot Segmentation
by: Karimi, Amin, et al.
Published: (2025)
by: Karimi, Amin, et al.
Published: (2025)
LatentKeypointGAN: Controlling Images via Latent Keypoints
by: He, Xingzhe, et al.
Published: (2021)
by: He, Xingzhe, et al.
Published: (2021)
GeoMotion: Rethinking Motion Segmentation via Latent 4D Geometry
by: He, Xiankang, et al.
Published: (2026)
by: He, Xiankang, et al.
Published: (2026)
Textured Geometry Evaluation: Perceptual 3D Textured Shape Metric via 3D Latent-Geometry Network
by: Luan, Tianyu, et al.
Published: (2025)
by: Luan, Tianyu, et al.
Published: (2025)
GeoTopoDiff: Learning Geometry--Topology Graph Priors through Boundary-Constrained Mixed Diffusion for Sparse-Slice 3D Porous Reconstruction
by: Shi, Yue, et al.
Published: (2026)
by: Shi, Yue, et al.
Published: (2026)
Joint Geometry-Appearance Human Reconstruction in a Unified Latent Space via Bridge Diffusion
by: Tang, Yingzhi, et al.
Published: (2026)
by: Tang, Yingzhi, et al.
Published: (2026)
Learning Latent Proxies for Controllable Single-Image Relighting
by: Zheng, Haoze, et al.
Published: (2026)
by: Zheng, Haoze, et al.
Published: (2026)
Aligning Latent Geometry for Spherical Flow Matching in Image Generation
by: Meral, Tuna Han Salih, et al.
Published: (2026)
by: Meral, Tuna Han Salih, et al.
Published: (2026)
Learning Compact Latent Space for Representing Neural Signed Distance Functions with High-fidelity Geometry Details
by: Bai, Qiang, et al.
Published: (2025)
by: Bai, Qiang, et al.
Published: (2025)
Localized Control in Diffusion Models via Latent Vector Prediction
by: Domingo-Gregorio, Pablo, et al.
Published: (2026)
by: Domingo-Gregorio, Pablo, et al.
Published: (2026)
Latent Harmony: Synergistic Unified UHD Image Restoration via Latent Space Regularization and Controllable Refinement
by: Liu, Yidi, et al.
Published: (2025)
by: Liu, Yidi, et al.
Published: (2025)
Graph Smoothing for Enhanced Local Geometry Learning in Point Cloud Analysis
by: Yuan, Shangbo, et al.
Published: (2026)
by: Yuan, Shangbo, et al.
Published: (2026)
Fus3D: Decoding Consolidated 3D Geometry from Feed-forward Geometry Transformer Latents
by: Fink, Laura, et al.
Published: (2026)
by: Fink, Laura, et al.
Published: (2026)
Accelerating Masked Image Generation by Learning Latent Controlled Dynamics
by: Zhu, Kaiwen, et al.
Published: (2026)
by: Zhu, Kaiwen, et al.
Published: (2026)
CLAD: Constrained Latent Action Diffusion for Vision-Language Procedure Planning
by: Shi, Lei, et al.
Published: (2025)
by: Shi, Lei, et al.
Published: (2025)
GeoSplat: A Deep Dive into Geometry-Constrained Gaussian Splatting
by: Li, Yangming, et al.
Published: (2025)
by: Li, Yangming, et al.
Published: (2025)
GraphMAR: Geometry-Aware Graph Learning Framework for Spatially Adaptive CT Metal Artifact Reduction
by: Li, Zilong, et al.
Published: (2026)
by: Li, Zilong, et al.
Published: (2026)
Hybrid Latents: Geometry-Appearance-Aware Surfel Splatting
by: Kelkar, Neel, et al.
Published: (2026)
by: Kelkar, Neel, et al.
Published: (2026)
Learning Semantic Latent Directions for Accurate and Controllable Human Motion Prediction
by: Xu, Guowei, et al.
Published: (2024)
by: Xu, Guowei, et al.
Published: (2024)
Geometry-Constrained Monocular Scale Estimation Using Semantic Segmentation for Dynamic Scenes
by: Zhang, Hui, et al.
Published: (2025)
by: Zhang, Hui, et al.
Published: (2025)
PAGCNet: A Pose-Aware and Geometry Constrained Framework for Panoramic Depth Estimation
by: Ning, Kanglin, et al.
Published: (2025)
by: Ning, Kanglin, et al.
Published: (2025)
MoRe: Monocular Geometry Refinement via Graph Optimization for Cross-View Consistency
by: Jung, Dongki, et al.
Published: (2025)
by: Jung, Dongki, et al.
Published: (2025)
Learning Object Permanence from Videos via Latent Imaginations
by: Traub, Manuel, et al.
Published: (2023)
by: Traub, Manuel, et al.
Published: (2023)
Reframing Long-Tailed Learning via Loss Landscape Geometry
by: Chen, Shenghan, et al.
Published: (2026)
by: Chen, Shenghan, et al.
Published: (2026)
OmniBooth: Learning Latent Control for Image Synthesis with Multi-modal Instruction
by: Li, Leheng, et al.
Published: (2024)
by: Li, Leheng, et al.
Published: (2024)
DriveVGGT: Calibration-Constrained Visual Geometry Transformers for Multi-Camera Autonomous Driving
by: Jia, Xiaosong, et al.
Published: (2025)
by: Jia, Xiaosong, et al.
Published: (2025)
SA-GS: Semantic-Aware Gaussian Splatting for Large Scene Reconstruction with Geometry Constrain
by: Xiong, Butian, et al.
Published: (2024)
by: Xiong, Butian, et al.
Published: (2024)
CARE: Training-Free Controllable Restoration for Medical Images via Dual-Latent Steering
by: Liu, Xu
Published: (2026)
by: Liu, Xu
Published: (2026)
Similar Items
-
Frequency-Guided Masking for Enhanced Vision Self-Supervised Learning
by: Monsefi, Amin Karimi, et al.
Published: (2024) -
DiReCT: Disentangled Regularization of Contrastive Trajectories for Physics-Refined Video Generation
by: Meyarian, Abolfazl, et al.
Published: (2026) -
DetailCLIP: Detail-Oriented CLIP for Fine-Grained Tasks
by: Monsefi, Amin Karimi, et al.
Published: (2024) -
KnobGen: Controlling the Sophistication of Artwork in Sketch-Based Diffusion Models
by: Navard, Pouyan, et al.
Published: (2024) -
TaxaDiffusion: Progressively Trained Diffusion Model for Fine-Grained Species Generation
by: Monsefi, Amin Karimi, et al.
Published: (2025)