Salvato in:
| Autori principali: | Goyal, Shreya, Khan, Naimul, Chattopadhyay, Chiranjoy, Bhatnagar, Gaurav |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2021
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2103.08297 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Knowledge driven Description Synthesis for Floor Plan Interpretation
di: Goyal, Shreya, et al.
Pubblicazione: (2021)
di: Goyal, Shreya, et al.
Pubblicazione: (2021)
Enhancing Fruit and Vegetable Detection in Unconstrained Environment with a Novel Dataset
di: Khanna, Sandeep, et al.
Pubblicazione: (2024)
di: Khanna, Sandeep, et al.
Pubblicazione: (2024)
DynaSeg: A Deep Dynamic Fusion Method for Unsupervised Image Segmentation Incorporating Feature Similarity and Spatial Continuity
di: Guermazi, Boujemaa, et al.
Pubblicazione: (2024)
di: Guermazi, Boujemaa, et al.
Pubblicazione: (2024)
DynaGuide: A Generalizable Dynamic Guidance Framework for Unsupervised Semantic Segmentation
di: Guermazi, Boujemaa, et al.
Pubblicazione: (2026)
di: Guermazi, Boujemaa, et al.
Pubblicazione: (2026)
FaceGemma: Enhancing Image Captioning with Facial Attributes for Portrait Images
di: Haque, Naimul, et al.
Pubblicazione: (2023)
di: Haque, Naimul, et al.
Pubblicazione: (2023)
Temporal Feature Weaving for Neonatal Echocardiographic Viewpoint Video Classification
di: French, Satchel, et al.
Pubblicazione: (2025)
di: French, Satchel, et al.
Pubblicazione: (2025)
Stress Classification from ECG Signals Using Vision Transformer
di: Ahmad, Zeeshan, et al.
Pubblicazione: (2026)
di: Ahmad, Zeeshan, et al.
Pubblicazione: (2026)
Bridging the Gap: Studio-like Avatar Creation from a Monocular Phone Capture
di: Athar, ShahRukh, et al.
Pubblicazione: (2024)
di: Athar, ShahRukh, et al.
Pubblicazione: (2024)
Capture, Canonicalize, Splat: Zero-Shot 3D Gaussian Avatars from Unstructured Phone Images
di: Garbin, Emanuel, et al.
Pubblicazione: (2025)
di: Garbin, Emanuel, et al.
Pubblicazione: (2025)
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
di: Chattopadhyay, Nandish, et al.
Pubblicazione: (2026)
di: Chattopadhyay, Nandish, et al.
Pubblicazione: (2026)
AUGCAL: Improving Sim2Real Adaptation by Uncertainty Calibration on Augmented Synthetic Images
di: Chattopadhyay, Prithvijit, et al.
Pubblicazione: (2023)
di: Chattopadhyay, Prithvijit, et al.
Pubblicazione: (2023)
GazeProphetV2: Head-Movement-Based Gaze Prediction Enabling Efficient Foveated Rendering on Mobile VR
di: Ebadulla, Farhaan, et al.
Pubblicazione: (2025)
di: Ebadulla, Farhaan, et al.
Pubblicazione: (2025)
Build-A-Scene: Interactive 3D Layout Control for Diffusion-Based Image Generation
di: Eldesokey, Abdelrahman, et al.
Pubblicazione: (2024)
di: Eldesokey, Abdelrahman, et al.
Pubblicazione: (2024)
Maximizing Generalization: The Effect of Different Augmentation Techniques on Lightweight Vision Transformer for Bengali Character Classification
di: Chowdhury, Rafi Hassan, et al.
Pubblicazione: (2026)
di: Chowdhury, Rafi Hassan, et al.
Pubblicazione: (2026)
Computer-Aided Layout Generation for Building Design: A Review
di: Liu, Jiachen, et al.
Pubblicazione: (2025)
di: Liu, Jiachen, et al.
Pubblicazione: (2025)
Image Demoiréing Using Dual Camera Fusion on Mobile Phones
di: Mei, Yanting, et al.
Pubblicazione: (2025)
di: Mei, Yanting, et al.
Pubblicazione: (2025)
uLayout: Unified Room Layout Estimation for Perspective and Panoramic Images
di: Lee, Jonathan, et al.
Pubblicazione: (2025)
di: Lee, Jonathan, et al.
Pubblicazione: (2025)
Efficient Hybrid Zoom using Camera Fusion on Mobile Phones
di: Wu, Xiaotong, et al.
Pubblicazione: (2024)
di: Wu, Xiaotong, et al.
Pubblicazione: (2024)
CreatiLayout: Siamese Multimodal Diffusion Transformer for Creative Layout-to-Image Generation
di: Zhang, Hui, et al.
Pubblicazione: (2024)
di: Zhang, Hui, et al.
Pubblicazione: (2024)
STAY Diffusion: Styled Layout Diffusion Model for Diverse Layout-to-Image Generation
di: Wang, Ruyu, et al.
Pubblicazione: (2025)
di: Wang, Ruyu, et al.
Pubblicazione: (2025)
Tokenizing Buildings: A Transformer for Layout Synthesis
di: de Guevara, Manuel Ladron, et al.
Pubblicazione: (2025)
di: de Guevara, Manuel Ladron, et al.
Pubblicazione: (2025)
Chat2Layout: Interactive 3D Furniture Layout with a Multimodal LLM
di: Wang, Can, et al.
Pubblicazione: (2024)
di: Wang, Can, et al.
Pubblicazione: (2024)
Consistent Image Layout Editing with Diffusion Models
di: Xia, Tao, et al.
Pubblicazione: (2025)
di: Xia, Tao, et al.
Pubblicazione: (2025)
Layout-to-Image Generation with Localized Descriptions using ControlNet with Cross-Attention Control
di: Lukovnikov, Denis, et al.
Pubblicazione: (2024)
di: Lukovnikov, Denis, et al.
Pubblicazione: (2024)
BornoViT: A Novel Efficient Vision Transformer for Bengali Handwritten Basic Characters Classification
di: Chowdhury, Rafi Hassan, et al.
Pubblicazione: (2026)
di: Chowdhury, Rafi Hassan, et al.
Pubblicazione: (2026)
ToLo: A Two-Stage, Training-Free Layout-To-Image Generation Framework For High-Overlap Layouts
di: Huang, Linhao, et al.
Pubblicazione: (2025)
di: Huang, Linhao, et al.
Pubblicazione: (2025)
LayoutFlow: Flow Matching for Layout Generation
di: Guerreiro, Julian Jorge Andrade, et al.
Pubblicazione: (2024)
di: Guerreiro, Julian Jorge Andrade, et al.
Pubblicazione: (2024)
Dual-Camera Smooth Zoom on Mobile Phones
di: Wu, Renlong, et al.
Pubblicazione: (2024)
di: Wu, Renlong, et al.
Pubblicazione: (2024)
SciPostLayout: A Dataset for Layout Analysis and Layout Generation of Scientific Posters
di: Tanaka, Shohei, et al.
Pubblicazione: (2024)
di: Tanaka, Shohei, et al.
Pubblicazione: (2024)
Griffin: Generative Reference and Layout Guided Image Composition
di: Mikaeili, Aryan, et al.
Pubblicazione: (2025)
di: Mikaeili, Aryan, et al.
Pubblicazione: (2025)
LayoutDETR: Detection Transformer Is a Good Multimodal Layout Designer
di: Yu, Ning, et al.
Pubblicazione: (2022)
di: Yu, Ning, et al.
Pubblicazione: (2022)
MDTv2: Masked Diffusion Transformer is a Strong Image Synthesizer
di: Gao, Shanghua, et al.
Pubblicazione: (2023)
di: Gao, Shanghua, et al.
Pubblicazione: (2023)
An Object-Centered Data Acquisition Method for 3D Gaussian Splatting using Mobile Phones
di: Zhang, Yuezhe, et al.
Pubblicazione: (2026)
di: Zhang, Yuezhe, et al.
Pubblicazione: (2026)
Synthesizing Iris Images using Generative Adversarial Networks: Survey and Comparative Analysis
di: Yadav, Shivangi, et al.
Pubblicazione: (2024)
di: Yadav, Shivangi, et al.
Pubblicazione: (2024)
Layout Agnostic Scene Text Image Synthesis with Diffusion Models
di: Zhangli, Qilong, et al.
Pubblicazione: (2024)
di: Zhangli, Qilong, et al.
Pubblicazione: (2024)
Rethinking The Training And Evaluation of Rich-Context Layout-to-Image Generation
di: Cheng, Jiaxin, et al.
Pubblicazione: (2024)
di: Cheng, Jiaxin, et al.
Pubblicazione: (2024)
Training-Free Layout-to-Image Generation with Marginal Attention Constraints
di: Chen, Huancheng, et al.
Pubblicazione: (2024)
di: Chen, Huancheng, et al.
Pubblicazione: (2024)
ConsistCompose: Unified Multimodal Layout Control for Image Composition
di: Shi, Xuanke, et al.
Pubblicazione: (2025)
di: Shi, Xuanke, et al.
Pubblicazione: (2025)
PLACE: Adaptive Layout-Semantic Fusion for Semantic Image Synthesis
di: Lv, Zhengyao, et al.
Pubblicazione: (2024)
di: Lv, Zhengyao, et al.
Pubblicazione: (2024)
IMAGHarmony: Controllable Image Editing with Consistent Object Quantity and Layout
di: Shen, Fei, et al.
Pubblicazione: (2025)
di: Shen, Fei, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Knowledge driven Description Synthesis for Floor Plan Interpretation
di: Goyal, Shreya, et al.
Pubblicazione: (2021) -
Enhancing Fruit and Vegetable Detection in Unconstrained Environment with a Novel Dataset
di: Khanna, Sandeep, et al.
Pubblicazione: (2024) -
DynaSeg: A Deep Dynamic Fusion Method for Unsupervised Image Segmentation Incorporating Feature Similarity and Spatial Continuity
di: Guermazi, Boujemaa, et al.
Pubblicazione: (2024) -
DynaGuide: A Generalizable Dynamic Guidance Framework for Unsupervised Semantic Segmentation
di: Guermazi, Boujemaa, et al.
Pubblicazione: (2026) -
FaceGemma: Enhancing Image Captioning with Facial Attributes for Portrait Images
di: Haque, Naimul, et al.
Pubblicazione: (2023)