Saved in:
| Main Authors: | Dosi, Dhruv, Meena, Rohit, Rajpura, Param, Meena, Yogesh Kumar |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2508.10449 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Optimising EEG Decoding using Post-hoc Explanations and Domain Knowledge
by: Rajpura, Param, et al.
Published: (2024)
by: Rajpura, Param, et al.
Published: (2024)
Quantifying Spatial Domain Explanations in BCI using Earth Mover's Distance
by: Rajpura, Param, et al.
Published: (2024)
by: Rajpura, Param, et al.
Published: (2024)
Can EEG resting state data benefit data-driven approaches for motor-imagery decoding?
by: Mehta, Rishan, et al.
Published: (2024)
by: Mehta, Rishan, et al.
Published: (2024)
MultiFusionNet: Multilayer Multimodal Fusion of Deep Neural Networks for Chest X-Ray Image Classification
by: Agarwal, Saurabh, et al.
Published: (2024)
by: Agarwal, Saurabh, et al.
Published: (2024)
CHOrD: Generation of Collision-Free, House-Scale, and Organized Digital Twins for 3D Indoor Scenes with Controllable Floor Plans and Optimal Layouts
by: Su, Chong, et al.
Published: (2025)
by: Su, Chong, et al.
Published: (2025)
Are We Ready for Out-of-Distribution Detection in Digital Pathology?
by: Oh, Ji-Hun, et al.
Published: (2024)
by: Oh, Ji-Hun, et al.
Published: (2024)
A Volumetric Saliency Guided Image Summarization for RGB-D Indoor Scene Classification
by: Meena, Preeti, et al.
Published: (2024)
by: Meena, Preeti, et al.
Published: (2024)
Masked Training for Robust Arrhythmia Detection from Digitalized Multiple Layout ECG Images
by: Zhang, Shanwei, et al.
Published: (2025)
by: Zhang, Shanwei, et al.
Published: (2025)
Automated Knot Detection and Pairing for Wood Analysis in the Timber Industry
by: Lin, Guohao, et al.
Published: (2025)
by: Lin, Guohao, et al.
Published: (2025)
Improving Medical Multi-modal Contrastive Learning with Expert Annotations
by: Kumar, Yogesh, et al.
Published: (2024)
by: Kumar, Yogesh, et al.
Published: (2024)
Disentangling Polysemantic Neurons with a Null-Calibrated Polysemanticity Index and Causal Patch Interventions
by: Gupta, Manan, et al.
Published: (2025)
by: Gupta, Manan, et al.
Published: (2025)
LayoutAgent: A Vision-Language Agent Guided Compositional Diffusion for Spatial Layout Planning
by: Fan, Zezhong, et al.
Published: (2025)
by: Fan, Zezhong, et al.
Published: (2025)
Scene Representation using 360° Saliency Graph and its Application in Vision-based Indoor Navigation
by: Meena, Preeti, et al.
Published: (2026)
by: Meena, Preeti, et al.
Published: (2026)
MALeR: Improving Compositional Fidelity in Layout-Guided Generation
by: Saxena, Shivank, et al.
Published: (2025)
by: Saxena, Shivank, et al.
Published: (2025)
Knowledge Detection by Relevant Question and Image Attributes in Visual Question Answering
by: Ahir, Param, et al.
Published: (2023)
by: Ahir, Param, et al.
Published: (2023)
Spot the Error: Non-autoregressive Graphic Layout Generation with Wireframe Locator
by: Lin, Jieru, et al.
Published: (2024)
by: Lin, Jieru, et al.
Published: (2024)
SpotActor: Training-Free Layout-Controlled Consistent Image Generation
by: Wang, Jiahao, et al.
Published: (2024)
by: Wang, Jiahao, et al.
Published: (2024)
ReLayout: Integrating Relation Reasoning for Content-aware Layout Generation with Multi-modal Large Language Models
by: Tian, Jiaxu, et al.
Published: (2025)
by: Tian, Jiaxu, et al.
Published: (2025)
Less is More: AMBER-AFNO -- a New Benchmark for Lightweight 3D Medical Image Segmentation
by: Dosi, Andrea, et al.
Published: (2025)
by: Dosi, Andrea, et al.
Published: (2025)
Uni-Layout: Integrating Human Feedback in Unified Layout Generation and Evaluation
by: Lu, Shuo, et al.
Published: (2025)
by: Lu, Shuo, et al.
Published: (2025)
Language-Guided Temporal Token Pruning for Efficient VideoLLM Processing
by: Kumar, Yogesh
Published: (2025)
by: Kumar, Yogesh
Published: (2025)
Perceive-then-Plan: Layout-as-Policy for Monocular 3D Scene Layout Estimation
by: Zhou, Junwei, et al.
Published: (2026)
by: Zhou, Junwei, et al.
Published: (2026)
DRiVE: Dynamic Recognition in VEhicles using snnTorch
by: Vora, Heerak, et al.
Published: (2025)
by: Vora, Heerak, et al.
Published: (2025)
Evaluating the Robustness of Off-Road Autonomous Driving Segmentation against Adversarial Attacks: A Dataset-Centric analysis
by: Deoli, Pankaj, et al.
Published: (2024)
by: Deoli, Pankaj, et al.
Published: (2024)
Disentangled 3D Scene Generation with Layout Learning
by: Epstein, Dave, et al.
Published: (2024)
by: Epstein, Dave, et al.
Published: (2024)
PP-DocLayout: A Unified Document Layout Detection Model to Accelerate Large-Scale Data Construction
by: Sun, Ting, et al.
Published: (2025)
by: Sun, Ting, et al.
Published: (2025)
Temporal Object-Aware Vision Transformer for Few-Shot Video Object Detection
by: Kumar, Yogesh, et al.
Published: (2025)
by: Kumar, Yogesh, et al.
Published: (2025)
Layout-Corrector: Alleviating Layout Sticking Phenomenon in Discrete Diffusion Model
by: Iwai, Shoma, et al.
Published: (2024)
by: Iwai, Shoma, et al.
Published: (2024)
A Self-Supervised Learning of a Foundation Model for Analog Layout Design Automation
by: Jeong, Sungyu, et al.
Published: (2025)
by: Jeong, Sungyu, et al.
Published: (2025)
EZ-CLIP: Efficient Zeroshot Video Action Recognition
by: Ahmad, Shahzad, et al.
Published: (2023)
by: Ahmad, Shahzad, et al.
Published: (2023)
Tokenizing Buildings: A Transformer for Layout Synthesis
by: de Guevara, Manuel Ladron, et al.
Published: (2025)
by: de Guevara, Manuel Ladron, et al.
Published: (2025)
EgoVITA: Learning to Plan and Verify for Egocentric Video Reasoning
by: Kulkarni, Yogesh, et al.
Published: (2025)
by: Kulkarni, Yogesh, et al.
Published: (2025)
Stable Mean Teacher for Semi-supervised Video Action Detection
by: Kumar, Akash, et al.
Published: (2024)
by: Kumar, Akash, et al.
Published: (2024)
The Effectiveness of Edge Detection Evaluation Metrics for Automated Coastline Detection
by: O'Sullivan, Conor, et al.
Published: (2024)
by: O'Sullivan, Conor, et al.
Published: (2024)
Harmonizing Geometry and Uncertainty: Diffusion with Hyperspheres
by: Dosi, Muskan, et al.
Published: (2025)
by: Dosi, Muskan, et al.
Published: (2025)
HyperSpaceX: Radial and Angular Exploration of HyperSpherical Dimensions
by: Chiranjeev, Chiranjeev, et al.
Published: (2024)
by: Chiranjeev, Chiranjeev, et al.
Published: (2024)
Mixture of Experts in Image Classification: What's the Sweet Spot?
by: Videau, Mathurin, et al.
Published: (2024)
by: Videau, Mathurin, et al.
Published: (2024)
MINT: Multimodal Imaging-to-Speech Knowledge Transfer for Early Alzheimer's Screening
by: Ahire, Vrushank, et al.
Published: (2026)
by: Ahire, Vrushank, et al.
Published: (2026)
LayoutDETR: Detection Transformer Is a Good Multimodal Layout Designer
by: Yu, Ning, et al.
Published: (2022)
by: Yu, Ning, et al.
Published: (2022)
Lay-Your-Scene: Natural Scene Layout Generation with Diffusion Transformers
by: Srivastava, Divyansh, et al.
Published: (2025)
by: Srivastava, Divyansh, et al.
Published: (2025)
Similar Items
-
Towards Optimising EEG Decoding using Post-hoc Explanations and Domain Knowledge
by: Rajpura, Param, et al.
Published: (2024) -
Quantifying Spatial Domain Explanations in BCI using Earth Mover's Distance
by: Rajpura, Param, et al.
Published: (2024) -
Can EEG resting state data benefit data-driven approaches for motor-imagery decoding?
by: Mehta, Rishan, et al.
Published: (2024) -
MultiFusionNet: Multilayer Multimodal Fusion of Deep Neural Networks for Chest X-Ray Image Classification
by: Agarwal, Saurabh, et al.
Published: (2024) -
CHOrD: Generation of Collision-Free, House-Scale, and Organized Digital Twins for 3D Indoor Scenes with Controllable Floor Plans and Optimal Layouts
by: Su, Chong, et al.
Published: (2025)