CLID: Controlled-Length Image Descriptions with Limited Data
Fuente:
arXiv
Guardado en:
| Autores principales: | Hirsch, Elad, Tal, Ayellet |
|---|---|
| Formato: | Preprint |
| Publicado: |
2022
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Image-aware Evaluation of Generated Medical Reports
por: Dawidowicz, Gefen, et al.
Publicado: (2024)
por: Dawidowicz, Gefen, et al.
Publicado: (2024)
MedRAT: Unpaired Medical Report Generation via Auxiliary Tasks
por: Hirsch, Elad, et al.
Publicado: (2024)
por: Hirsch, Elad, et al.
Publicado: (2024)
MedCycle: Unpaired Medical Report Generation via Cycle-Consistency
por: Hirsch, Elad, et al.
Publicado: (2024)
por: Hirsch, Elad, et al.
Publicado: (2024)
SFMNet: Sparse Focal Modulation for 3D Object Detection
por: Shrout, Oren, et al.
Publicado: (2025)
por: Shrout, Oren, et al.
Publicado: (2025)
MME: Mixture of Mesh Experts with Random Walk Transformer Gating
por: Belder, Amir, et al.
Publicado: (2026)
por: Belder, Amir, et al.
Publicado: (2026)
Random Walks in Self-supervised Learning for Triangular Meshes
por: Yefet, Gal, et al.
Publicado: (2025)
por: Yefet, Gal, et al.
Publicado: (2025)
Concept Retrieval -- What and How?
por: Nizan, Ori, et al.
Publicado: (2025)
por: Nizan, Ori, et al.
Publicado: (2025)
VALD: Multi-Stage Vision Attack Detection for Efficient LVLM Defense
por: Kadvil, Nadav, et al.
Publicado: (2026)
por: Kadvil, Nadav, et al.
Publicado: (2026)
GraVoS: Voxel Selection for 3D Point-Cloud Detection
por: Shrout, Oren, et al.
Publicado: (2022)
por: Shrout, Oren, et al.
Publicado: (2022)
ArcAid: Analysis of Archaeological Artifacts using Drawings
por: Hayon, Offry, et al.
Publicado: (2022)
por: Hayon, Offry, et al.
Publicado: (2022)
PatchContrast: Self-Supervised Pre-training for 3D Object Detection
por: Shrout, Oren, et al.
Publicado: (2023)
por: Shrout, Oren, et al.
Publicado: (2023)
LICA: Layered Image Composition Annotations for Graphic Design Research
por: Hirsch, Elad, et al.
Publicado: (2026)
por: Hirsch, Elad, et al.
Publicado: (2026)
Text-to-Image Generation Via Energy-Based CLIP
por: Ganz, Roy, et al.
Publicado: (2024)
por: Ganz, Roy, et al.
Publicado: (2024)
Sample- and Parameter-Efficient Auto-Regressive Image Models
por: Amrani, Elad, et al.
Publicado: (2024)
por: Amrani, Elad, et al.
Publicado: (2024)
Palette Aligned Image Diffusion
por: Aharoni, Elad, et al.
Publicado: (2025)
por: Aharoni, Elad, et al.
Publicado: (2025)
Layout-to-Image Generation with Localized Descriptions using ControlNet with Cross-Attention Control
por: Lukovnikov, Denis, et al.
Publicado: (2024)
por: Lukovnikov, Denis, et al.
Publicado: (2024)
TASTE: A Designer-Annotated Multi-Dimensional Preference Dataset for AI-Generated Graphic Design
por: Zhu, Haonan, et al.
Publicado: (2026)
por: Zhu, Haonan, et al.
Publicado: (2026)
Evaluating Design Video Generation: Metrics for Compositional Fidelity
por: Deganutti, Adrienne, et al.
Publicado: (2026)
por: Deganutti, Adrienne, et al.
Publicado: (2026)
Advancing Image Classification with Discrete Diffusion Classification Modeling
por: Belhasin, Omer, et al.
Publicado: (2025)
por: Belhasin, Omer, et al.
Publicado: (2025)
Images are Worth Variable Length of Representations
por: Mao, Lingjun, et al.
Publicado: (2025)
por: Mao, Lingjun, et al.
Publicado: (2025)
4-LEGS: 4D Language Embedded Gaussian Splatting
por: Fiebelman, Gal, et al.
Publicado: (2024)
por: Fiebelman, Gal, et al.
Publicado: (2024)
CoatFusion: Controllable Material Coating in Images
por: Levy, Sagie, et al.
Publicado: (2025)
por: Levy, Sagie, et al.
Publicado: (2025)
BBQ-to-Image: Numeric Bounding Box and Qolor Control in Large-Scale Text-to-Image Models
por: Kachlon, Eliran, et al.
Publicado: (2026)
por: Kachlon, Eliran, et al.
Publicado: (2026)
Leveraging Vision-Language Foundation Models to Reveal Hidden Image-Attribute Relationships in Medical Imaging
por: Kumar, Amar, et al.
Publicado: (2025)
por: Kumar, Amar, et al.
Publicado: (2025)
Principal Uncertainty Quantification with Spatial Correlation for Image Restoration Problems
por: Belhasin, Omer, et al.
Publicado: (2023)
por: Belhasin, Omer, et al.
Publicado: (2023)
Class-Conditioned Transformation for Enhanced Robust Image Classification
por: Blau, Tsachi, et al.
Publicado: (2023)
por: Blau, Tsachi, et al.
Publicado: (2023)
RL4Med-DDPO: Reinforcement Learning for Controlled Guidance Towards Diverse Medical Image Generation using Vision-Language Foundation Models
por: Saremi, Parham, et al.
Publicado: (2025)
por: Saremi, Parham, et al.
Publicado: (2025)
Self-supervised Learning of Dense Hierarchical Representations for Medical Image Segmentation
por: Kats, Eytan, et al.
Publicado: (2024)
por: Kats, Eytan, et al.
Publicado: (2024)
Graphic-Design-Bench: A Comprehensive Benchmark for Evaluating AI on Graphic Design Tasks
por: Deganutti, Adrienne, et al.
Publicado: (2026)
por: Deganutti, Adrienne, et al.
Publicado: (2026)
Performance of Human Annotators in Object Detection and Segmentation of Remotely Sensed Data
por: Blushtein-Livnon, Roni, et al.
Publicado: (2024)
por: Blushtein-Livnon, Roni, et al.
Publicado: (2024)
Real Classification by Description: Extending CLIP's Limits of Part Attributes Recognition
por: Baron, Ethan, et al.
Publicado: (2024)
por: Baron, Ethan, et al.
Publicado: (2024)
FastJAM: a Fast Joint Alignment Model for Images
por: Hirsch, Omri, et al.
Publicado: (2025)
por: Hirsch, Omri, et al.
Publicado: (2025)
Language-Guided Trajectory Traversal in Disentangled Stable Diffusion Latent Space for Factorized Medical Image Generation
por: TehraniNasab, Zahra, et al.
Publicado: (2025)
por: TehraniNasab, Zahra, et al.
Publicado: (2025)
Detailed Object Description with Controllable Dimensions
por: Wang, Xinran, et al.
Publicado: (2024)
por: Wang, Xinran, et al.
Publicado: (2024)
Enhancing Consistency-Based Image Generation via Adversarialy-Trained Classification and Energy-Based Discrimination
por: Golan, Shelly, et al.
Publicado: (2024)
por: Golan, Shelly, et al.
Publicado: (2024)
Keypoint Detection and Description for Raw Bayer Images
por: Lin, Jiakai, et al.
Publicado: (2025)
por: Lin, Jiakai, et al.
Publicado: (2025)
Personalized Image Descriptions from Attention Sequences
por: Xue, Ruoyu, et al.
Publicado: (2025)
por: Xue, Ruoyu, et al.
Publicado: (2025)
DiffuSAM: Diffusion-Based Prompt-Free SAM2 for Few-Shot and Source-Free Medical Image Segmentation
por: Grossman, Tal, et al.
Publicado: (2026)
por: Grossman, Tal, et al.
Publicado: (2026)
Internal Organ Localization Using Depth Images
por: Kats, Eytan, et al.
Publicado: (2025)
por: Kats, Eytan, et al.
Publicado: (2025)
Editor's Choice: Evaluating Abstract Intent in Image Editing through Atomic Entity Analysis
por: Ventura, Mor, et al.
Publicado: (2026)
por: Ventura, Mor, et al.
Publicado: (2026)
Ejemplares similares
-
Image-aware Evaluation of Generated Medical Reports
por: Dawidowicz, Gefen, et al.
Publicado: (2024) -
MedRAT: Unpaired Medical Report Generation via Auxiliary Tasks
por: Hirsch, Elad, et al.
Publicado: (2024) -
MedCycle: Unpaired Medical Report Generation via Cycle-Consistency
por: Hirsch, Elad, et al.
Publicado: (2024) -
SFMNet: Sparse Focal Modulation for 3D Object Detection
por: Shrout, Oren, et al.
Publicado: (2025) -
MME: Mixture of Mesh Experts with Random Walk Transformer Gating
por: Belder, Amir, et al.
Publicado: (2026)