ProtoSnap: Prototype Alignment for Cuneiform Signs
Fuente:
arXiv
Saved in:
| Main Authors: | Mikulinsky, Rachel, Alper, Morris, Gordin, Shai, Jiménez, Enrique, Cohen, Yoram, Averbuch-Elor, Hadar |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Kiki or Bouba? Sound Symbolism in Vision-and-Language Models
by: Alper, Morris, et al.
Published: (2023)
by: Alper, Morris, et al.
Published: (2023)
WAFFLE: Multimodal Floorplan Understanding in the Wild
by: Ganon, Keren, et al.
Published: (2024)
by: Ganon, Keren, et al.
Published: (2024)
Dynamic Scene Understanding from Vision-Language Representations
by: Pruss, Shahaf, et al.
Published: (2025)
by: Pruss, Shahaf, et al.
Published: (2025)
ICC: Quantifying Image Caption Concreteness for Multimodal Dataset Curation
by: Yanuka, Moran, et al.
Published: (2024)
by: Yanuka, Moran, et al.
Published: (2024)
Emergent Visual-Semantic Hierarchies in Image-Text Representations
by: Alper, Morris, et al.
Published: (2024)
by: Alper, Morris, et al.
Published: (2024)
Supercharging Floorplan Localization with Semantic Rays
by: Grader, Yuval, et al.
Published: (2025)
by: Grader, Yuval, et al.
Published: (2025)
ReNoise: Real Image Inversion Through Iterative Noising
by: Garibi, Daniel, et al.
Published: (2024)
by: Garibi, Daniel, et al.
Published: (2024)
WildCAT3D: Appearance-Aware Multi-View Diffusion in the Wild
by: Alper, Morris, et al.
Published: (2025)
by: Alper, Morris, et al.
Published: (2025)
Raster2Seq: Polygon Sequence Generation for Floorplan Reconstruction
by: Phung, Hao, et al.
Published: (2026)
by: Phung, Hao, et al.
Published: (2026)
Shaping History: Advanced Machine Learning Techniques for the Analysis and Dating of Cuneiform Tablets over Three Millennia
by: Kapon, Danielle, et al.
Published: (2024)
by: Kapon, Danielle, et al.
Published: (2024)
Lang3D-XL: Language Embedded 3D Gaussians for Large-scale Scenes
by: Krakovsky, Shai, et al.
Published: (2025)
by: Krakovsky, Shai, et al.
Published: (2025)
Scene Grounding In the Wild
by: Cohen, Tamir, et al.
Published: (2026)
by: Cohen, Tamir, et al.
Published: (2026)
Mitigating Open-Vocabulary Caption Hallucinations
by: Ben-Kish, Assaf, et al.
Published: (2023)
by: Ben-Kish, Assaf, et al.
Published: (2023)
HaLo-NeRF: Learning Geometry-Guided Semantics for Exploring Unconstrained Photo Collections
by: Dudai, Chen, et al.
Published: (2024)
by: Dudai, Chen, et al.
Published: (2024)
InstanceGen: Image Generation with Instance-level Instructions
by: Sella, Etai, et al.
Published: (2025)
by: Sella, Etai, et al.
Published: (2025)
Composing People Together: Iterative Pose-Image Generation for Multi-Person Interaction Scenes
by: Peng, Wenxuan, et al.
Published: (2026)
by: Peng, Wenxuan, et al.
Published: (2026)
Let it Snow! Animating 3D Gaussian Scenes with Dynamic Weather Effects via Physics-Guided Score Distillation
by: Fiebelman, Gal, et al.
Published: (2025)
by: Fiebelman, Gal, et al.
Published: (2025)
A Joint Study of Phrase Grounding and Task Performance in Vision and Language Models
by: Kojima, Noriyuki, et al.
Published: (2023)
by: Kojima, Noriyuki, et al.
Published: (2023)
4-LEGS: 4D Language Embedded Gaussian Splatting
by: Fiebelman, Gal, et al.
Published: (2024)
by: Fiebelman, Gal, et al.
Published: (2024)
Extreme Rotation Estimation in the Wild
by: Bezalel, Hana, et al.
Published: (2024)
by: Bezalel, Hana, et al.
Published: (2024)
Spice-E : Structural Priors in 3D Diffusion using Cross-Entity Attention
by: Sella, Etai, et al.
Published: (2023)
by: Sella, Etai, et al.
Published: (2023)
Blended Point Cloud Diffusion for Localized Text-guided Shape Editing
by: Sella, Etai, et al.
Published: (2025)
by: Sella, Etai, et al.
Published: (2025)
ProtoTTA: Prototype-Guided Test-Time Adaptation
by: Abootorabi, Mohammad Mahdi, et al.
Published: (2026)
by: Abootorabi, Mohammad Mahdi, et al.
Published: (2026)
Color Bind: Exploring Color Perception in Text-to-Image Models
by: Shomer-Chai, Shay, et al.
Published: (2025)
by: Shomer-Chai, Shay, et al.
Published: (2025)
Thinking in Blender: Staged Executable Inverse Graphics with Vision-Language Models
by: He, Guangzhao, et al.
Published: (2026)
by: He, Guangzhao, et al.
Published: (2026)
ProtoP-OD: Explainable Object Detection with Prototypical Parts
by: Rath-Manakidis, Pavlos, et al.
Published: (2024)
by: Rath-Manakidis, Pavlos, et al.
Published: (2024)
Emergent Extreme-View Geometry in 3D Foundation Models
by: Zhang, Yiwen, et al.
Published: (2025)
by: Zhang, Yiwen, et al.
Published: (2025)
Long-tail Internet photo reconstruction
by: Li, Yuan, et al.
Published: (2026)
by: Li, Yuan, et al.
Published: (2026)
MeshOn: Intersection-Free Mesh-to-Mesh Composition
by: Kim, Hyunwoo, et al.
Published: (2026)
by: Kim, Hyunwoo, et al.
Published: (2026)
Systole-Conditioned Generative Cardiac Motion
by: Zuler, Shahar, et al.
Published: (2025)
by: Zuler, Shahar, et al.
Published: (2025)
ProtoAda: Prototype-Guided Adaptive Adapter Expansion and Geometric Consolidation for Multimodal Continual Instruction Tuning
by: Shi, Yu-Cheng, et al.
Published: (2026)
by: Shi, Yu-Cheng, et al.
Published: (2026)
Deformable ProtoPNet: An Interpretable Image Classifier Using Deformable Prototypes
by: Donnelly, Jon, et al.
Published: (2021)
by: Donnelly, Jon, et al.
Published: (2021)
ProtoVQA: An Adaptable Prototypical Framework for Explainable Fine-Grained Visual Question Answering
by: Diao, Xingjian, et al.
Published: (2025)
by: Diao, Xingjian, et al.
Published: (2025)
ProtoEFNet: Dynamic Prototype Learning for Inherently Interpretable Ejection Fraction Estimation in Echocardiography
by: Ghamary, Yeganeh, et al.
Published: (2025)
by: Ghamary, Yeganeh, et al.
Published: (2025)
ProtoMedX: Towards Explainable Multi-Modal Prototype Learning for Bone Health Classification
by: Pellicer, Alvaro Lopez, et al.
Published: (2025)
by: Pellicer, Alvaro Lopez, et al.
Published: (2025)
Closing the Alignment-Maturity Gap in Federated Prototype Learning
by: Casado-Diez, Mario, et al.
Published: (2026)
by: Casado-Diez, Mario, et al.
Published: (2026)
ProtoCLIP: Prototype-Aligned Latent Refinement for Robust Zero-Shot Chest X-Ray Classification
by: Kittler, Florian, et al.
Published: (2026)
by: Kittler, Florian, et al.
Published: (2026)
Segmentation by Factorization: Unsupervised Semantic Segmentation for Pathology by Factorizing Foundation Model Features
by: Gildenblat, Jacob, et al.
Published: (2024)
by: Gildenblat, Jacob, et al.
Published: (2024)
Visual Diffusion Models are Geometric Solvers
by: Goren, Nir, et al.
Published: (2025)
by: Goren, Nir, et al.
Published: (2025)
ProtoS-ViT: Visual foundation models for sparse self-explainable classifications
by: Turbé, Hugues, et al.
Published: (2024)
by: Turbé, Hugues, et al.
Published: (2024)
Similar Items
-
Kiki or Bouba? Sound Symbolism in Vision-and-Language Models
by: Alper, Morris, et al.
Published: (2023) -
WAFFLE: Multimodal Floorplan Understanding in the Wild
by: Ganon, Keren, et al.
Published: (2024) -
Dynamic Scene Understanding from Vision-Language Representations
by: Pruss, Shahaf, et al.
Published: (2025) -
ICC: Quantifying Image Caption Concreteness for Multimodal Dataset Curation
by: Yanuka, Moran, et al.
Published: (2024) -
Emergent Visual-Semantic Hierarchies in Image-Text Representations
by: Alper, Morris, et al.
Published: (2024)