Better STEP, a format and dataset for boundary representation
Fuente:
arXiv
Saved in:
| Main Authors: | Izadyar, Nafiseh, Madduri, Sai Chandra, Schneider, Teseo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLM-Guided Material Inference for 3D Point Clouds
by: Izadyar, Nafiseh, et al.
Published: (2025)
by: Izadyar, Nafiseh, et al.
Published: (2025)
STEP: Spatial Temporal Graph Convolutional Networks for Emotion Perception from Gaits
by: Bhattacharya, Uttaran, et al.
Published: (2019)
by: Bhattacharya, Uttaran, et al.
Published: (2019)
STEP: Simultaneous Tracking and Estimation of Pose for Animals and Humans
by: Verma, Shashikant, et al.
Published: (2025)
by: Verma, Shashikant, et al.
Published: (2025)
STEP3-VL-10B Technical Report
by: Huang, Ailin, et al.
Published: (2026)
by: Huang, Ailin, et al.
Published: (2026)
SODAWideNet++: Combining Attention and Convolutions for Salient Object Detection
by: Dulam, Rohit Venkata Sai, et al.
Published: (2024)
by: Dulam, Rohit Venkata Sai, et al.
Published: (2024)
Invertible generative models for inverse problems: mitigating representation error and dataset bias
by: Asim, Muhammad, et al.
Published: (2019)
by: Asim, Muhammad, et al.
Published: (2019)
Can Better Text Semantics in Prompt Tuning Improve VLM Generalization?
by: Kuchibhotla, Hari Chandana, et al.
Published: (2024)
by: Kuchibhotla, Hari Chandana, et al.
Published: (2024)
CNN-based explanation ensembling for dataset, representation and explanations evaluation
by: Hryniewska-Guzik, Weronika, et al.
Published: (2024)
by: Hryniewska-Guzik, Weronika, et al.
Published: (2024)
Where Do Tokens Go? Understanding Pruning Behaviors in STEP at High Resolutions
by: Szczepanski, Michal, et al.
Published: (2025)
by: Szczepanski, Michal, et al.
Published: (2025)
GlotOCR Bench: OCR Models Still Struggle Beyond a Handful of Unicode Scripts
by: Kargaran, Amir Hossein, et al.
Published: (2026)
by: Kargaran, Amir Hossein, et al.
Published: (2026)
ABC: Achieving Better Control of Multimodal Embeddings using VLMs
by: Schneider, Benjamin, et al.
Published: (2025)
by: Schneider, Benjamin, et al.
Published: (2025)
Mitigating representation bias caused by missing pixels in methane plume detection
by: Wąsala, Julia, et al.
Published: (2025)
by: Wąsala, Julia, et al.
Published: (2025)
STEP: Enhancing Video-LLMs' Compositional Reasoning by Spatio-Temporal Graph-guided Self-Training
by: Qiu, Haiyi, et al.
Published: (2024)
by: Qiu, Haiyi, et al.
Published: (2024)
Better Sampling, towards Better End-to-end Small Object Detection
by: Huang, Zile, et al.
Published: (2024)
by: Huang, Zile, et al.
Published: (2024)
Inverse problems with diffusion models: MAP estimation via mode-seeking loss
by: Gutha, Sai Bharath Chandra, et al.
Published: (2025)
by: Gutha, Sai Bharath Chandra, et al.
Published: (2025)
Inverse Problems with Diffusion Models: A MAP Estimation Perspective
by: Gutha, Sai Bharath Chandra, et al.
Published: (2024)
by: Gutha, Sai Bharath Chandra, et al.
Published: (2024)
Objective drives the consistency of representational similarity across datasets
by: Ciernik, Laure, et al.
Published: (2024)
by: Ciernik, Laure, et al.
Published: (2024)
Topology-First B-Rep Meshing
by: Zhou, YunFan, et al.
Published: (2026)
by: Zhou, YunFan, et al.
Published: (2026)
ConvFormer3D-TAP: Phase/Uncertainty-Aware Front-End Fusion for Cine CMR View Classification Pipelines
by: Nia, Nafiseh Ghaffar, et al.
Published: (2026)
by: Nia, Nafiseh Ghaffar, et al.
Published: (2026)
Better Tokens for Better 3D: Advancing Vision-Language Modeling in 3D Medical Imaging
by: Hamamci, Ibrahim Ethem, et al.
Published: (2025)
by: Hamamci, Ibrahim Ethem, et al.
Published: (2025)
BOGausS: Better Optimized Gaussian Splatting
by: Pateux, Stéphane, et al.
Published: (2025)
by: Pateux, Stéphane, et al.
Published: (2025)
CoTracker: It is Better to Track Together
by: Karaev, Nikita, et al.
Published: (2023)
by: Karaev, Nikita, et al.
Published: (2023)
Decomposition Betters Tracking Everything Everywhere
by: Li, Rui, et al.
Published: (2024)
by: Li, Rui, et al.
Published: (2024)
CDG-MAE: Learning Correspondences from Diffusion Generated Views
by: Belagali, Varun, et al.
Published: (2025)
by: Belagali, Varun, et al.
Published: (2025)
Mixup Helps Understanding Multimodal Video Better
by: Ma, Xiaoyu, et al.
Published: (2025)
by: Ma, Xiaoyu, et al.
Published: (2025)
Boosting the Local Invariance for Better Adversarial Transferability
by: Liu, Bohan, et al.
Published: (2025)
by: Liu, Bohan, et al.
Published: (2025)
A Simple and Better Baseline for Visual Grounding
by: Wang, Jingchao, et al.
Published: (2025)
by: Wang, Jingchao, et al.
Published: (2025)
Gen-SIS: Generative Self-augmentation Improves Self-supervised Learning
by: Belagali, Varun, et al.
Published: (2024)
by: Belagali, Varun, et al.
Published: (2024)
Diffusion Feedback Helps CLIP See Better
by: Wang, Wenxuan, et al.
Published: (2024)
by: Wang, Wenxuan, et al.
Published: (2024)
8-Calves Image dataset
by: Fang, Xuyang, et al.
Published: (2025)
by: Fang, Xuyang, et al.
Published: (2025)
Deep video representation learning: a survey
by: Ravanbakhsh, Elham, et al.
Published: (2024)
by: Ravanbakhsh, Elham, et al.
Published: (2024)
Better Coherence, Better Height: Fusing Physical Models and Deep Learning for Forest Height Estimation from Interferometric SAR Data
by: Mahesh, Ragini Bal, et al.
Published: (2025)
by: Mahesh, Ragini Bal, et al.
Published: (2025)
Rethinking FID: Towards a Better Evaluation Metric for Image Generation
by: Jayasumana, Sadeep, et al.
Published: (2023)
by: Jayasumana, Sadeep, et al.
Published: (2023)
Towards a Better Understanding of the Computer Vision Research Community in Africa
by: Omotayo, Abdul-Hakeem, et al.
Published: (2023)
by: Omotayo, Abdul-Hakeem, et al.
Published: (2023)
An evaluation of Deep Learning based stereo dense matching dataset shift from aerial images and a large scale stereo dataset
by: Wu, Teng, et al.
Published: (2024)
by: Wu, Teng, et al.
Published: (2024)
Disrupting Semantic and Abstract Features for Better Adversarial Transferability
by: Luo, Yuyang, et al.
Published: (2025)
by: Luo, Yuyang, et al.
Published: (2025)
Towards Better Optimization For Listwise Preference in Diffusion Models
by: Bai, Jiamu, et al.
Published: (2025)
by: Bai, Jiamu, et al.
Published: (2025)
Enhance-A-Video: Better Generated Video for Free
by: Luo, Yang, et al.
Published: (2025)
by: Luo, Yang, et al.
Published: (2025)
Anomize: Better Open Vocabulary Video Anomaly Detection
by: Li, Fei, et al.
Published: (2025)
by: Li, Fei, et al.
Published: (2025)
Toward A Better Understanding of Monocular Depth Evaluation
by: Wu, Siyang, et al.
Published: (2025)
by: Wu, Siyang, et al.
Published: (2025)
Similar Items
-
LLM-Guided Material Inference for 3D Point Clouds
by: Izadyar, Nafiseh, et al.
Published: (2025) -
STEP: Spatial Temporal Graph Convolutional Networks for Emotion Perception from Gaits
by: Bhattacharya, Uttaran, et al.
Published: (2019) -
STEP: Simultaneous Tracking and Estimation of Pose for Animals and Humans
by: Verma, Shashikant, et al.
Published: (2025) -
STEP3-VL-10B Technical Report
by: Huang, Ailin, et al.
Published: (2026) -
SODAWideNet++: Combining Attention and Convolutions for Salient Object Detection
by: Dulam, Rohit Venkata Sai, et al.
Published: (2024)