Positional Encoding Field
Fuente:
arXiv
Salvato in:
| Autori principali: | Bai, Yunpeng, Li, Haoxiang, Huang, Qixing |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
FiffDepth: Feed-forward Transformation of Diffusion-Based Generators for Detailed Depth Estimation
di: Bai, Yunpeng, et al.
Pubblicazione: (2024)
di: Bai, Yunpeng, et al.
Pubblicazione: (2024)
GeoVideo: Introducing Geometric Regularization into Video Generation Model
di: Bai, Yunpeng, et al.
Pubblicazione: (2025)
di: Bai, Yunpeng, et al.
Pubblicazione: (2025)
WorldReel: 4D Video Generation with Consistent Geometry and Motion Modeling
di: Fang, Shaoheng, et al.
Pubblicazione: (2025)
di: Fang, Shaoheng, et al.
Pubblicazione: (2025)
Revisiting Multimodal Positional Encoding in Vision-Language Models
di: Huang, Jie, et al.
Pubblicazione: (2025)
di: Huang, Jie, et al.
Pubblicazione: (2025)
Learning Convex Decomposition via Feature Fields
di: Yang, Yuezhi, et al.
Pubblicazione: (2026)
di: Yang, Yuezhi, et al.
Pubblicazione: (2026)
Information-Regularized Constrained Inversion for Stable Avatar Editing from Sparse Supervision
di: Liang, Zhenxiao, et al.
Pubblicazione: (2026)
di: Liang, Zhenxiao, et al.
Pubblicazione: (2026)
UGG: Unified Generative Grasping
di: Lu, Jiaxin, et al.
Pubblicazione: (2023)
di: Lu, Jiaxin, et al.
Pubblicazione: (2023)
CoFie: Learning Compact Neural Surface Representations with Coordinate Fields
di: Jiang, Hanwen, et al.
Pubblicazione: (2024)
di: Jiang, Hanwen, et al.
Pubblicazione: (2024)
Jigsaw++: Imagining Complete Shape Priors for Object Reassembly
di: Lu, Jiaxin, et al.
Pubblicazione: (2024)
di: Lu, Jiaxin, et al.
Pubblicazione: (2024)
Real3D: Scaling Up Large Reconstruction Models with Real-World Images
di: Jiang, Hanwen, et al.
Pubblicazione: (2024)
di: Jiang, Hanwen, et al.
Pubblicazione: (2024)
An Investigation on The Position Encoding in Vision-Based Dynamics Prediction
di: Zhu, Jiageng, et al.
Pubblicazione: (2024)
di: Zhu, Jiageng, et al.
Pubblicazione: (2024)
Unified Camera Positional Encoding for Controlled Video Generation
di: Zhang, Cheng, et al.
Pubblicazione: (2025)
di: Zhang, Cheng, et al.
Pubblicazione: (2025)
Cameras as Relative Positional Encoding
di: Li, Ruilong, et al.
Pubblicazione: (2025)
di: Li, Ruilong, et al.
Pubblicazione: (2025)
LEDiT: Your Length-Extrapolatable Diffusion Transformer without Positional Encoding
di: Zhang, Shen, et al.
Pubblicazione: (2025)
di: Zhang, Shen, et al.
Pubblicazione: (2025)
Head-wise Adaptive Rotary Positional Encoding for Fine-Grained Image Generation
di: Li, Jiaye, et al.
Pubblicazione: (2025)
di: Li, Jiaye, et al.
Pubblicazione: (2025)
Multi-View Representation is What You Need for Point-Cloud Pre-Training
di: Yan, Siming, et al.
Pubblicazione: (2023)
di: Yan, Siming, et al.
Pubblicazione: (2023)
PPLNs: Parametric Piecewise Linear Networks for Event-Based Temporal Modeling and Beyond
di: Song, Chen, et al.
Pubblicazione: (2024)
di: Song, Chen, et al.
Pubblicazione: (2024)
KeyPoint Relative Position Encoding for Face Recognition
di: Kim, Minchul, et al.
Pubblicazione: (2024)
di: Kim, Minchul, et al.
Pubblicazione: (2024)
Seamless Human Motion Composition with Blended Positional Encodings
di: Barquero, German, et al.
Pubblicazione: (2024)
di: Barquero, German, et al.
Pubblicazione: (2024)
Multi-View Large Reconstruction Model via Geometry-Aware Positional Encoding and Attention
di: Li, Mengfei, et al.
Pubblicazione: (2024)
di: Li, Mengfei, et al.
Pubblicazione: (2024)
GenCorres: Consistent Shape Matching via Coupled Implicit-Explicit Shape Generative Models
di: Yang, Haitao, et al.
Pubblicazione: (2023)
di: Yang, Haitao, et al.
Pubblicazione: (2023)
ViGoR: Improving Visual Grounding of Large Vision Language Models with Fine-Grained Reward Modeling
di: Yan, Siming, et al.
Pubblicazione: (2024)
di: Yan, Siming, et al.
Pubblicazione: (2024)
LiteGE: Lightweight Geodesic Embedding for Efficient Geodesics Computation and Non-Isometric Shape Correspondence
di: Adikusuma, Yohanes Yudhi, et al.
Pubblicazione: (2025)
di: Adikusuma, Yohanes Yudhi, et al.
Pubblicazione: (2025)
Boosting Resolution Generalization of Diffusion Transformers with Randomized Positional Encodings
di: Hou, Liang, et al.
Pubblicazione: (2025)
di: Hou, Liang, et al.
Pubblicazione: (2025)
Quantum Visual Fields with Neural Amplitude Encoding
di: Wang, Shuteng, et al.
Pubblicazione: (2025)
di: Wang, Shuteng, et al.
Pubblicazione: (2025)
InfoGCN++: Learning Representation by Predicting the Future for Online Human Skeleton-based Action Recognition
di: Chi, Seunggeun, et al.
Pubblicazione: (2023)
di: Chi, Seunggeun, et al.
Pubblicazione: (2023)
OMEGA: Optimized Multimodal Position Encoding Index Derivation with Global Adaptive Scaling for Vision-Language Models
di: Huang, Ruoxiang, et al.
Pubblicazione: (2025)
di: Huang, Ruoxiang, et al.
Pubblicazione: (2025)
Focus on BEV: Self-calibrated Cycle View Transformation for Monocular Birds-Eye-View Segmentation
di: Zhao, Jiawei, et al.
Pubblicazione: (2024)
di: Zhao, Jiawei, et al.
Pubblicazione: (2024)
PuzzleBoard: A New Camera Calibration Pattern with Position Encoding
di: Stelldinger, Peer, et al.
Pubblicazione: (2024)
di: Stelldinger, Peer, et al.
Pubblicazione: (2024)
Beyond Sequential Distance: Inter-Modal Distance Invariant Position Encoding
di: Chen, Lin, et al.
Pubblicazione: (2026)
di: Chen, Lin, et al.
Pubblicazione: (2026)
Weierstrass Positional Encoding for Vision Transformers
di: Xin, Zhihang, et al.
Pubblicazione: (2026)
di: Xin, Zhihang, et al.
Pubblicazione: (2026)
Reconstructing Humans with a Biomechanically Accurate Skeleton
di: Xia, Yan, et al.
Pubblicazione: (2025)
di: Xia, Yan, et al.
Pubblicazione: (2025)
TutteNet: Injective 3D Deformations by Composition of 2D Mesh Deformations
di: Sun, Bo, et al.
Pubblicazione: (2024)
di: Sun, Bo, et al.
Pubblicazione: (2024)
OmniGlue: Generalizable Feature Matching with Foundation Model Guidance
di: Jiang, Hanwen, et al.
Pubblicazione: (2024)
di: Jiang, Hanwen, et al.
Pubblicazione: (2024)
UCM: Unifying Camera Control and Memory with Time-aware Positional Encoding Warping for World Models
di: Xu, Tianxing, et al.
Pubblicazione: (2026)
di: Xu, Tianxing, et al.
Pubblicazione: (2026)
HUMOTO: A 4D Dataset of Mocap Human Object Interactions
di: Lu, Jiaxin, et al.
Pubblicazione: (2025)
di: Lu, Jiaxin, et al.
Pubblicazione: (2025)
Occlusion Handling in 3D Human Pose Estimation with Perturbed Positional Encoding
di: Azizi, Niloofar, et al.
Pubblicazione: (2024)
di: Azizi, Niloofar, et al.
Pubblicazione: (2024)
A 2D Semantic-Aware Position Encoding for Vision Transformers
di: Chen, Xi, et al.
Pubblicazione: (2025)
di: Chen, Xi, et al.
Pubblicazione: (2025)
Seeing the Unseen: Mask-Driven Positional Encoding and Strip-Convolution Context Modeling for Cross-View Object Geo-Localization
di: Hu, Shuhan, et al.
Pubblicazione: (2025)
di: Hu, Shuhan, et al.
Pubblicazione: (2025)
BrepLLM: Native Boundary Representation Understanding with Large Language Models
di: Deng, Liyuan, et al.
Pubblicazione: (2025)
di: Deng, Liyuan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
FiffDepth: Feed-forward Transformation of Diffusion-Based Generators for Detailed Depth Estimation
di: Bai, Yunpeng, et al.
Pubblicazione: (2024) -
GeoVideo: Introducing Geometric Regularization into Video Generation Model
di: Bai, Yunpeng, et al.
Pubblicazione: (2025) -
WorldReel: 4D Video Generation with Consistent Geometry and Motion Modeling
di: Fang, Shaoheng, et al.
Pubblicazione: (2025) -
Revisiting Multimodal Positional Encoding in Vision-Language Models
di: Huang, Jie, et al.
Pubblicazione: (2025) -
Learning Convex Decomposition via Feature Fields
di: Yang, Yuezhi, et al.
Pubblicazione: (2026)