QORT-Former: Query-optimized Real-time Transformer for Understanding Two Hands Manipulating Objects
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ismayilzada, Elkhan, Sayem, MD Khalequzzaman Chowdhury, Tiruneh, Yihalem Yimolal, Chowdhury, Mubarrat Tajoar, Boboev, Muhammadjon, Baek, Seungryul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HandVQA: Diagnosing and Improving Fine-Grained Spatial Reasoning about Hands in Vision-Language Models
von: Sayem, MD Khalequzzaman Chowdhury, et al.
Veröffentlicht: (2026)
von: Sayem, MD Khalequzzaman Chowdhury, et al.
Veröffentlicht: (2026)
THOM: Generating Physically Plausible Hand-Object Meshes From Text
von: Jeong, Uyoung, et al.
Veröffentlicht: (2026)
von: Jeong, Uyoung, et al.
Veröffentlicht: (2026)
SDDGR: Stable Diffusion-based Deep Generative Replay for Class Incremental Object Detection
von: Kim, Junsu, et al.
Veröffentlicht: (2024)
von: Kim, Junsu, et al.
Veröffentlicht: (2024)
PAD-Hand: Physics-Aware Diffusion for Hand Motion Recovery
von: Ismayilzada, Elkhan, et al.
Veröffentlicht: (2026)
von: Ismayilzada, Elkhan, et al.
Veröffentlicht: (2026)
GeoBotsVR: A Robotics Learning Game for Beginners with Hands-on Learning Simulation
von: Mubarrat, Syed T.
Veröffentlicht: (2024)
von: Mubarrat, Syed T.
Veröffentlicht: (2024)
Text2HOI: Text-guided 3D Motion Generation for Hand-Object Interaction
von: Cha, Junuk, et al.
Veröffentlicht: (2024)
von: Cha, Junuk, et al.
Veröffentlicht: (2024)
Robust Long-Form Bangla Speech Processing: Automatic Speech Recognition and Speaker Diarization
von: Chowdhury, MD. Sagor, et al.
Veröffentlicht: (2026)
von: Chowdhury, MD. Sagor, et al.
Veröffentlicht: (2026)
Can Synthetic Images Conquer Forgetting? Beyond Unexplored Doubts in Few-Shot Class-Incremental Learning
von: Kim, Junsu, et al.
Veröffentlicht: (2025)
von: Kim, Junsu, et al.
Veröffentlicht: (2025)
O'ZBEKISTON ISLOM SIVILIZATSIYASI MEROSINING ZAMONAVIY DAVLAT SIYOSATIDAGI O'RNI VA XALQARO MODELI
von: Xotamxonov, Muhammadjon
Veröffentlicht: (2026)
von: Xotamxonov, Muhammadjon
Veröffentlicht: (2026)
ПPSIXOLOGIK ADAPTATSIYA: XODIMLAR SAMARADORLIGINI OSHIRISHNING PSIXOLOGIK IMKONIYATLARI
von: Muhammadjon Abdukarimov
Veröffentlicht: (2025)
von: Muhammadjon Abdukarimov
Veröffentlicht: (2025)
VLM-PL: Advanced Pseudo Labeling Approach for Class Incremental Object Detection via Vision-Language Model
von: Kim, Junsu, et al.
Veröffentlicht: (2024)
von: Kim, Junsu, et al.
Veröffentlicht: (2024)
On the Monotonicity of Information Aging
von: Shisher, MD Kamran Chowdhury, et al.
Veröffentlicht: (2024)
von: Shisher, MD Kamran Chowdhury, et al.
Veröffentlicht: (2024)
Incompatibility of optimized protection of entanglement and teleportation fdelity in the presence of decoherence
von: Chowdhury, Priyanka
Veröffentlicht: (2022)
von: Chowdhury, Priyanka
Veröffentlicht: (2022)
BOBUR VA BOBURIYLAR DAVLATINING IQTISODIY ASOSLARI
von: Mahkamov Muhammadjon Dadajonovich
Veröffentlicht: (2025)
von: Mahkamov Muhammadjon Dadajonovich
Veröffentlicht: (2025)
AGRICULTURE COMMUNICATION AND ITS CONSEQUENCES
von: G'aniyev, Muhammadjon
Veröffentlicht: (2025)
von: G'aniyev, Muhammadjon
Veröffentlicht: (2025)
Beyond Synthetic Replays: Turning Diffusion Features into Few-Shot Class-Incremental Learning Knowledge
von: Kim, Junsu, et al.
Veröffentlicht: (2025)
von: Kim, Junsu, et al.
Veröffentlicht: (2025)
CNN+ViT+SigLiP: A Real-Time Framework for Smart Urban Mobility
von: Chowdhury, Koushik
Veröffentlicht: (2025)
von: Chowdhury, Koushik
Veröffentlicht: (2025)
Design and Application of Multimodal Large Language Model Based System for End to End Automation of Accident Dataset Generation
von: Chowdhury, MD Thamed Bin Zaman, et al.
Veröffentlicht: (2025)
von: Chowdhury, MD Thamed Bin Zaman, et al.
Veröffentlicht: (2025)
ALIGN: A Vision-Language Framework for High-Accuracy Accident Location Inference through Geo-Spatial Neural Reasoning
von: Chowdhury, MD Thamed Bin Zaman, et al.
Veröffentlicht: (2025)
von: Chowdhury, MD Thamed Bin Zaman, et al.
Veröffentlicht: (2025)
HOIST-Former: Hand-held Objects Identification, Segmentation, and Tracking in the Wild
von: Narasimhaswamy, Supreeth, et al.
Veröffentlicht: (2024)
von: Narasimhaswamy, Supreeth, et al.
Veröffentlicht: (2024)
XXI ASRDA TA'LIM VA TA'LIMNING KELAJAGI
von: Husanova Odinaxon Muhammadjon qizi
Veröffentlicht: (2025)
von: Husanova Odinaxon Muhammadjon qizi
Veröffentlicht: (2025)
REGIONAL ECONOMIC ZONES AND THEIR ROLE IN INDUSTRIAL DEVELOPMENT
von: Samadkulov Muhammadjon, et al.
Veröffentlicht: (2025)
von: Samadkulov Muhammadjon, et al.
Veröffentlicht: (2025)
THE ROLE OF MASS MEDIA IN THE FORMATION AND DISSEMINATION OF NEOLOGISMS
von: Mirmusayev, Mashrabjon, et al.
Veröffentlicht: (2026)
von: Mirmusayev, Mashrabjon, et al.
Veröffentlicht: (2026)
Understanding Investor Sentiment: Analyzing Its Influence on Stock and Cryptocurrency Markets During the Russia–Ukraine War
von: Emon Kalyan Chowdhury, et al.
Veröffentlicht: (2024)
von: Emon Kalyan Chowdhury, et al.
Veröffentlicht: (2024)
BoIR: Box-Supervised Instance Representation for Multi-Person Pose Estimation
von: Jeong, Uyoung, et al.
Veröffentlicht: (2023)
von: Jeong, Uyoung, et al.
Veröffentlicht: (2023)
ShapeGraFormer: GraFormer-Based Network for Hand-Object Reconstruction from a Single Depth Map
von: Aboukhadra, Ahmed Tawfik, et al.
Veröffentlicht: (2023)
von: Aboukhadra, Ahmed Tawfik, et al.
Veröffentlicht: (2023)
IsingFormer: Augmenting Parallel Tempering With Learned Proposals
von: Bunaiyan, Saleh, et al.
Veröffentlicht: (2025)
von: Bunaiyan, Saleh, et al.
Veröffentlicht: (2025)
B-RIGHT: Benchmark Re-evaluation for Integrity in Generalized Human-Object Interaction Testing
von: Jang, Yoojin, et al.
Veröffentlicht: (2025)
von: Jang, Yoojin, et al.
Veröffentlicht: (2025)
Exploiting Style Latent Flows for Generalizing Deepfake Video Detection
von: Choi, Jongwook, et al.
Veröffentlicht: (2024)
von: Choi, Jongwook, et al.
Veröffentlicht: (2024)
EmoTalkingGaussian: Continuous Emotion-conditioned Talking Head Synthesis
von: Cha, Junuk, et al.
Veröffentlicht: (2025)
von: Cha, Junuk, et al.
Veröffentlicht: (2025)
Benchmarks and Challenges in Pose Estimation for Egocentric Hand Interactions with Objects
von: Fan, Zicong, et al.
Veröffentlicht: (2024)
von: Fan, Zicong, et al.
Veröffentlicht: (2024)
Two-Stage Stochastic Optimal Power Flow for Microgrids With Uncertain Wildfire Effects
von: Chowdhury, Sifat, et al.
Veröffentlicht: (2024)
von: Chowdhury, Sifat, et al.
Veröffentlicht: (2024)
In-Hand Manipulation of Articulated Tools with Dexterous Robot Hands with Sim-to-Real Transfer
von: Atar, Soofiyan, et al.
Veröffentlicht: (2025)
von: Atar, Soofiyan, et al.
Veröffentlicht: (2025)
SpokenNativQA: Multilingual Everyday Spoken Queries for LLMs
von: Alam, Firoj, et al.
Veröffentlicht: (2025)
von: Alam, Firoj, et al.
Veröffentlicht: (2025)
Digital Library Research: Major Issues and Trends.
von: Chowdhury, Suddatta, et al.
Veröffentlicht: (1999)
von: Chowdhury, Suddatta, et al.
Veröffentlicht: (1999)
PoseBH: Prototypical Multi-Dataset Training Beyond Human Pose Estimation
von: Jeong, Uyoung, et al.
Veröffentlicht: (2025)
von: Jeong, Uyoung, et al.
Veröffentlicht: (2025)
MF-GCN: A Multi-Frequency Graph Convolutional Network for Tri-Modal Depression Detection Using Eye-Tracking, Facial, and Acoustic Features
von: Rahman, Sejuti, et al.
Veröffentlicht: (2025)
von: Rahman, Sejuti, et al.
Veröffentlicht: (2025)
Vibration-based Full State In-Hand Manipulation of Thin Objects
von: Binyamin, Oron, et al.
Veröffentlicht: (2024)
von: Binyamin, Oron, et al.
Veröffentlicht: (2024)
DyTact: Capturing Dynamic Contacts in Hand-Object Manipulation
von: Cong, Xiaoyan, et al.
Veröffentlicht: (2025)
von: Cong, Xiaoyan, et al.
Veröffentlicht: (2025)
A dual physics-informed neural network for topology optimization
von: Singh, Ajendra, et al.
Veröffentlicht: (2024)
von: Singh, Ajendra, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
HandVQA: Diagnosing and Improving Fine-Grained Spatial Reasoning about Hands in Vision-Language Models
von: Sayem, MD Khalequzzaman Chowdhury, et al.
Veröffentlicht: (2026) -
THOM: Generating Physically Plausible Hand-Object Meshes From Text
von: Jeong, Uyoung, et al.
Veröffentlicht: (2026) -
SDDGR: Stable Diffusion-based Deep Generative Replay for Class Incremental Object Detection
von: Kim, Junsu, et al.
Veröffentlicht: (2024) -
PAD-Hand: Physics-Aware Diffusion for Hand Motion Recovery
von: Ismayilzada, Elkhan, et al.
Veröffentlicht: (2026) -
GeoBotsVR: A Robotics Learning Game for Beginners with Hands-on Learning Simulation
von: Mubarrat, Syed T.
Veröffentlicht: (2024)