QuIIL at T3 challenge: Towards Automation in Life-Saving Intervention Procedures from First-Person View
Fuente:
arXiv
Saved in:
| Main Authors: | Vuong, Trinh T. L., Bui, Doanh C., Kwak, Jin Tae |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FALFormer: Feature-aware Landmarks self-attention for Whole-slide Image Classification
by: Bui, Doanh C., et al.
Published: (2024)
by: Bui, Doanh C., et al.
Published: (2024)
Welcome New Doctor: Continual Learning with Expert Consultation and Autoregressive Inference for Whole Slide Image Analysis
by: Bui, Doanh Cao, et al.
Published: (2025)
by: Bui, Doanh Cao, et al.
Published: (2025)
MECFormer: Multi-task Whole Slide Image Classification with Expert Consultation Network
by: Bui, Doanh C., et al.
Published: (2024)
by: Bui, Doanh C., et al.
Published: (2024)
ViDRiP-LLaVA: A Dataset and Benchmark for Diagnostic Reasoning from Pathology Videos
by: Vuong, Trinh T. L., et al.
Published: (2025)
by: Vuong, Trinh T. L., et al.
Published: (2025)
MoMA: Momentum Contrastive Learning with Multi-head Attention-based Knowledge Distillation for Histopathology Image Analysis
by: Vuong, Trinh Thi Le, et al.
Published: (2023)
by: Vuong, Trinh Thi Le, et al.
Published: (2023)
Towards a text-based quantitative and explainable histopathology image analysis
by: Nguyen, Anh Tien, et al.
Published: (2024)
by: Nguyen, Anh Tien, et al.
Published: (2024)
CLEAR: Cross-Transformers with Pre-trained Language Model is All you need for Person Attribute Recognition and Retrieval
by: Bui, Doanh C., et al.
Published: (2024)
by: Bui, Doanh C., et al.
Published: (2024)
HiGDA: Hierarchical Graph of Nodes to Learn Local-to-Global Topology for Semi-Supervised Domain Adaptation
by: Ngo, Ba Hung, et al.
Published: (2024)
by: Ngo, Ba Hung, et al.
Published: (2024)
USegMix: Unsupervised Segment Mix for Efficient Data Augmentation in Pathology Images
by: Wang, Jiamu, et al.
Published: (2025)
by: Wang, Jiamu, et al.
Published: (2025)
HEXST: Hexagonal Shifted-Window Transformer for Spatial Transcriptomics Gene Expression Prediction
by: Byeon, Keunho, et al.
Published: (2026)
by: Byeon, Keunho, et al.
Published: (2026)
NucleiMix: Realistic Data Augmentation for Nuclei Instance Segmentation
by: Wang, Jiamu, et al.
Published: (2024)
by: Wang, Jiamu, et al.
Published: (2024)
EgoTwin: Dreaming Body and View in First Person
by: Xiu, Jingqiao, et al.
Published: (2025)
by: Xiu, Jingqiao, et al.
Published: (2025)
Sequence-Based Identification of First-Person Camera Wearers in Third-Person Views
by: Zhao, Ziwei, et al.
Published: (2025)
by: Zhao, Ziwei, et al.
Published: (2025)
Is Tracking really more challenging in First Person Egocentric Vision?
by: Dunnhofer, Matteo, et al.
Published: (2025)
by: Dunnhofer, Matteo, et al.
Published: (2025)
Towards Comprehensive Scene Understanding: Integrating First and Third-Person Views for LVLMs
by: Lee, Insu, et al.
Published: (2025)
by: Lee, Insu, et al.
Published: (2025)
GPC: Generative and General Pathology Image Classifier
by: Nguyen, Anh Tien, et al.
Published: (2024)
by: Nguyen, Anh Tien, et al.
Published: (2024)
Benchmarking Pathology Foundation Models: Adaptation Strategies and Scenarios
by: Lee, Jeaung, et al.
Published: (2024)
by: Lee, Jeaung, et al.
Published: (2024)
Can Geometry Save Central Views for Sports Field Registration?
by: Magera, Floriane, et al.
Published: (2025)
by: Magera, Floriane, et al.
Published: (2025)
VLEER: Vision and Language Embeddings for Explainable Whole Slide Image Representation
by: Nguyen, Anh Tien, et al.
Published: (2025)
by: Nguyen, Anh Tien, et al.
Published: (2025)
CCPA: Long-term Person Re-Identification via Contrastive Clothing and Pose Augmentation
by: Nguyen, Vuong D., et al.
Published: (2024)
by: Nguyen, Vuong D., et al.
Published: (2024)
SignEye: Traffic Sign Interpretation from Vehicle First-Person View
by: Yang, Chuang, et al.
Published: (2024)
by: Yang, Chuang, et al.
Published: (2024)
Novel View Synthesis as Video Completion
by: Wu, Qi, et al.
Published: (2026)
by: Wu, Qi, et al.
Published: (2026)
QuRe: Query-Relevant Retrieval through Hard Negative Sampling in Composed Image Retrieval
by: Kwak, Jaehyun, et al.
Published: (2025)
by: Kwak, Jaehyun, et al.
Published: (2025)
MergeSlide: Continual Model Merging and Task-to-Class Prompt-Aligned Inference for Lifelong Learning on Whole Slide Images
by: Bui, Doanh C., et al.
Published: (2025)
by: Bui, Doanh C., et al.
Published: (2025)
EvHand-FPV: Efficient Event-Based 3D Hand Tracking from First-Person View
by: Xu, Zhen, et al.
Published: (2025)
by: Xu, Zhen, et al.
Published: (2025)
Patch Stitching Data Augmentation for Cancer Classification in Pathology Images
by: Wang, Jiamu, et al.
Published: (2025)
by: Wang, Jiamu, et al.
Published: (2025)
Hierarchical Classification for Improved Histopathology Image Analysis
by: Byeon, Keunho, et al.
Published: (2026)
by: Byeon, Keunho, et al.
Published: (2026)
AerialMegaDepth: Learning Aerial-Ground Reconstruction and View Synthesis
by: Vuong, Khiem, et al.
Published: (2025)
by: Vuong, Khiem, et al.
Published: (2025)
ZeroSlide: Is Zero-Shot Classification Adequate for Lifelong Learning in Whole-Slide Image Analysis in the Era of Pathology Vision-Language Foundation Models?
by: Bui, Doanh C., et al.
Published: (2025)
by: Bui, Doanh C., et al.
Published: (2025)
SRHand: Super-Resolving Hand Images and 3D Shapes via View/Pose-aware Neural Image Representations and Explicit 3D Meshes
by: Kim, Minje, et al.
Published: (2025)
by: Kim, Minje, et al.
Published: (2025)
Towards Robust Text-to-Image Person Retrieval: Multi-View Reformulation for Semantic Compensation
by: Yuan, Chao, et al.
Published: (2026)
by: Yuan, Chao, et al.
Published: (2026)
Q-Save: Towards Scoring and Attribution for Generated Video Evaluation
by: Wu, Xiele, et al.
Published: (2025)
by: Wu, Xiele, et al.
Published: (2025)
VI3DRM:Towards meticulous 3D Reconstruction from Sparse Views via Photo-Realistic Novel View Synthesis
by: Chen, Hao, et al.
Published: (2024)
by: Chen, Hao, et al.
Published: (2024)
Towards Unified 3D Hair Reconstruction from Single-View Portraits
by: Zheng, Yujian, et al.
Published: (2024)
by: Zheng, Yujian, et al.
Published: (2024)
Lifelong Whole Slide Image Analysis: Online Vision-Language Adaptation and Past-to-Present Gradient Distillation
by: Bui, Doanh C., et al.
Published: (2025)
by: Bui, Doanh C., et al.
Published: (2025)
PGDS: Pose-Guidance Deep Supervision for Mitigating Clothes-Changing in Person Re-Identification
by: Trinh, Quoc-Huy, et al.
Published: (2023)
by: Trinh, Quoc-Huy, et al.
Published: (2023)
Aligned Novel View Image and Geometry Synthesis via Cross-modal Attention Instillation
by: Kwak, Min-Seop, et al.
Published: (2025)
by: Kwak, Min-Seop, et al.
Published: (2025)
Controllable Human Image Generation with Personalized Multi-Garments
by: Choi, Yisol, et al.
Published: (2024)
by: Choi, Yisol, et al.
Published: (2024)
CAMEO: Correspondence-Attention Alignment for Multi-View Diffusion Models
by: Kwon, Minkyung, et al.
Published: (2025)
by: Kwon, Minkyung, et al.
Published: (2025)
Projected Representation Conditioning for High-fidelity Novel View Synthesis
by: Kwak, Min-Seop, et al.
Published: (2026)
by: Kwak, Min-Seop, et al.
Published: (2026)
Similar Items
-
FALFormer: Feature-aware Landmarks self-attention for Whole-slide Image Classification
by: Bui, Doanh C., et al.
Published: (2024) -
Welcome New Doctor: Continual Learning with Expert Consultation and Autoregressive Inference for Whole Slide Image Analysis
by: Bui, Doanh Cao, et al.
Published: (2025) -
MECFormer: Multi-task Whole Slide Image Classification with Expert Consultation Network
by: Bui, Doanh C., et al.
Published: (2024) -
ViDRiP-LLaVA: A Dataset and Benchmark for Diagnostic Reasoning from Pathology Videos
by: Vuong, Trinh T. L., et al.
Published: (2025) -
MoMA: Momentum Contrastive Learning with Multi-head Attention-based Knowledge Distillation for Histopathology Image Analysis
by: Vuong, Trinh Thi Le, et al.
Published: (2023)