SituationalLLM: Proactive language models with scene awareness for dynamic, contextual task guidance
Fuente:
arXiv
Guardado en:
| Autores principales: | Khan, Muhammad Saif Ullah, Afzal, Muhammad Zeshan, Stricker, Didier |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Estimating Human Poses Across Datasets: A Unified Skeleton and Multi-Teacher Distillation Approach
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2024)
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2024)
Shape2.5D: A Dataset of Texture-less Surfaces for Depth and Normals Estimation
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2024)
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2024)
Enhanced Bank Check Security: Introducing a Novel Dataset and Transformer-Based Approach for Detection and Verification
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2024)
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2024)
CICA: Content-Injected Contrastive Alignment for Zero-Shot Document Image Classification
por: Sinha, Sankalp, et al.
Publicado: (2024)
por: Sinha, Sankalp, et al.
Publicado: (2024)
NURBGen: High-Fidelity Text-to-CAD Generation through LLM-Driven NURBS Modeling
por: Usama, Muhammad, et al.
Publicado: (2025)
por: Usama, Muhammad, et al.
Publicado: (2025)
Classroom-Inspired Multi-Mentor Distillation with Adaptive Learning Strategies
por: Sarode, Shalini, et al.
Publicado: (2024)
por: Sarode, Shalini, et al.
Publicado: (2024)
PoseAdapt: Sustainable Human Pose Estimation via Continual Learning Benchmarks and Toolkit
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2024)
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2024)
SIMSPINE: A Biomechanics-Aware Simulation Framework for 3D Spine Motion Annotation and Benchmarking
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2026)
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2026)
ReConText3D: Replay-based Continual Text-to-3D Generation
por: Khan, Muhammad Ahmed Ullah, et al.
Publicado: (2026)
por: Khan, Muhammad Ahmed Ullah, et al.
Publicado: (2026)
A Hybrid Approach for Document Layout Analysis in Document images
por: Shehzadi, Tahira, et al.
Publicado: (2024)
por: Shehzadi, Tahira, et al.
Publicado: (2024)
Human Pose Descriptions and Subject-Focused Attention for Improved Zero-Shot Transfer in Human-Centric Classification Tasks
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2024)
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2024)
Semi-Supervised Object Detection: A Survey on Progress from CNN to Transformer
por: Shehzadi, Tahira, et al.
Publicado: (2024)
por: Shehzadi, Tahira, et al.
Publicado: (2024)
Towards Unconstrained 2D Pose Estimation of the Human Spine
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2025)
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2025)
Towards End-to-End Semi-Supervised Table Detection with Semantic Aligned Matching Transformer
por: Shehzadi, Tahira, et al.
Publicado: (2024)
por: Shehzadi, Tahira, et al.
Publicado: (2024)
End-to-End Semi-Supervised approach with Modulated Object Queries for Table Detection in Documents
por: Ehsan, Iqraa, et al.
Publicado: (2024)
por: Ehsan, Iqraa, et al.
Publicado: (2024)
SemAttNet: Towards Attention-based Semantic Aware Guided Depth Completion
por: Nazir, Danish, et al.
Publicado: (2022)
por: Nazir, Danish, et al.
Publicado: (2022)
Sparse Semi-DETR: Sparse Learnable Queries for Semi-Supervised Object Detection
por: Shehzadi, Tahira, et al.
Publicado: (2024)
por: Shehzadi, Tahira, et al.
Publicado: (2024)
VVitCutLER: Towards Unsupervised Object Detection and Segmentation in Videos
por: Lu, Zhijing, et al.
Publicado: (2026)
por: Lu, Zhijing, et al.
Publicado: (2026)
Beyond Boxes: Mask-Guided Spatio-Temporal Feature Aggregation for Video Object Detection
por: Hashmi, Khurram Azeem, et al.
Publicado: (2024)
por: Hashmi, Khurram Azeem, et al.
Publicado: (2024)
UnSupDLA: Towards Unsupervised Document Layout Analysis
por: Sheikh, Talha Uddin, et al.
Publicado: (2024)
por: Sheikh, Talha Uddin, et al.
Publicado: (2024)
DreamCAD: Scaling Multi-modal CAD Generation using Differentiable Parametric Surfaces
por: Khan, Mohammad Sadil, et al.
Publicado: (2026)
por: Khan, Mohammad Sadil, et al.
Publicado: (2026)
Text2CAD: Generating Sequential CAD Models from Beginner-to-Expert Level Text Prompts
por: Khan, Mohammad Sadil, et al.
Publicado: (2024)
por: Khan, Mohammad Sadil, et al.
Publicado: (2024)
MARVEL-40M+: Multi-Level Visual Elaboration for High-Fidelity Text-to-3D Content Creation
por: Sinha, Sankalp, et al.
Publicado: (2024)
por: Sinha, Sankalp, et al.
Publicado: (2024)
Amortized Inverse Kinematics via Graph Attention for Real-Time Human Avatar Animation
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2026)
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2026)
Jointly Learning Spatial, Angular, and Temporal Information for Enhanced Lane Detection
por: Alam, Muhammad Zeshan
Publicado: (2024)
por: Alam, Muhammad Zeshan
Publicado: (2024)
TreeNet: Layered Decision Ensembles
por: Khan, Zeshan
Publicado: (2025)
por: Khan, Zeshan
Publicado: (2025)
A Light Weight Multi-Features-View Convolution Neural Network For Plant Disease Identification
por: Khan, Muhammad Kaleem Ullah
Publicado: (2026)
por: Khan, Muhammad Kaleem Ullah
Publicado: (2026)
Abnormalities and Disease Detection in Gastro-Intestinal Tract Images
por: Khan, Zeshan, et al.
Publicado: (2026)
por: Khan, Zeshan, et al.
Publicado: (2026)
Beyond Averages: Open-Vocabulary 3D Scene Understanding with Gaussian Splatting and Bag of Embeddings
por: Arafa, Abdalla, et al.
Publicado: (2025)
por: Arafa, Abdalla, et al.
Publicado: (2025)
Object-Centric 2D Gaussian Splatting: Background Removal and Occlusion-Aware Pruning for Compact Object Models
por: Rogge, Marcel, et al.
Publicado: (2025)
por: Rogge, Marcel, et al.
Publicado: (2025)
Efficient scene text image super-resolution with semantic guidance
por: TomyEnrique, LeoWu, et al.
Publicado: (2024)
por: TomyEnrique, LeoWu, et al.
Publicado: (2024)
Consistent text-to-image generation via scene de-contextualization
por: Tang, Song, et al.
Publicado: (2025)
por: Tang, Song, et al.
Publicado: (2025)
Trade-off Between Spatial and Angular Resolution in Facial Recognition
por: Alam, Muhammad Zeshan, et al.
Publicado: (2024)
por: Alam, Muhammad Zeshan, et al.
Publicado: (2024)
Saturation-Aware Space-Variant Blind Image Deblurring
por: Alam, Muhammad Z., et al.
Publicado: (2026)
por: Alam, Muhammad Z., et al.
Publicado: (2026)
Light Field Spatial Resolution Enhancement Framework
por: Shabbir, Javeria, et al.
Publicado: (2024)
por: Shabbir, Javeria, et al.
Publicado: (2024)
RMS-FlowNet++: Efficient and Robust Multi-Scale Scene Flow Estimation for Large-Scale Point Clouds
por: Battrawy, Ramy, et al.
Publicado: (2024)
por: Battrawy, Ramy, et al.
Publicado: (2024)
ShapeAug: Occlusion Augmentation for Event Camera Data
por: Bendig, Katharina, et al.
Publicado: (2024)
por: Bendig, Katharina, et al.
Publicado: (2024)
ShapeAug++: More Realistic Shape Augmentation for Event Data
por: Bendig, Katharina, et al.
Publicado: (2024)
por: Bendig, Katharina, et al.
Publicado: (2024)
G3FA: Geometry-guided GAN for Face Animation
por: Javanmardi, Alireza, et al.
Publicado: (2024)
por: Javanmardi, Alireza, et al.
Publicado: (2024)
EgoFlowNet: Non-Rigid Scene Flow from Point Clouds with Ego-Motion Support
por: Battrawy, Ramy, et al.
Publicado: (2024)
por: Battrawy, Ramy, et al.
Publicado: (2024)
Ejemplares similares
-
Estimating Human Poses Across Datasets: A Unified Skeleton and Multi-Teacher Distillation Approach
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2024) -
Shape2.5D: A Dataset of Texture-less Surfaces for Depth and Normals Estimation
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2024) -
Enhanced Bank Check Security: Introducing a Novel Dataset and Transformer-Based Approach for Detection and Verification
por: Khan, Muhammad Saif Ullah, et al.
Publicado: (2024) -
CICA: Content-Injected Contrastive Alignment for Zero-Shot Document Image Classification
por: Sinha, Sankalp, et al.
Publicado: (2024) -
NURBGen: High-Fidelity Text-to-CAD Generation through LLM-Driven NURBS Modeling
por: Usama, Muhammad, et al.
Publicado: (2025)