Optimizing Multitask Industrial Processes with Predictive Action Guidance
Fuente:
arXiv
Saved in:
| Main Authors: | Mehta, Naval Kishore, Arvind, Prasad, Shyam Sunder, Saurav, Sumeet, Singh, Sanjay |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Multimodal Dataset for Enhancing Industrial Task Monitoring and Engagement Prediction
by: Mehta, Naval Kishore, et al.
Published: (2025)
by: Mehta, Naval Kishore, et al.
Published: (2025)
Gaze-Vector Estimation in the Dark with Temporally Encoded Event-driven Neural Networks
by: Banerjee, Abeer, et al.
Published: (2024)
by: Banerjee, Abeer, et al.
Published: (2024)
Dual Guidance Semi-Supervised Action Detection
by: Singh, Ankit, et al.
Published: (2025)
by: Singh, Ankit, et al.
Published: (2025)
fine-CLIP: Enhancing Zero-Shot Fine-Grained Surgical Action Recognition with Vision-Language Models
by: Sharma, Saurav, et al.
Published: (2025)
by: Sharma, Saurav, et al.
Published: (2025)
Uncertainty-Guided Appearance-Motion Association Network for Out-of-Distribution Action Detection
by: Fang, Xiang, et al.
Published: (2024)
by: Fang, Xiang, et al.
Published: (2024)
Multi-Level LVLM Guidance for Untrimmed Video Action Recognition
by: Peng, Liyang, et al.
Published: (2025)
by: Peng, Liyang, et al.
Published: (2025)
Zero-Shot Temporal Action Localization Through Textual Guidance
by: Liberatori, Benedetta, et al.
Published: (2026)
by: Liberatori, Benedetta, et al.
Published: (2026)
Open-Vocabulary Temporal Action Localization using Multimodal Guidance
by: Gupta, Akshita, et al.
Published: (2024)
by: Gupta, Akshita, et al.
Published: (2024)
Action Recognition based Industrial Safety Violation Detection
by: Reddy, Surya N, et al.
Published: (2024)
by: Reddy, Surya N, et al.
Published: (2024)
State-Change Learning for Prediction of Future Events in Endoscopic Videos
by: Sharma, Saurav, et al.
Published: (2025)
by: Sharma, Saurav, et al.
Published: (2025)
Cross-Task Affinity Learning for Multitask Dense Scene Predictions
by: Sinodinos, Dimitrios, et al.
Published: (2024)
by: Sinodinos, Dimitrios, et al.
Published: (2024)
Multitasking Embedding for Embryo Blastocyst Grading Prediction (MEmEBG)
by: Angabini, Nahid Khoshk, et al.
Published: (2026)
by: Angabini, Nahid Khoshk, et al.
Published: (2026)
NutritionVerse-Direct: Exploring Deep Neural Networks for Multitask Nutrition Prediction from Food Images
by: Keller, Matthew, et al.
Published: (2024)
by: Keller, Matthew, et al.
Published: (2024)
Egocentric Action-aware Inertial Localization in Point Clouds with Vision-Language Guidance
by: Zhang, Mingfang, et al.
Published: (2025)
by: Zhang, Mingfang, et al.
Published: (2025)
Frequency Guidance Matters: Skeletal Action Recognition by Frequency-Aware Mixed Transformer
by: Wu, Wenhan, et al.
Published: (2024)
by: Wu, Wenhan, et al.
Published: (2024)
Efficient Multitask Dense Predictor via Binarization
by: Shang, Yuzhang, et al.
Published: (2024)
by: Shang, Yuzhang, et al.
Published: (2024)
Group Diffusion Transformers are Unsupervised Multitask Learners
by: Huang, Lianghua, et al.
Published: (2024)
by: Huang, Lianghua, et al.
Published: (2024)
BeLLA: End-to-End Birds Eye View Large Language Assistant for Autonomous Driving
by: Mohan, Karthik, et al.
Published: (2025)
by: Mohan, Karthik, et al.
Published: (2025)
World Guidance: World Modeling in Condition Space for Action Generation
by: Su, Yue, et al.
Published: (2026)
by: Su, Yue, et al.
Published: (2026)
HQ-JEPA: Hybrid Quantum Joint-Embedding Predictive Architecture for Cross-Modal Remote Sensing Representation Learning
by: Hossain, Md Aminur, et al.
Published: (2026)
by: Hossain, Md Aminur, et al.
Published: (2026)
IPAD: Industrial Process Anomaly Detection Dataset
by: Liu, Jinfan, et al.
Published: (2024)
by: Liu, Jinfan, et al.
Published: (2024)
CrunchLLM: Multitask LLMs for Structured Business Reasoning and Outcome Prediction
by: Sadia, Rabeya Tus, et al.
Published: (2025)
by: Sadia, Rabeya Tus, et al.
Published: (2025)
Automatic Discovery and Assessment of Interpretable Systematic Errors in Semantic Segmentation
by: Singh, Jaisidh, et al.
Published: (2024)
by: Singh, Jaisidh, et al.
Published: (2024)
Learning Streaming Video Representation via Multitask Training
by: Yan, Yibin, et al.
Published: (2025)
by: Yan, Yibin, et al.
Published: (2025)
Efficient Inter-Task Attention for Multitask Transformer Models
by: Bohn, Christian, et al.
Published: (2025)
by: Bohn, Christian, et al.
Published: (2025)
Factored Classifier-Free Guidance
by: Xia, Tian, et al.
Published: (2025)
by: Xia, Tian, et al.
Published: (2025)
ImPoster: Text and Frequency Guidance for Subject Driven Action Personalization using Diffusion Models
by: Kothandaraman, Divya, et al.
Published: (2024)
by: Kothandaraman, Divya, et al.
Published: (2024)
EAGLE: Expert-Augmented Attention Guidance for Tuning-Free Industrial Anomaly Detection in Multimodal Large Language Models
by: Peng, Xiaomeng, et al.
Published: (2026)
by: Peng, Xiaomeng, et al.
Published: (2026)
Cross-Domain Identity Representation for Skull to Face Matching with Benchmark DataSet
by: Prasad, Ravi Shankar, et al.
Published: (2025)
by: Prasad, Ravi Shankar, et al.
Published: (2025)
FCR: Investigating Generative AI models for Forensic Craniofacial Reconstruction
by: Prasad, Ravi Shankar, et al.
Published: (2025)
by: Prasad, Ravi Shankar, et al.
Published: (2025)
SPOT-Face: Forensic Face Identification using Attention Guided Optimal Transport
by: Prasad, Ravi Shankar, et al.
Published: (2026)
by: Prasad, Ravi Shankar, et al.
Published: (2026)
Can Multi-Modal LLMs Provide Live Step-by-Step Task Guidance?
by: Bhattacharyya, Apratim, et al.
Published: (2025)
by: Bhattacharyya, Apratim, et al.
Published: (2025)
TextOCVP: Object-Centric Video Prediction with Language Guidance
by: Villar-Corrales, Angel, et al.
Published: (2025)
by: Villar-Corrales, Angel, et al.
Published: (2025)
IAP: Invisible Adversarial Patch Attack through Perceptibility-Aware Localization and Perturbation Optimization
by: Dutta, Subrat Kishore, et al.
Published: (2025)
by: Dutta, Subrat Kishore, et al.
Published: (2025)
OUSAC: Optimized Guidance Scheduling with Adaptive Caching for DiT Acceleration
by: Sun, Ruitong, et al.
Published: (2025)
by: Sun, Ruitong, et al.
Published: (2025)
Guidance Contrastive Token Credit Assignment for Discrete Policy Optimization
by: Li, Shufan, et al.
Published: (2026)
by: Li, Shufan, et al.
Published: (2026)
HandDreamer: Zero-Shot Text to 3D Hand Model Generation using Corrective Hand Shape Guidance
by: Rosh, Green, et al.
Published: (2026)
by: Rosh, Green, et al.
Published: (2026)
EgoM2P: Egocentric Multimodal Multitask Pretraining
by: Li, Gen, et al.
Published: (2025)
by: Li, Gen, et al.
Published: (2025)
Document Image Rectification Bases on Self-Adaptive Multitask Fusion
by: Li, Heng, et al.
Published: (2025)
by: Li, Heng, et al.
Published: (2025)
Multitask Learning in Minimally Invasive Surgical Vision: A Review
by: Alabi, Oluwatosin, et al.
Published: (2024)
by: Alabi, Oluwatosin, et al.
Published: (2024)
Similar Items
-
A Multimodal Dataset for Enhancing Industrial Task Monitoring and Engagement Prediction
by: Mehta, Naval Kishore, et al.
Published: (2025) -
Gaze-Vector Estimation in the Dark with Temporally Encoded Event-driven Neural Networks
by: Banerjee, Abeer, et al.
Published: (2024) -
Dual Guidance Semi-Supervised Action Detection
by: Singh, Ankit, et al.
Published: (2025) -
fine-CLIP: Enhancing Zero-Shot Fine-Grained Surgical Action Recognition with Vision-Language Models
by: Sharma, Saurav, et al.
Published: (2025) -
Uncertainty-Guided Appearance-Motion Association Network for Out-of-Distribution Action Detection
by: Fang, Xiang, et al.
Published: (2024)