A Tactical Behaviour Recognition Framework Based on Causal Multimodal Reasoning: A Study on Covert Audio-Video Analysis Combining GAN Structure Enhancement and Phonetic Accent Modelling
Fuente:
arXiv
Saved in:
| Main Author: | Meng, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Classifier Calibration at Scale: An Empirical Study of Model-Agnostic Post-Hoc Methods
by: Manokhin, Valery, et al.
Published: (2026)
by: Manokhin, Valery, et al.
Published: (2026)
Self-Attention And Beyond the Infinite: Towards Linear Transformers with Infinite Self-Attention
by: Roffo, Giorgio, et al.
Published: (2026)
by: Roffo, Giorgio, et al.
Published: (2026)
TerraSeg: Self-Supervised Ground Segmentation for Any LiDAR
by: Lentsch, Ted, et al.
Published: (2026)
by: Lentsch, Ted, et al.
Published: (2026)
UNION: Unsupervised 3D Object Detection using Object Appearance-based Pseudo-Classes
by: Lentsch, Ted, et al.
Published: (2024)
by: Lentsch, Ted, et al.
Published: (2024)
See What You Need: Query-Aware Visual Intelligence through Reasoning-Perception Loops
by: Dong, Zixuan, et al.
Published: (2025)
by: Dong, Zixuan, et al.
Published: (2025)
GeoJEPA: Towards Eliminating Augmentation- and Sampling Bias in Multimodal Geospatial Learning
by: Lundqvist, Theodor, et al.
Published: (2025)
by: Lundqvist, Theodor, et al.
Published: (2025)
DeepShade: Enable Shade Simulation by Text-conditioned Image Generation
by: Da, Longchao, et al.
Published: (2025)
by: Da, Longchao, et al.
Published: (2025)
Dense Video Understanding with Gated Residual Tokenization
by: Zhang, Haichao, et al.
Published: (2025)
by: Zhang, Haichao, et al.
Published: (2025)
Active Negative Loss: A Robust Framework for Learning with Noisy Labels
by: Ye, Xichen, et al.
Published: (2024)
by: Ye, Xichen, et al.
Published: (2024)
Mitigating Catastrophic Forgetting in Streaming Generative and Predictive Learning via Stateful Replay
by: Du, Wenzhang
Published: (2025)
by: Du, Wenzhang
Published: (2025)
Enhancing Diversity in Multi-objective Feature Selection
by: Miyandoab, Sevil Zanjani, et al.
Published: (2024)
by: Miyandoab, Sevil Zanjani, et al.
Published: (2024)
Adaptive Machine Learning for Resource-Constrained Environments
by: Ordóñez, Sebastián A. Cajas, et al.
Published: (2025)
by: Ordóñez, Sebastián A. Cajas, et al.
Published: (2025)
Akasha 2: Hamiltonian State Space Duality and Visual-Language Joint Embedding Predictive Architectur
by: Meziani, Yani
Published: (2026)
by: Meziani, Yani
Published: (2026)
LightPFP: A Lightweight Route to Ab Initio Accuracy at Scale
by: Li, Wenwen, et al.
Published: (2025)
by: Li, Wenwen, et al.
Published: (2025)
AI-Powered Augmented Reality for Satellite Assembly, Integration and Test
by: Patricio, Alvaro, et al.
Published: (2024)
by: Patricio, Alvaro, et al.
Published: (2024)
Measuring Similarity in Causal Graphs: A Framework for Semantic and Structural Analysis
by: Liu, Ning-Yuan Georgia, et al.
Published: (2025)
by: Liu, Ning-Yuan Georgia, et al.
Published: (2025)
IAUNet: Instance-Aware U-Net
by: Prytula, Yaroslav, et al.
Published: (2025)
by: Prytula, Yaroslav, et al.
Published: (2025)
Vis-CoT: A Human-in-the-Loop Framework for Interactive Visualization and Intervention in LLM Chain-of-Thought Reasoning
by: Pather, Kaviraj, et al.
Published: (2025)
by: Pather, Kaviraj, et al.
Published: (2025)
Visualizing the Evolution of Twitter (X.com) Conversations: A Comprehensive Methodology Applied to AI Training Discussions on ChatGPT
by: Jess, Nicole, et al.
Published: (2024)
by: Jess, Nicole, et al.
Published: (2024)
Network Analysis of the Egyptian Reddit Community
by: Shaawat, Samy, et al.
Published: (2026)
by: Shaawat, Samy, et al.
Published: (2026)
Time Aggregation Features for XGBoost Models
by: Pinchuk, Mykola
Published: (2026)
by: Pinchuk, Mykola
Published: (2026)
Semantic-Constrained Federated Aggregation: Convergence Theory and Privacy-Utility Bounds for Knowledge-Enhanced Distributed Learning
by: Arafat, Jahidul
Published: (2025)
by: Arafat, Jahidul
Published: (2025)
Enhanced Single-Cell RNA-seq Embedding through Gene Expression and Data-Driven Gene-Gene Interaction Integration
by: Goudarzi, Hojjat Torabi, et al.
Published: (2025)
by: Goudarzi, Hojjat Torabi, et al.
Published: (2025)
Modeling and Visualization Reasoning for Stakeholders in Education and Industry Integration Systems: Research on Structured Synthetic Dialogue Data Generation Based on NIST Standards
by: Meng, Wei
Published: (2025)
by: Meng, Wei
Published: (2025)
LinkedOut: Linking World Knowledge Representation Out of Video LLM for Next-Generation Video Recommendation
by: Zhang, Haichao, et al.
Published: (2025)
by: Zhang, Haichao, et al.
Published: (2025)
VideoMind: An Omni-Modal Video Dataset with Intent Grounding for Deep-Cognitive Video Understanding
by: Yang, Baoyao, et al.
Published: (2025)
by: Yang, Baoyao, et al.
Published: (2025)
Deep Learning-Based Multi-Object Tracking: A Comprehensive Survey from Foundations to State-of-the-Art
by: Adžemović, Momir
Published: (2025)
by: Adžemović, Momir
Published: (2025)
Uniqueness ratio as a predictor of a privacy leakage
by: AlKhashti, Danah A. AlSalem
Published: (2025)
by: AlKhashti, Danah A. AlSalem
Published: (2025)
A deep learning approach to track eye movements based on events
by: Seth, Chirag, et al.
Published: (2025)
by: Seth, Chirag, et al.
Published: (2025)
iLTM: Integrated Large Tabular Model
by: Bonet, David, et al.
Published: (2025)
by: Bonet, David, et al.
Published: (2025)
AVATAAR: Agentic Video Answering via Temporal Adaptive Alignment and Reasoning
by: Patel, Urjitkumar, et al.
Published: (2025)
by: Patel, Urjitkumar, et al.
Published: (2025)
Symphonym: Universal Phonetic Embeddings for Cross-Script Name Matching
by: Gadd, Stephen
Published: (2026)
by: Gadd, Stephen
Published: (2026)
WSCIF: A Weakly-Supervised Color Intelligence Framework for Tactical Anomaly Detection in Surveillance Keyframes
by: Meng, Wei
Published: (2025)
by: Meng, Wei
Published: (2025)
Pixel-Wise Multimodal Contrastive Learning for Remote Sensing Images
by: Stival, Leandro, et al.
Published: (2026)
by: Stival, Leandro, et al.
Published: (2026)
Toward precision soil health: A regional framework for site-specific management across Missouri
by: Shah, Dipal, et al.
Published: (2025)
by: Shah, Dipal, et al.
Published: (2025)
Task Memory Engine (TME): Enhancing State Awareness for Multi-Step LLM Agent Tasks
by: Ye, Ye
Published: (2025)
by: Ye, Ye
Published: (2025)
DRIFT open dataset: A drone-derived intelligence for traffic analysis in urban environment
by: Lee, Hyejin, et al.
Published: (2025)
by: Lee, Hyejin, et al.
Published: (2025)
MATEX: Multi-scale Attention and Text-guided Explainability of Medical Vision-Language Models
by: Imran, Muhammad, et al.
Published: (2026)
by: Imran, Muhammad, et al.
Published: (2026)
Adaptive Negative Scheduling for Graph Contrastive Learning
by: Ali, Adnan, et al.
Published: (2026)
by: Ali, Adnan, et al.
Published: (2026)
The Impact of Data Characteristics on GNN Evaluation for Detecting Fake News
by: Karn, Isha, et al.
Published: (2025)
by: Karn, Isha, et al.
Published: (2025)
Similar Items
-
Classifier Calibration at Scale: An Empirical Study of Model-Agnostic Post-Hoc Methods
by: Manokhin, Valery, et al.
Published: (2026) -
Self-Attention And Beyond the Infinite: Towards Linear Transformers with Infinite Self-Attention
by: Roffo, Giorgio, et al.
Published: (2026) -
TerraSeg: Self-Supervised Ground Segmentation for Any LiDAR
by: Lentsch, Ted, et al.
Published: (2026) -
UNION: Unsupervised 3D Object Detection using Object Appearance-based Pseudo-Classes
by: Lentsch, Ted, et al.
Published: (2024) -
See What You Need: Query-Aware Visual Intelligence through Reasoning-Perception Loops
by: Dong, Zixuan, et al.
Published: (2025)