MediaClaw: Multimodal Intelligent-Agent Platform Technical Report
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Shaoan, Gao, Huanlin, Hui, Qiang, Lu, Ting, Guo, Xueqiang, Li, Yantao, Su, Xinpei, Shi, Fuyuan, Tan, Chao, Zhao, Fang, Wang, Kai, Lian, Shiguo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LeMiCa: Lexicographic Minimax Path Caching for Efficient Diffusion-Based Video Generation
by: Gao, Huanlin, et al.
Published: (2025)
by: Gao, Huanlin, et al.
Published: (2025)
MeanCache: From Instantaneous to Average Velocity for Accelerating Flow Matching Inference
by: Gao, Huanlin, et al.
Published: (2026)
by: Gao, Huanlin, et al.
Published: (2026)
HiMo-CLIP: Modeling Semantic Hierarchy and Monotonicity in Vision-Language Alignment
by: Wu, Ruijia, et al.
Published: (2025)
by: Wu, Ruijia, et al.
Published: (2025)
PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment
by: Li, Yantao, et al.
Published: (2026)
by: Li, Yantao, et al.
Published: (2026)
From Assistant to Double Agent: Formalizing and Benchmarking Attacks on OpenClaw for Personalized Local AI Agent
by: Wang, Yuhang, et al.
Published: (2026)
by: Wang, Yuhang, et al.
Published: (2026)
X-OmniClaw Technical Report: A Unified Mobile Agent for Multimodal Understanding and Interaction
by: Ren, Xiaoming, et al.
Published: (2026)
by: Ren, Xiaoming, et al.
Published: (2026)
A Systematic Security Evaluation of OpenClaw and Its Variants
by: Wang, Yuhang, et al.
Published: (2026)
by: Wang, Yuhang, et al.
Published: (2026)
MITS: A Large-Scale Multimodal Benchmark Dataset for Intelligent Traffic Surveillance
by: Zhao, Kaikai, et al.
Published: (2025)
by: Zhao, Kaikai, et al.
Published: (2025)
StreamingClaw Technical Report
by: Chen, Jiawei, et al.
Published: (2026)
by: Chen, Jiawei, et al.
Published: (2026)
NeuroClaw Technical Report
by: Wang, Cheng, et al.
Published: (2026)
by: Wang, Cheng, et al.
Published: (2026)
KAConvNet: Kolmogorov-Arnold Convolutional Networks for Vision Recognition
by: Liu, Zhaoxiang, et al.
Published: (2026)
by: Liu, Zhaoxiang, et al.
Published: (2026)
ClawGym: A Scalable Framework for Building Effective Claw Agents
by: Bai, Fei, et al.
Published: (2026)
by: Bai, Fei, et al.
Published: (2026)
Fatigue Life Analysis of Cranes Based on Load Spectrum Prediction and Fracture Mechanics
by: Mantang Hu, et al.
Published: (2025)
by: Mantang Hu, et al.
Published: (2025)
A S‐shaped feedrate based NURBS interpolator for synchronization of robot tool tip trajectory and attitude
by: Guanghui Liu, et al.
Published: (2025)
by: Guanghui Liu, et al.
Published: (2025)
ClawSafety: "Safe" LLMs, Unsafe Agents
by: Wei, Bowen, et al.
Published: (2026)
by: Wei, Bowen, et al.
Published: (2026)
A Multimodal Benchmark Dataset and Model for Crop Disease Diagnosis
by: Liu, Xiang, et al.
Published: (2025)
by: Liu, Xiang, et al.
Published: (2025)
ClawEnvKit: Automatic Environment Generation for Claw-Like Agents
by: Li, Xirui, et al.
Published: (2026)
by: Li, Xirui, et al.
Published: (2026)
ClawTrap: A MITM-Based Red-Teaming Framework for Real-World OpenClaw Security Evaluation
by: Zhao, Haochen, et al.
Published: (2026)
by: Zhao, Haochen, et al.
Published: (2026)
TIR-Agent: Training an Explorative and Efficient Agent for Image Restoration
by: Zhang, Yisheng, et al.
Published: (2026)
by: Zhang, Yisheng, et al.
Published: (2026)
Firm theories in neoclassical institutional economics
by: Shaoan Huang
Published: (2024)
by: Shaoan Huang
Published: (2024)
CHiSafetyBench: A Chinese Hierarchical Safety Benchmark for Large Language Models
by: Zhang, Wenjing, et al.
Published: (2024)
by: Zhang, Wenjing, et al.
Published: (2024)
Claw AI Lab: An Autonomous Multi-Agent Research Team
by: Wu, Fan, et al.
Published: (2026)
by: Wu, Fan, et al.
Published: (2026)
ClawTrace: Cost-Aware Tracing for LLM Agent Skill Distillation
by: Yuan, Boqin, et al.
Published: (2026)
by: Yuan, Boqin, et al.
Published: (2026)
ALIVE: Awakening LLM Reasoning via Adversarial Learning and Instructive Verbal Evaluation
by: Duan, Yiwen, et al.
Published: (2026)
by: Duan, Yiwen, et al.
Published: (2026)
Patch-wise Auto-Encoder for Visual Anomaly Detection
by: Cui, Yajie, et al.
Published: (2023)
by: Cui, Yajie, et al.
Published: (2023)
Data-Driven Deepfake Image Detection Method -- The 2024 Global Deepfake Image Detection Challenge
by: Zhu, Xiaoya, et al.
Published: (2025)
by: Zhu, Xiaoya, et al.
Published: (2025)
TP3M: Transformer-based Pseudo 3D Image Matching with Reference Image
by: Han, Liming, et al.
Published: (2024)
by: Han, Liming, et al.
Published: (2024)
Multi-Modal Artificial Intelligence of Embryo Grading and Pregnancy Prediction in Assisted Reproductive Technology: A Review
by: Ouyang, Xueqiang, et al.
Published: (2025)
by: Ouyang, Xueqiang, et al.
Published: (2025)
SentiMM: A Multimodal Multi-Agent Framework for Sentiment Analysis in Social Media
by: Xu, Xilai, et al.
Published: (2025)
by: Xu, Xilai, et al.
Published: (2025)
OPTIMA: Optimized Policy for Intelligent Multi-Agent Systems Enables Coordination-Aware Autonomous Vehicles
by: Du, Rui, et al.
Published: (2024)
by: Du, Rui, et al.
Published: (2024)
AnomalyClaw: A Universal Visual Anomaly Detection Agent via Tool-Grounded Refutation
by: Jiang, Xi, et al.
Published: (2026)
by: Jiang, Xi, et al.
Published: (2026)
Clawdrain: Exploiting Tool-Calling Chains for Stealthy Token Exhaustion in OpenClaw Agents
by: Dong, Ben, et al.
Published: (2026)
by: Dong, Ben, et al.
Published: (2026)
Number of independent transversals in multipartite graphs
by: Tang, Yantao, et al.
Published: (2025)
by: Tang, Yantao, et al.
Published: (2025)
Wound Bed Temperature as a Biomarker: Clinical Utility, Limitations and Future Directions
by: Di Xiao, et al.
Published: (2025)
by: Di Xiao, et al.
Published: (2025)
A Large Vision-Language Model based Environment Perception System for Visually Impaired People
by: Chen, Zezhou, et al.
Published: (2025)
by: Chen, Zezhou, et al.
Published: (2025)
Piculet: Specialized Models-Guided Hallucination Decrease for MultiModal Large Language Models
by: Wang, Kohou, et al.
Published: (2024)
by: Wang, Kohou, et al.
Published: (2024)
ClawKeeper: Comprehensive Safety Protection for OpenClaw Agents Through Skills, Plugins, and Watchers
by: Liu, Songyang, et al.
Published: (2026)
by: Liu, Songyang, et al.
Published: (2026)
ClawMark: A Living-World Benchmark for Multi-Turn, Multi-Day, Multimodal Coworker Agents
by: Meng, Fanqing, et al.
Published: (2026)
by: Meng, Fanqing, et al.
Published: (2026)
Do Multimodal Agents Really Benefit from Tool Use? A Systematic Study of Capability Gains
by: Guo, Garvin, et al.
Published: (2026)
by: Guo, Garvin, et al.
Published: (2026)
AR2: Attention-Guided Repair for the Robustness of CNNs Against Common Corruptions
by: Zhang, Fuyuan, et al.
Published: (2025)
by: Zhang, Fuyuan, et al.
Published: (2025)
Similar Items
-
LeMiCa: Lexicographic Minimax Path Caching for Efficient Diffusion-Based Video Generation
by: Gao, Huanlin, et al.
Published: (2025) -
MeanCache: From Instantaneous to Average Velocity for Accelerating Flow Matching Inference
by: Gao, Huanlin, et al.
Published: (2026) -
HiMo-CLIP: Modeling Semantic Hierarchy and Monotonicity in Vision-Language Alignment
by: Wu, Ruijia, et al.
Published: (2025) -
PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment
by: Li, Yantao, et al.
Published: (2026) -
From Assistant to Double Agent: Formalizing and Benchmarking Attacks on OpenClaw for Personalized Local AI Agent
by: Wang, Yuhang, et al.
Published: (2026)