Privacy-Aware Camera 2.0 Technical Report
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Song, Huan, Tian, Shuyu, Long, Ting, Liu, Jiang, Yuan, Cheng, Jia, Zhenyu, Shao, Jiawei, Li, Xuelong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Ruyi2.5 Technical Report
von: Song, Huan, et al.
Veröffentlicht: (2026)
von: Song, Huan, et al.
Veröffentlicht: (2026)
Ruyi2 Technical Report
von: Song, Huan, et al.
Veröffentlicht: (2026)
von: Song, Huan, et al.
Veröffentlicht: (2026)
VARCO-VISION-2.0 Technical Report
von: Cha, Young-rok, et al.
Veröffentlicht: (2025)
von: Cha, Young-rok, et al.
Veröffentlicht: (2025)
Enhancing Neural Video Compression of Static Scenes with Positive-Incentive Noise
von: Yuan, Cheng, et al.
Veröffentlicht: (2026)
von: Yuan, Cheng, et al.
Veröffentlicht: (2026)
Theoretical Foundations of Scaling Law in Familial Models
von: Song, Huan, et al.
Veröffentlicht: (2025)
von: Song, Huan, et al.
Veröffentlicht: (2025)
PhysBrain 1.0 Technical Report
von: Lian, Shijie, et al.
Veröffentlicht: (2026)
von: Lian, Shijie, et al.
Veröffentlicht: (2026)
Mano Technical Report
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
KlingAvatar 2.0 Technical Report
von: Kling Team, et al.
Veröffentlicht: (2025)
von: Kling Team, et al.
Veröffentlicht: (2025)
Qwen2.5-VL Technical Report
von: Bai, Shuai, et al.
Veröffentlicht: (2025)
von: Bai, Shuai, et al.
Veröffentlicht: (2025)
Single-Pixel Vision-Language Model for Intrinsic Privacy-Preserving Behavioral Intelligence
von: An, Hongjun, et al.
Veröffentlicht: (2026)
von: An, Hongjun, et al.
Veröffentlicht: (2026)
Skywork-R1V3 Technical Report
von: Shen, Wei, et al.
Veröffentlicht: (2025)
von: Shen, Wei, et al.
Veröffentlicht: (2025)
Arctic-Extract Technical Report
von: Chiliński, Mateusz, et al.
Veröffentlicht: (2025)
von: Chiliński, Mateusz, et al.
Veröffentlicht: (2025)
Baichuan-Omni Technical Report
von: Li, Yadong, et al.
Veröffentlicht: (2024)
von: Li, Yadong, et al.
Veröffentlicht: (2024)
VLMQ: Token Saliency-Driven Post-Training Quantization for Vision-language Models
von: Xue, Yufei, et al.
Veröffentlicht: (2025)
von: Xue, Yufei, et al.
Veröffentlicht: (2025)
Falcon2-11B Technical Report
von: Malartic, Quentin, et al.
Veröffentlicht: (2024)
von: Malartic, Quentin, et al.
Veröffentlicht: (2024)
Qwen-Image-2.0 Technical Report
von: Zhao, Bing, et al.
Veröffentlicht: (2026)
von: Zhao, Bing, et al.
Veröffentlicht: (2026)
Debiasing Multimodal Large Language Models via Noise-Aware Preference Optimization
von: Zhang, Zefeng, et al.
Veröffentlicht: (2025)
von: Zhang, Zefeng, et al.
Veröffentlicht: (2025)
SliceWorld: A Predictive and Controllable World-State Model for CT Report Generation
von: Tian, Yuanhe, et al.
Veröffentlicht: (2026)
von: Tian, Yuanhe, et al.
Veröffentlicht: (2026)
Watch Out Your Album! On the Inadvertent Privacy Memorization in Multi-Modal Large Language Models
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
von: Ju, Tianjie, et al.
Veröffentlicht: (2025)
Docling Technical Report
von: Auer, Christoph, et al.
Veröffentlicht: (2024)
von: Auer, Christoph, et al.
Veröffentlicht: (2024)
UI-Venus-1.5 Technical Report
von: Venus Team, et al.
Veröffentlicht: (2026)
von: Venus Team, et al.
Veröffentlicht: (2026)
MiMo-Embodied: X-Embodied Foundation Model Technical Report
von: Hao, Xiaoshuai, et al.
Veröffentlicht: (2025)
von: Hao, Xiaoshuai, et al.
Veröffentlicht: (2025)
Recurrent Visual Feature Extraction and Stereo Attentions for CT Report Generation
von: Tian, Yuanhe, et al.
Veröffentlicht: (2025)
von: Tian, Yuanhe, et al.
Veröffentlicht: (2025)
A Systematic Review of Deep Learning-based Research on Radiology Report Generation
von: Liu, Chang, et al.
Veröffentlicht: (2023)
von: Liu, Chang, et al.
Veröffentlicht: (2023)
CameraBench: Benchmarking Visual Reasoning in MLLMs via Photography
von: Fang, I-Sheng, et al.
Veröffentlicht: (2025)
von: Fang, I-Sheng, et al.
Veröffentlicht: (2025)
MedGemma Technical Report
von: Sellergren, Andrew, et al.
Veröffentlicht: (2025)
von: Sellergren, Andrew, et al.
Veröffentlicht: (2025)
Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model
von: Ma, Guoqing, et al.
Veröffentlicht: (2025)
von: Ma, Guoqing, et al.
Veröffentlicht: (2025)
Phoenix-VL 1.5 Medium Technical Report
von: Phoenix, Team, et al.
Veröffentlicht: (2026)
von: Phoenix, Team, et al.
Veröffentlicht: (2026)
Qwen-Image-VAE-2.0 Technical Report
von: Zhang, Zekai, et al.
Veröffentlicht: (2026)
von: Zhang, Zekai, et al.
Veröffentlicht: (2026)
Dynamic Token Reweighting for Robust Vision-Language Models
von: Jiang, Tanqiu, et al.
Veröffentlicht: (2025)
von: Jiang, Tanqiu, et al.
Veröffentlicht: (2025)
Multimodal Structured Generation: CVPR's 2nd MMFM Challenge Technical Report
von: Cesista, Franz Louis
Veröffentlicht: (2024)
von: Cesista, Franz Louis
Veröffentlicht: (2024)
Step-Video-TI2V Technical Report: A State-of-the-Art Text-Driven Image-to-Video Generation Model
von: Huang, Haoyang, et al.
Veröffentlicht: (2025)
von: Huang, Haoyang, et al.
Veröffentlicht: (2025)
VLN-NF: Feasibility-Aware Vision-and-Language Navigation with False-Premise Instructions
von: Su, Hung-Ting, et al.
Veröffentlicht: (2026)
von: Su, Hung-Ting, et al.
Veröffentlicht: (2026)
Octopus v3: Technical Report for On-device Sub-billion Multimodal AI Agent
von: Chen, Wei, et al.
Veröffentlicht: (2024)
von: Chen, Wei, et al.
Veröffentlicht: (2024)
Qwen2.5-Omni Technical Report
von: Xu, Jin, et al.
Veröffentlicht: (2025)
von: Xu, Jin, et al.
Veröffentlicht: (2025)
ICON: Improving Inter-Report Consistency in Radiology Report Generation via Lesion-aware Mixup Augmentation
von: Hou, Wenjun, et al.
Veröffentlicht: (2024)
von: Hou, Wenjun, et al.
Veröffentlicht: (2024)
Little Data, Big Impact: Privacy-Aware Visual Language Models via Minimal Tuning
von: Samson, Laurens, et al.
Veröffentlicht: (2024)
von: Samson, Laurens, et al.
Veröffentlicht: (2024)
Pegasus-v1 Technical Report
von: Jung, Raehyuk, et al.
Veröffentlicht: (2024)
von: Jung, Raehyuk, et al.
Veröffentlicht: (2024)
Technical Report: Quantifying and Analyzing the Generalization Power of a DNN
von: He, Yuxuan, et al.
Veröffentlicht: (2025)
von: He, Yuxuan, et al.
Veröffentlicht: (2025)
Multimodal Reasoning for Science: Technical Report and 1st Place Solution to the ICML 2025 SeePhys Challenge
von: Liang, Hao, et al.
Veröffentlicht: (2025)
von: Liang, Hao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Ruyi2.5 Technical Report
von: Song, Huan, et al.
Veröffentlicht: (2026) -
Ruyi2 Technical Report
von: Song, Huan, et al.
Veröffentlicht: (2026) -
VARCO-VISION-2.0 Technical Report
von: Cha, Young-rok, et al.
Veröffentlicht: (2025) -
Enhancing Neural Video Compression of Static Scenes with Positive-Incentive Noise
von: Yuan, Cheng, et al.
Veröffentlicht: (2026) -
Theoretical Foundations of Scaling Law in Familial Models
von: Song, Huan, et al.
Veröffentlicht: (2025)