Differentiable JPEG: The Devil is in the Details
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Reich, Christoph, Debnath, Biplob, Patel, Deep, Chakradhar, Srimat |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Deep Video Codec Control for Vision Models
von: Reich, Christoph, et al.
Veröffentlicht: (2023)
von: Reich, Christoph, et al.
Veröffentlicht: (2023)
A Perspective on Deep Vision Performance with Standard Image and Video Codecs
von: Reich, Christoph, et al.
Veröffentlicht: (2024)
von: Reich, Christoph, et al.
Veröffentlicht: (2024)
TrafficLens: Multi-Camera Traffic Video Analysis Using LLMs
von: Arefeen, Md Adnan, et al.
Veröffentlicht: (2025)
von: Arefeen, Md Adnan, et al.
Veröffentlicht: (2025)
iRAG: Advancing RAG for Videos with an Incremental Approach
von: Arefeen, Md Adnan, et al.
Veröffentlicht: (2024)
von: Arefeen, Md Adnan, et al.
Veröffentlicht: (2024)
Open-SAT: LLM-Guided Query Embedding Refinement for Open-Vocabulary Object Retrieval in Satellite Imagery
von: Arefeen, Md Adnan, et al.
Veröffentlicht: (2026)
von: Arefeen, Md Adnan, et al.
Veröffentlicht: (2026)
Visual Alignment of Medical Vision-Language Models for Grounded Radiology Report Generation
von: Bose, Sarosij, et al.
Veröffentlicht: (2025)
von: Bose, Sarosij, et al.
Veröffentlicht: (2025)
JPEG AI Image Compression Visual Artifacts: Detection Methods and Dataset
von: Tsereh, Daria, et al.
Veröffentlicht: (2024)
von: Tsereh, Daria, et al.
Veröffentlicht: (2024)
Multi-Scale and Detail-Enhanced Segment Anything Model for Salient Object Detection
von: Gao, Shixuan, et al.
Veröffentlicht: (2024)
von: Gao, Shixuan, et al.
Veröffentlicht: (2024)
MSCT: Differential Cross-Modal Attention for Deepfake Detection
von: Wei, Fangda, et al.
Veröffentlicht: (2026)
von: Wei, Fangda, et al.
Veröffentlicht: (2026)
DetailVerifyBench: A Benchmark for Dense Hallucination Localization in Long Image Captions
von: Wang, Xinran, et al.
Veröffentlicht: (2026)
von: Wang, Xinran, et al.
Veröffentlicht: (2026)
Accelerated Event-Based Feature Detection and Compression for Surveillance Video Systems
von: Freeman, Andrew C., et al.
Veröffentlicht: (2023)
von: Freeman, Andrew C., et al.
Veröffentlicht: (2023)
Look, Compare and Draw: Differential Query Transformer for Automatic Oil Painting
von: Liu, Lingyu, et al.
Veröffentlicht: (2026)
von: Liu, Lingyu, et al.
Veröffentlicht: (2026)
StreamingRAG: Real-time Contextual Retrieval and Generation Framework
von: Sankaradas, Murugan, et al.
Veröffentlicht: (2025)
von: Sankaradas, Murugan, et al.
Veröffentlicht: (2025)
Mesquite MoCap: Democratizing Real-Time Motion Capture with Affordable, Bodyworn IoT Sensors and WebXR SLAM
von: Vanani, Poojan, et al.
Veröffentlicht: (2025)
von: Vanani, Poojan, et al.
Veröffentlicht: (2025)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
von: Zhang, Zhenxing, et al.
Veröffentlicht: (2024)
von: Zhang, Zhenxing, et al.
Veröffentlicht: (2024)
Omni-Captioner: Data Pipeline, Models, and Benchmark for Omni Detailed Perception
von: Ma, Ziyang, et al.
Veröffentlicht: (2025)
von: Ma, Ziyang, et al.
Veröffentlicht: (2025)
Tackle CSM in JPEG Steganalysis with Data Adaptation
von: Abecidan, Rony, et al.
Veröffentlicht: (2026)
von: Abecidan, Rony, et al.
Veröffentlicht: (2026)
Reference-Guided Identity Preserving Face Restoration
von: Zhou, Mo, et al.
Veröffentlicht: (2025)
von: Zhou, Mo, et al.
Veröffentlicht: (2025)
Improving Accuracy and Generalization for Efficient Visual Tracking
von: Zaveri, Ram, et al.
Veröffentlicht: (2024)
von: Zaveri, Ram, et al.
Veröffentlicht: (2024)
SequencePAR: Understanding Pedestrian Attributes via A Sequence Generation Paradigm
von: Jin, Jiandong, et al.
Veröffentlicht: (2023)
von: Jin, Jiandong, et al.
Veröffentlicht: (2023)
CASR: Refining Action Segmentation via Marginalizing Frame-levle Causal Relationships
von: Du, Keqing, et al.
Veröffentlicht: (2023)
von: Du, Keqing, et al.
Veröffentlicht: (2023)
Study of Subjective and Objective Quality Assessment of Mobile Cloud Gaming Videos
von: Saha, Avinab, et al.
Veröffentlicht: (2023)
von: Saha, Avinab, et al.
Veröffentlicht: (2023)
Unraveling Instance Associations: A Closer Look for Audio-Visual Segmentation
von: Chen, Yuanhong, et al.
Veröffentlicht: (2023)
von: Chen, Yuanhong, et al.
Veröffentlicht: (2023)
Bridging the Gap: Sketch-Aware Interpolation Network for High-Quality Animation Sketch Inbetweening
von: Shen, Jiaming, et al.
Veröffentlicht: (2023)
von: Shen, Jiaming, et al.
Veröffentlicht: (2023)
Joint Explicit and Implicit Cross-Modal Interaction Network for Anterior Chamber Inflammation Diagnosis
von: Shao, Qian, et al.
Veröffentlicht: (2023)
von: Shao, Qian, et al.
Veröffentlicht: (2023)
VIoTGPT: Learning to Schedule Vision Tools in LLMs towards Intelligent Video Internet of Things
von: Zhong, Yaoyao, et al.
Veröffentlicht: (2023)
von: Zhong, Yaoyao, et al.
Veröffentlicht: (2023)
CLIP Brings Better Features to Visual Aesthetics Learners
von: Xu, Liwu, et al.
Veröffentlicht: (2023)
von: Xu, Liwu, et al.
Veröffentlicht: (2023)
QGFace: Quality-Guided Joint Training For Mixed-Quality Face Recognition
von: Song, Youzhe, et al.
Veröffentlicht: (2023)
von: Song, Youzhe, et al.
Veröffentlicht: (2023)
Noisy-Correspondence Learning for Text-to-Image Person Re-identification
von: Qin, Yang, et al.
Veröffentlicht: (2023)
von: Qin, Yang, et al.
Veröffentlicht: (2023)
TALDS-Net: Task-Aware Adaptive Local Descriptors Selection for Few-shot Image Classification
von: Qiao, Qian, et al.
Veröffentlicht: (2023)
von: Qiao, Qian, et al.
Veröffentlicht: (2023)
Towards Natural Language-Guided Drones: GeoText-1652 Benchmark with Spatial Relation Matching
von: Chu, Meng, et al.
Veröffentlicht: (2023)
von: Chu, Meng, et al.
Veröffentlicht: (2023)
Using Saliency and Cropping to Improve Video Memorability
von: Mudgal, Vaibhav, et al.
Veröffentlicht: (2023)
von: Mudgal, Vaibhav, et al.
Veröffentlicht: (2023)
Vision-Language Models Learn Super Images for Efficient Partially Relevant Video Retrieval
von: Nishimura, Taichi, et al.
Veröffentlicht: (2023)
von: Nishimura, Taichi, et al.
Veröffentlicht: (2023)
GPT-4V with Emotion: A Zero-shot Benchmark for Generalized Emotion Recognition
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
von: Lian, Zheng, et al.
Veröffentlicht: (2023)
FlowVid: Taming Imperfect Optical Flows for Consistent Video-to-Video Synthesis
von: Liang, Feng, et al.
Veröffentlicht: (2023)
von: Liang, Feng, et al.
Veröffentlicht: (2023)
Modularized Zero-shot VQA with Pre-trained Models
von: Cao, Rui, et al.
Veröffentlicht: (2023)
von: Cao, Rui, et al.
Veröffentlicht: (2023)
PathAsst: A Generative Foundation AI Assistant Towards Artificial General Intelligence of Pathology
von: Sun, Yuxuan, et al.
Veröffentlicht: (2023)
von: Sun, Yuxuan, et al.
Veröffentlicht: (2023)
Test-Time Adaptation with CLIP Reward for Zero-Shot Generalization in Vision-Language Models
von: Zhao, Shuai, et al.
Veröffentlicht: (2023)
von: Zhao, Shuai, et al.
Veröffentlicht: (2023)
AdaMesh: Personalized Facial Expressions and Head Poses for Adaptive Speech-Driven 3D Facial Animation
von: Chen, Liyang, et al.
Veröffentlicht: (2023)
von: Chen, Liyang, et al.
Veröffentlicht: (2023)
UniPT: Universal Parallel Tuning for Transfer Learning with Efficient Parameter and Memory
von: Diao, Haiwen, et al.
Veröffentlicht: (2023)
von: Diao, Haiwen, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Deep Video Codec Control for Vision Models
von: Reich, Christoph, et al.
Veröffentlicht: (2023) -
A Perspective on Deep Vision Performance with Standard Image and Video Codecs
von: Reich, Christoph, et al.
Veröffentlicht: (2024) -
TrafficLens: Multi-Camera Traffic Video Analysis Using LLMs
von: Arefeen, Md Adnan, et al.
Veröffentlicht: (2025) -
iRAG: Advancing RAG for Videos with an Incremental Approach
von: Arefeen, Md Adnan, et al.
Veröffentlicht: (2024) -
Open-SAT: LLM-Guided Query Embedding Refinement for Open-Vocabulary Object Retrieval in Satellite Imagery
von: Arefeen, Md Adnan, et al.
Veröffentlicht: (2026)